✳flâneur — a map of the web's best reading
Generative modelling in latent space – Sander Dieleman
sander.ai · 13,056 words · saved by 6 readers
Latent representations for generative models.
Most contemporary generative models of images, sound and video do not operate directly on pixels or waveforms. They consist of two stages : first, a compact, higher-level latent representation is extracted, and then an iterative generative process operates on this representation instead. How does this work, and why is this approach so popular? Generative models that make use of latent representations are everywhere nowadays, so I thought it was high time to dedicate a blog post to them. In what follows, I will talk at length about latents as a plural noun, which is the usual shorthand for late
Explore this link on the map →saved by
related reading
- What are Diffusion Models? | Lil'Loglilianweng.github.io
- Understanding VQ-VAE (DALL-E Explained Pt. 1)mlberkeley.substack.com
- Energy-Based Modelsenergy-based-model.github.io
- Difference between AutoEncoder (AE) and Variational AutoEncoder (VAE) | Towards Data Sciencetowardsdatascience.com
- Generating music in the waveform domain – Sander Dielemansander.ai
- The Principles of Diffusion Modelsarxiv.org
- 2409.02908arxiv.org
- Diffusion is spectral autoregression – Sander Dielemansander.ai
- Large Language Diffusion Modelsarxiv.org
- ⭐️ Diffusion Modelsandrewkchan.dev
- Variational autoencoders.jeremyjordan.me
- francesco215.github.io/autoregressive_diffusion/francesco215.github.io