Every major generative AI model for images, video, and physical AI runs on latent diffusion — invented by Robin Rombach and his co-founders as PhD students in Munich. The insight: compress natural data like images and video into efficient representations, then train a transformer on that. JPEG and MP3 principles, applied to generative intelligence.