FLUX, Open Research, and the Future of Visual AI — Stephen Batifol, Black Forest Labs
May 8, 2026 · 22:32
Black Forest Labs (BFL), the team behind Stable Diffusion and Latent Diffusion, has released a series of open FLUX image models pushing toward visual intelligence. After FLUX.1 and the first open-source editing model FLUX Kontext (7–8 second edits), FLUX.2 achieved state-of-the-art text-to-image and multi-reference editing, and FLUX.2 Klein generates and edits in 300–500 milliseconds for near real-time use. BFL also published Self-Flow, a scalable self-supervised approach that trains multimodal models across images, video, audio, and actions without external encoders, outperforming baselines in all modalities. The episode explains how Self-Flow reduces artifacts (e.g., corrects text rendering and anatomy) and converges 70× faster with representation alignment. BFL’s roadmap includes world models that simulate geometry and interaction, aiming to train agents for robotics and automation.