March 5, 2026

Black Forest Labs' new Self-Flow technique makes training multimodal AI models 2.8x more efficient

low angle photo of beige concrete building under cloudy sky
Anthony Esau / Unsplash

To create coherent images or videos, generative AI diffusion models like Stable Diffusion or FLUX have typically relied on external "teachers"—frozen encoders like CLIP or DINOv2—to provide the semantic understanding they couldn't lear...