Black Forest Labs has released FLUX 3, a multimodal foundation model that learns from images, video, and audio in a single architecture. The model can generate 20-second video clips with native audio and, in human preference tests, was preferred over Luma Ray 3.2 in 93% of comparisons. The same backbone also powers a robot policy, FLUX-mimic. Access is being rolled out in stages via early access.
Aug 10, 2026 · 8 sources
Aug 9, 2026 · 2 sources
Aug 8, 2026 · 1 source
Aug 8, 2026 · 7 sources
Story comments
Loading comments…