7/26/2026, 1:05:14 PM · video-generation

Germany's Black Forest Labs Unveils FLUX 3 Multimodal Model, Claims 77% Win Rate Over Runway Gen-4.5

The Freiburg-based lab's first video-generation model jointly trains on image, video, and audio within a single architecture, extending the same backbone to robotic action prediction.

Background

<cite index="20-2">Black Forest Labs was founded in August 2024 in Freiburg im Breisgau, Germany, by Robin Rombach, Andreas Blattmann, Patrick Esser, and Dominik Lorenz — all principal researchers behind the latent diffusion technology that powered Stable Diffusion.</cite> <cite index="25-4">The company has raised significant capital across multiple rounds, culminating in a $300 million Series B financing at a $3.25 billion valuation in February 2026, backed by investors including Andreessen Horowitz, General Catalyst, and Salesforce Ventures.</cite>

The Announcement

<cite index="7-2,7-3">On July 23, 2026, Black Forest Labs introduced FLUX 3, its new multimodal frontier model that jointly learns from images, video, and audio within a unified architecture, and can also be extended to predict actions.</cite> <cite index="4-6">This marks the company's first-ever video generation model — until now, Black Forest Labs had shipped image-only models, from the original FLUX.1 through the FLUX.2 line.</cite>

Architecture: Self-Flow

<cite index="12-12,12-13,12-14">FLUX 3 builds on Self-Flow, Black Forest Labs' approach for aligning multimodal generation and understanding within the same underlying architecture. Based on this approach, the company scaled up compute and data resources to train FLUX 3 across video, images, and audio simultaneously. Self-Flow is a training technique that unifies representation learning and generation in one pass, without relying on frozen external encoders such as CLIP or DINO.</cite> <cite index="7-8">Testing showed that generative video generation and action prediction do not require separate foundations, as the same underlying architecture could be extended to action prediction without sacrificing capabilities learned from videos.</cite>

Capabilities

<cite index="6-11">The model's most immediate feature is video generation: FLUX 3 can produce 20-second video clips with synchronized audio, where sound effects, dialogue, and ambient noise align with the visuals.</cite> <cite index="14-3">It supports text-to-video, image-to-video, video-to-video, and keyframe-to-video, with multi-shot sequences chained agentically.</cite> A companion product, FLUX-mimic, extends the same backbone to physical AI: <cite index="3-17">it is being tested with Mimic Robotics on real production tasks at Audi's manufacturing facilities.</cite>

Benchmark Claims

<cite index="10-6,10-7">Black Forest Labs has published several benchmark comparisons, qualified as preliminary, with full methodology to be published later during broader general availability. In early head-to-head preference testing on 10-second, 720p text-to-video clips with audio, the company says FLUX 3 was preferred over Luma Ray 3.2 in 93% of comparisons, Runway Gen-4.5 in 77%, Grok Imagine Video in 69%, Kling v3 Pro in 60%, Happy Horse v1 in 59%, Happy Horse 1.1 in 57%, and both Seedance 2.0 and Google's Gemini Omni Flash in 52%.</cite>

Those figures require context. <cite index="9-5,9-6">Black Forest Labs has not published the evaluation methodology, sample size, rater selection criteria, or the specific prompt set used; the clips used in comparisons were 10-second generations, not the full 20-second maximum.</cite> <cite index="9-7">Independent benchmarks from researchers or publication-grade review sites have not yet been conducted.</cite>

Availability and Pricing

<cite index="4-9">FLUX 3 Video and FLUX 3 Action entered a gated early-access program on launch day — anyone can apply, but Black Forest Labs must approve each applicant.</cite> <cite index="1-8">Image generation is to follow in the coming weeks, and an open-weight FLUX 3 Dev backbone is planned for later.</cite> <cite index="12-17">No pricing has been disclosed for any tier as of this writing.</cite>

Partnerships and Ecosystem

<cite index="7-4">The FLUX family already powers generative features inside leading platforms including Adobe Photoshop, Picsart, and Nous Research's Hermes Agent.</cite> <cite index="14-8">Canva, Burda, Magnific (formerly Freepik), Krea, and Picsart are testing FLUX 3, per Black Forest Labs.</cite> The lab's dual-track strategy — gated commercial access paired with a planned open-weight release — mirrors the approach that made earlier FLUX models among the most-downloaded image generation systems on Hugging Face.

Cross-references

Sources

  1. [1]
    What Is FLUX 3? BFL's Multimodal Omni Model
  2. [2]
    Black Forest Labs launches Flux 3, its new multimodal model for video and audio generation | AlternativeTo
  3. [3]
    FLUX 3 Is Here: Black Forest Labs Unveils a Multimodal Model That Generates Video, Image, and Audio Together | Wan 2.7
  4. [4]
    FLUX 3: Black Forest Labs Goes Multimodal Frontier
  5. [5]
    Black Forest Labs Unveils FLUX 3 AI: Ditches Stills for Video—And Robot Hands - Decrypt
  6. [6]
    Germany's Black Forest Labs Launches Flux 3: Debuts Video Generation and Expands into Robot Manipulation — BigGo Finance
  7. [7]
    Black Forest Labs Unveils FLUX 3, A New Multimodal Frontier Model For Visual Intelligence
  8. [8]
    Black Forest Labs launches FLUX 3 capable of generating images and 20-second video with audio — but in limited release to start | VentureBeat
  9. [9]
    FLUX 3 Launches: Black Forest Labs Enters Video, Audio, and Physical AI in One Model
  10. [10]
    Black Forest Labs' FLUX 3 Promises Images, Video and Audio; Which Features Actually Work Today
  11. [11]
    Black Forest Labs' FLUX 3 Beats Runway Gen-4.5 and Now Runs Factory Robots | AlphaSignal
  12. [12]
    Runway Research | Introducing Runway Gen-4.5
  13. [13]
    How to Generate Videos with FLUX 3 (2026) | Flick
  14. [14]
    FLUX 3 — Real World Multimodal AI Model Showcase
  15. [15]
    FLUX 3 Specs — Model Capabilities, Benchmarks & FAQ
  16. [16]
    German startup Black Forest Labs secures $300M Series B round | Vestbee
  17. [17]
    Black Forest Labs - Learn AI
  18. [18]
    Germany’s Black Forest Labs to raise $300M at $4B valuation — TFN
  19. [19]
    From The Forest to The Stratosphere | Artifical Intelligence
  20. [20]
    German AI startup Black Forest Labs raises $300m Series B | Sifted
  21. [21]
    Used by Adobe, Canva and Meta, Germany's Black Forest Labs lands €258 million to scale its visual-AI platform | EU-Startups
  22. [22]
    Black Forest Labs: Funding, Team & Investors | Startup Intros
  23. [23]
    Robin Rombach Co-founder & CEO Black Forest Labs
  24. [24]
    Black Forest Labs: Europe’s most-hyped — and elusive — startup? | Sifted
  25. [25]
    Black Forest Labs | AI Wiki