# Black Forest Labs releases FLUX 3 multimodal video model

> Black Forest Labs, the German startup behind the popular FLUX image models, has unveiled FLUX 3, and the ambition is notable: a single architecture trained jointly across images, video, and audio, plus robot action prediction. Rather than bolting separate systems together, the model learns these modalities in one place and can generate up to 20 seconds of video with native dialogue, sound effects, and ambient noise produced in the same pass. Beyond text-to-video, it supports image-to-video and video-to-video with character consistency and keyframe control, the kind of features creators need for coherent longer clips. Early access is open now, with an API, private weights, and an open FLUX 3 Dev release promised later in 2026. The interesting part is the convergence. Folding audio and even robot control into one generative model points toward systems that treat pixels, sound, and action as versions of the same prediction problem, and it keeps an independent lab in serious contention against much larger video efforts.

_Section: [Daily AI Updates](https://www.wortins.com/daily-ai) · Source: MarkTechPost · Published Tuesday, August 18, 2026_

## Wortins' read

Black Forest Labs, the German startup behind the popular FLUX image models, has unveiled FLUX 3, and the ambition is notable: a single architecture trained jointly across images, video, and audio, plus robot action prediction. Rather than bolting separate systems together, the model learns these modalities in one place and can generate up to 20 seconds of video with native dialogue, sound effects, and ambient noise produced in the same pass. Beyond text-to-video, it supports image-to-video and video-to-video with character consistency and keyframe control, the kind of features creators need for coherent longer clips. Early access is open now, with an API, private weights, and an open FLUX 3 Dev release promised later in 2026. The interesting part is the convergence. Folding audio and even robot control into one generative model points toward systems that treat pixels, sound, and action as versions of the same prediction problem, and it keeps an independent lab in serious contention against much larger video efforts.

## Source

[Read the full story at MarkTechPost](https://www.marktechpost.com/2026/07/26/black-forest-labs-releases-flux-3-a-multimodal-flow-model-for-image-video-audio-and-robot-action-prediction/)

## Related coverage

- [Nvidia Agrees to Acquire Hugging Face for $13 Billion](https://www.wortins.com/story/nvidia-agrees-to-acquire-hugging-face-for-13-billion-17f12bb8) — [TechCrunch](https://techcrunch.com/2026/08/26/nvidia-closes-in-on-hugging-face-acquisition/)
- [80% of Organizations Report Measurable ROI from AI Agents in Production](https://www.wortins.com/story/80-of-organizations-report-measurable-roi-from-ai-agents-in--0f1c628d) — [Anthropic](https://resources.anthropic.com/2026-state-of-ai-agents)
- [Google DeepMind Releases Gemini Robotics 2 With Whole-Body Control](https://www.wortins.com/story/google-deepmind-releases-gemini-robotics-2-with-whole-body-c-578b43eb) — [Google DeepMind](https://deepmind.google/blog/gemini-robotics-2-brings-whole-body-intelligence-to-robots/)
- [Cohere Launches Command A+ Mixture-of-Experts Model](https://www.wortins.com/story/cohere-launches-command-a-mixture-of-experts-model-5d840970) — [Cohere](https://docs.cohere.com/docs/command-a-plus)
- [DARPA and US Air Force Successfully Fly F-16 Fighter Jet Under Full AI Control](https://www.wortins.com/story/darpa-and-us-air-force-successfully-fly-f-16-fighter-jet-und-b659da37) — [DARPA](https://www.darpa.mil/news/2026/darpa-us-air-force-fly-ai-controlled-f-16)
- [Moonshot AI Releases Kimi K3, World's Largest Open-Source AI Model at 2.8 Trillion Parameters](https://www.wortins.com/story/moonshot-ai-releases-kimi-k3-world-s-largest-open-source-ai--cde823ba) — [Tom's Hardware](https://www.tomshardware.com/tech-industry/artificial-intelligence/moonshot-releases-2-8-trillion-parameter-kimi-k3)

---

_Curated and written by [Wortins](https://www.wortins.com) — The daily AI briefing. Every story links to its original source; the "Wortins read" on each is our own original analysis. [About Wortins & our editorial approach](https://www.wortins.com/about)._
