FLUX 3 x mimic: The Next Generation of Video-Action Models | Black Forest Labs
FLUX 3 is running robots. Built with mimic and deployed at Audi, FLUX-mimic shows content creation and physical AI share one foundation: a model that understands the world.
An early version of FLUX 3, our new multimodal foundation model, is now running on robots. We gave mimic robotics early access to FLUX.3. Their strength in robot learning and deployment, combined with the model's world knowledge and BFL's foundation model expertise, produced FLUX-mimic: the next generation of video-action models. FLUX FLUX 1 and FLUX 2 generate images. FLUX 3 expands into multimodality and generates audio-visual content jointly - and, at the same time, provides the foundation of FLUX-mimic: A video-action model, developed in collaboration with mimic, running robots that…
saved by
related reading
- Generalist - GEN-1: Scaling Embodied Foundation Models to Masterygeneralistai.com
- The First Fully General Computer Action Model | blogsi.inc
- The Model That Dreams the Worldmoe-capital.com
- 1X World Model | From Video to Action: A New Way Robots Learn1x.tech
- A Functional Taxonomy of World Models - Dr. Fei-Fei Lidrfeifei.substack.com
- [2606.02800] Cosmos 3: Omnimodal World Models for Physical AIarxiv.org
- World Models | Rohit Bandarurohitbandaru.github.io
- Causal Video Models Are Data-Efficient Robot Policy Learners | Rhoda AIrhoda.ai
- How Claude Performs on Robotics Tasks \ Anthropicanthropic.com
- Explore | alphaXivalphaxiv.org
- Pretrained to Imagine, Fine-Tuned to Act: The Rise of World-Action Models | NVIDIA Technical Blogdeveloper.nvidia.com
- General Instinct | Any frontier model. Any edge device.general-instinct.com