FLUX 3 x mimic: The Next Generation of Video-Action Models

FLUX 3 x mimic: The Next Generation of Video-Action Models

We combined our FLUX 3 multimodal foundation model with mimic robotics' expertise to create FLUX-mimic, a new video-action model for real-world automation. By treating actions, video, and audio as views of a single physical reality, we proved that one backbone can master both content creation and robot control. This approach allows robots to learn complex manipulation tasks with minimal data, scaling our world model from the lab to actual production lines at Audi.

If one model does both, it was never really only a content creation model. It is a model of how the world behaves, and content creation is one thing one can do with it.

More from this day

2026-07-24