← Back to Tech Radar
Hacker News tech

Flux 3 X Mimic: The Next Generation of Video-Action Models

Trending on Hacker News: Flux 3 X Mimic: The Next Generation of Video-Action Models (297 points / 47 comments, via bfl.ai)

In one line

FLUX 3 is running robots. Built with mimic and deployed at Audi, FLUX-mimic shows content creation and physical AI share one foundation: a model that understands the world.

Opening excerpt

Back to blog Research Models FLUX 3 x mimic: The Next Generation of Video-Action Models An early version of FLUX 3, our new multimodal foundation model , is now running on robots. We gave mimic robotics early access to FLUX.3. Their strength in robot learning and deployment, combined with the model’s world knowledge and BFL’s foundation model expertise, produced FLUX-mimic: the next generation of video-action models.

FLUX 1 and FLUX 2 generate images. FLUX 3 expands into multimodality and generates audio-visual content jointly - and, at the same time, provides the foundation of FLUX-mimic: A video-action model, developed in collaboration with mimic, running robots that have been tested and deployed at Audi.

At first glance, producing convincing visual content and controlling robots seem to have little in common. One requires generating pixels, the other an understanding of how the physical world responds when you touch and manipulate it. If one model does both, it was never really only a content creation model.

(Excerpted from the original; full article via the source link below.)

This story hit the Hacker News front page today (297 points / 47 comments, via bfl.ai). Our Tech Radar aggregates daily signals on AI engineering, backend architecture and DevOps — browse the related services and further reading below, or get in touch with our team.

Source: Hacker News

Related Services

Related Reading