← 返回技术雷达
Hacker News tech

Flux 3 X Mimic: The Next Generation of Video-Action Models

Hacker News 热议:Flux 3 X Mimic: The Next Generation of Video-Action Models(297 赞 / 47 评论,来源 bfl.ai)

一句话概要

FLUX 3 is running robots. Built with mimic and deployed at Audi, FLUX-mimic shows content creation and physical AI share one foundation: a model that understands the world.

原文开头节选

Back to blog Research Models FLUX 3 x mimic: The Next Generation of Video-Action Models An early version of FLUX 3, our new multimodal foundation model , is now running on robots. We gave mimic robotics early access to FLUX.3. Their strength in robot learning and deployment, combined with the model’s world knowledge and BFL’s foundation model expertise, produced FLUX-mimic: the next generation of video-action models.

FLUX 1 and FLUX 2 generate images. FLUX 3 expands into multimodality and generates audio-visual content jointly - and, at the same time, provides the foundation of FLUX-mimic: A video-action model, developed in collaboration with mimic, running robots that have been tested and deployed at Audi.

At first glance, producing convincing visual content and controlling robots seem to have little in common. One requires generating pixels, the other an understanding of how the physical world responds when you touch and manipulate it. If one model does both, it was never really only a content creation model.

(以上为原文节选,完整内容见下方”原文来源”)

这条动态今日登上 Hacker News 首页(297 赞 / 47 评论,来源 bfl.ai)。技术雷达每日自动聚合 AI 工程、后端架构、DevOps 方向的前沿动态;相关工程落地可浏览下方的相关服务与延伸阅读,或直接与我们团队交流。

原文来源: Hacker News

相关服务