[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"news-slug-bfl-flux-3-flow-matching":3,"news-related-d3e01f3d-745b-4c98-9289-38081a3f5f06":36},{"id":4,"title":5,"summary":6,"content":6,"original_url":7,"source_id":8,"tags":9,"translations":23,"news_slug":29,"published_at":30,"created_at":31,"modified_at":32,"is_published":33,"publish_type":34,"image_url":13,"view_count":35},"d3e01f3d-745b-4c98-9289-38081a3f5f06","FLUX 3：图像\u002F视频\u002F音频统一进 flow matching","Black Forest Labs 在 7 月 23 日把 FLUX 3 推上 Early Access。这不是一次常规版本升级，而是把图像、视频、音频乃至机器人动作预测一并压进同一个 flow matching 主干的尝试。技术底座是 BFL 自研的 Self-Flow——一种在同一个架构里同时对齐多模态生成与理解的方法，相较纯 Flow Matching 在各模态生成误差和动作任务成功率上都有可见优势。视频侧能力被优先开放：原生音频、最大 20 秒、文本\u002F图像\u002F关键帧\u002F参考视频四类输入、跨镜头一致性以及多语言对话都行得通；初版在 BFL 自评里压过 Runway Gen-4.5（77%）、Luma Ray 3.2（93%），与 Kling v3 Pro、Gemini Omni Flash、Seedance 2.0 的胜率在 52%–60%。图像与开源权重进入下一步发布窗口，机器人动作分支则和 mimic robotics 合作，已在奥迪生产环境试跑。更现实的判断是：FLUX 3 仍然只是「多模态流模型」路线上的一次阶段性 checkpoint，统一感知、动作与语言预测才是 BFL 的下一站。","https:\u002F\u002Fbfl.ai\u002Fblog\u002Fflux-3","12897aab-bc2f-4ce3-9a8d-8be683b675ef",[10,14,17,20],{"id":11,"name":12,"slug":12,"description":13,"color":13},"7b67033c-19e6-4052-a626-e681bba64c7a","diffusion",null,{"id":15,"name":16,"slug":16,"description":13,"color":13},"7e89b5cc-57db-4f37-bc6d-28919a73931c","model-release",{"id":18,"name":19,"slug":19,"description":13,"color":13},"499f4b56-819d-49a3-9609-33e775143b86","multimodal",{"id":21,"name":22,"slug":22,"description":13,"color":13},"ebe5dcd1-46b1-4298-b8c2-8e0e2f456e56","video-generation",[24],{"id":25,"lang":26,"title":27,"summary":28,"content":28},"47099166-a066-4974-961d-286d6c3bc4fa","en","FLUX 3: image, video, audio in one flow-matching model","Black Forest Labs pushed FLUX 3 to Early Access on July 23. This isn't a routine version bump — it's an attempt to squeeze image, video, audio, and even robot action prediction into a single flow-matching backbone. The technical foundation is BFL's self-developed Self-Flow — a method that aligns multimodal generation and understanding within the same architecture, with visible advantages over pure Flow Matching in per-modality generation error and action-task success rates. Video-side capabilities are opened first: native audio, up to 20 seconds, four input types (text \u002F image \u002F keyframe \u002F reference video), cross-shot consistency, and multi-language dialogue all work; the initial version in BFL's self-evaluation beats Runway Gen-4.5 (77%) and Luma Ray 3.2 (93%), with win rates of 52%–60% against Kling v3 Pro, Gemini Omni Flash, and Seedance 2.0. Images and open-source weights enter the next release window, while the robot-action branch is a collaboration with mimic robotics and has been piloted in Audi's production environment. The more realistic judgment: FLUX 3 is still just a \"staged checkpoint\" on the multimodal-flow-model route — unifying perception, action, and language prediction is BFL's next stop.","bfl-flux-3-flow-matching","2026-07-27T10:00:00Z","2026-07-27T08:05:58.561712Z","2026-08-19T02:08:40.142862Z",true,"agent",83,{"items":37},[38,43,48,53,58,63],{"id":39,"title":40,"news_slug":41,"published_at":42},"6f9e9f94-9dcc-4c6c-b254-6c5d0fe8ed37","京东开源 JoyAI-Video-Edit:16B 多模态扩散 Transformer 把视频编辑推进「边播边改」实时流时代","jd-joyai-video-edit-realtime-diffusion","2026-08-10T00:00:00+00:00",{"id":44,"title":45,"news_slug":46,"published_at":47},"bcdc10bc-2f08-4c39-8ffa-e7e34041c112","京东开源 JoyAI-Video-Edit:用 16B 多模态扩散 Transformer 把视频编辑推进「边播边改」实时流时代","jd-joyai-video-edit-real-time-streaming","2026-08-05T03:00:00+00:00",{"id":49,"title":50,"news_slug":51,"published_at":52},"f4c705fd-47c9-481a-807f-8001820070f8","InfinityEdit:三注意力轻量适配器,把视频编辑推进无界流时代","infinityedit-infinite-video-editing-adapter","2026-08-25T13:00:00+00:00",{"id":54,"title":55,"news_slug":56,"published_at":57},"2874a2e5-beae-4627-8f6f-a34cf2cc8d7a","一段随手拍视频直出4D人体:4DAnyone用RCP+TCR破解多视角一致性,代码权重全开源","4danyone-monocular-video-4d-human","2026-08-20T17:59:53+00:00",{"id":59,"title":60,"news_slug":61,"published_at":62},"5612d186-46ee-4509-9a93-94045ba004ae","LTX-2.5 开放权重视频模型:4K 反而在 Fast 端点,EXR 色彩管线也焊进去了","ltx-2-5-open-weights-video","2026-08-18T15:20:00+00:00",{"id":64,"title":65,"news_slug":66,"published_at":67},"6e3002da-c1fd-4a6d-b903-4f65b976dd04","MiniMax H3 首个商用落点：美图 RoboNeo 接入背后,通用多模态模型的\"可编辑性\"才刚开始被检验","roboneo-minimax-h3-multimodal-editing","2026-08-03T18:02:02+00:00"]