[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"news-slug-meta-muse-image":3,"news-related-19566223-1b02-4e48-8c44-518694edb049":36},{"id":4,"title":5,"summary":6,"content":6,"original_url":7,"source_id":8,"tags":9,"translations":23,"news_slug":29,"published_at":30,"created_at":31,"modified_at":32,"is_published":33,"publish_type":34,"image_url":13,"view_count":35},"19566223-1b02-4e48-8c44-518694edb049","Meta Muse Image 落地：Superintelligence Labs 把多模态推理与图生能力拧成一股","2026 年 7 月 7 日,Meta 在官方 Newsroom 正式推出 Muse Image——这是 Meta 超级智能实验室(MSL)的首个原生图像生成模型,已集成进 Meta AI 助手,并为 Instagram Stories 带来 30 余种 AI 创作特效,WhatsApp 聊天中也已开放图像生成能力(首批有限地区上线)。\n\n与 4 月发布的推理模型 Muse Spark 配对,Muse Image 并非简单的\"按文字画图\":它会先做多步规划——先理清画面版面、再调用实时 Web 上下文,最后融合多张视觉参考图。这意味着用户上传自拍加旅行照合成明信片、把宠物送进名画里,甚至根据一段历史地标文字直接渲染出照片级的画面,模型都能端到端完成,而不是停留在风格滤镜的层面。\n\n技术细节有三个看点:其一,可读文本渲染,模型能直接生成内嵌清晰文字的信息图与说明图;其二,@-mention Instagram 账号,系统会拉取该账号公开照片参与合成;其三,用户可在生成图上直接圈画修改,模型基于完整对话记忆持续迭代,而不会让上下文断裂。\n\n商业化方面,Meta 计划先把 Muse Image 推给广告主(通过 Advantage+ creative),再扩展到 Facebook 和 Messenger;视频侧,Muse Video 已经在路上,延续同一套\"推理 + 生成\"的双模型思路。\n\n如果说 GPT-5.6、Claude Sonnet 5 把\"主力 LLM\"打到新的成本\u002F能力曲线,那 Muse Image 的意义在于:Meta 终于补齐了\"原生图生 + 多模态推理 + 巨型应用分发\"这一完整闭环——把模型能力直接灌进三十亿用户的聊天框。","https:\u002F\u002Fabout.fb.com\u002Fnews\u002F2026\u002F07\u002Fintroducing-muse-image-meta-ai\u002F","788268d5-e014-4faa-976e-813ed5bce335",[10,14,17,20],{"id":11,"name":12,"slug":12,"description":13,"color":13},"5e628969-6d2a-437f-998a-104e4b16cfb1","ai-progress",null,{"id":15,"name":16,"slug":16,"description":13,"color":13},"7e89b5cc-57db-4f37-bc6d-28919a73931c","model-release",{"id":18,"name":19,"slug":19,"description":13,"color":13},"499f4b56-819d-49a3-9609-33e775143b86","multimodal",{"id":21,"name":22,"slug":22,"description":13,"color":13},"c883fd20-1d66-4fb7-9fc7-320fa7f87023","text-to-image",[24],{"id":25,"lang":26,"title":27,"summary":28,"content":13},"4db17b1f-56aa-4cd7-aa74-bbd2936e9c5a","en","Meta Muse Image: multimodal reasoning meets image generation","On July 7, 2026, Meta officially launched Muse Image in its official Newsroom — the first native image generation model from Meta's Superintelligence Labs (MSL), already integrated into the Meta AI assistant, bringing 30+ AI creative effects to Instagram Stories, with image generation also opened in WhatsApp chat (first launch in limited regions). Paired with the April-released reasoning model Muse Spark, Muse Image isn't simply \"draw from text\": it does multi-step planning first — first figuring out the layout, then invoking real-time web context, and finally fusing multiple visual reference images. This means users can upload selfies plus travel photos to synthesize postcards, send their pets into famous paintings, or even directly render photo-grade scenes from a piece of historical landmark text — the model handles it end-to-end, rather than staying at the level of style filters. Three technical highlights: first, readable text rendering — the model can directly generate info-graphics and explanatory images with clear text embedded; second, @-mention Instagram accounts — the system pulls that account's public photos for synthesis; third, users can circle-edit directly on the generated image, with the model iterating based on complete conversation memory rather than breaking the context. On commercialization, Meta plans to push Muse Image to advertisers first (via Advantage+ creative), then expand to Facebook and Messenger; on the video side, Muse Video is already on the way, continuing the same \"reasoning + generation\" dual-model thinking. If GPT-5.6 and Claude Sonnet 5 hit a new cost\u002Fcapability curve for the \"main LLM\", Muse Image's significance is: Meta finally completes the closed loop of \"native image generation + multimodal reasoning + mega-app distribution\" — directly pouring model capability into the chat windows of three billion users.","meta-muse-image","2026-07-07T20:01:00Z","2026-07-07T20:06:41.350434Z","2026-08-19T02:08:40.142862Z",true,"agent",277,{"items":37},[38,43,48,53,58,63],{"id":39,"title":40,"news_slug":41,"published_at":42},"619ad304-0d2a-4dba-b91e-19414d036746","Grok Imagine Image 2.0：文生图 Arena 双榜第二","grok-imagine-image-2-0-arena-second","2026-08-13T02:00:00+00:00",{"id":44,"title":45,"news_slug":46,"published_at":47},"dfc3dec4-2211-4c7e-b6ff-9e0d9a479ec4","微软与 Mistral 签下数十亿美元协议:Vera Rubin GPU 上的「欧洲主权云」开始落地","microsoft-mistral-vera-rubin-sovereign","2026-07-22T02:00:00+00:00",{"id":49,"title":50,"news_slug":51,"published_at":52},"f1397080-206a-469f-846c-932a4b3ab8f9","京东开源 JoyAI-Image：统一多模态基础模型，把「理解-生成-编辑」拧成一个闭环","jd-joyai-image","2026-07-20T06:00:00+00:00",{"id":54,"title":55,"news_slug":56,"published_at":57},"747917d5-e65b-46dd-b0db-40dfa119cdd1","Reve 2.1 用 Layout-First 架构 + 4K 输出登顶 Arena #2：用不到头部 1\u002F10 算力做独立图像生成实验室","reve-2-1-layout-first","2026-07-16T02:14:00+00:00",{"id":59,"title":60,"news_slug":61,"published_at":62},"a6119d74-007a-4692-bb4e-d85b562a9d66","字节 SpectraReward：自我奖励 T2I 干翻 30× 大模型","bytedance-spectra-reward","2026-07-15T04:30:00+00:00",{"id":64,"title":65,"news_slug":66,"published_at":67},"8865aca6-336a-4dbc-964a-de4afecb25c1","GenCeption 把视频生成模型改造成「通用视觉大脑」：Kaiming He 也在作者里","genception-kaiming-he","2026-07-13T10:01:00+00:00"]