[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"news-slug-gemini-omni-leak-google-io-video":3,"topics-all":36,"news-related-2b76a3ac-58f3-4fd1-9d99-c5c7f51b3a7d":55},{"id":4,"title":5,"summary":6,"content":6,"original_url":7,"source_id":8,"tags":9,"translations":23,"news_slug":29,"published_at":30,"created_at":31,"modified_at":32,"is_published":33,"publish_type":34,"image_url":13,"view_count":35},"2b76a3ac-58f3-4fd1-9d99-c5c7f51b3a7d","Google I\u002FO 前夕泄密：Gemini Omni 或将统一多模态生成","Google I\u002FO 2026（5月19-20日）即将开幕之际，一份来自 TestingCatalog 的泄密报告在AI社区引发广泛关注：Google 正在测试一款名为 Omni 的全新视频生成模型，并将作为 Gemini 的内置功能向消费者开放。\\n\\n泄密截图显示，Gemin视频生成界面的标识从 \"Powered by Veo 3.1\" 悄然换成了 \"Powered by Omni\"。这一变化远不止品牌名称更替那么简单：它可能标志着 Google 正式告别运行两年之久的「多模型分工」策略——Veo 负责视频、Imagen\u002FNano Banana 负责图像，各自独立运作。Omni 则有望将视频和图像生成统一在同一个模型架构之下。\\n\\n这一策略转向不难理解。OpenAI 早已用 Sora 统一了视频生成的品牌出口，而 Google 始终维持着三品牌四工具的分散格局——对普通用户而言，很难说清楚 Gemini 的视频功能究竟「住在哪里」。Omni 若如泄密所示落地，Google 将首次拥有比 OpenAI 更简洁的消费级多模态生成命名体系。\\n\\n不过目前信息仍相当有限。泄密仅透露了品牌归属，尚不清楚 Omni 是基于 Veo 4 的全新模型，还是现有 Nano Banana 系列与 Veo 的组合包装。定价、API 开放时间、输出时长等关键细节也未见披露。\\n\\n随着 Google I\u002FO 的脚步临近，Omni 是 Google 在多模态生成领域憋出的大招，还是一次产品层面的品牌整合？答案即将揭晓。","https:\u002F\u002Fwww.testingcatalog.com\u002Fgoogle-is-testing-new-omni-model-for-video-generation-ahead-of-i-o\u002F","2ae55717-b5ff-48b8-a142-d85d0d854a72",[10,14,17,20],{"id":11,"name":12,"slug":12,"description":13,"color":13},"a9524a82-a7c5-4daa-bb4b-a7ee77bb0b94","gemini",null,{"id":15,"name":16,"slug":16,"description":13,"color":13},"8cf7490f-2449-4ba7-be19-61befa0d92b4","google",{"id":18,"name":19,"slug":19,"description":13,"color":13},"499f4b56-819d-49a3-9609-33e775143b86","multimodal",{"id":21,"name":22,"slug":22,"description":13,"color":13},"ebe5dcd1-46b1-4298-b8c2-8e0e2f456e56","video-generation",[24],{"id":25,"lang":26,"title":27,"summary":28,"content":13},"15ce3245-047c-4be1-9e60-755298c0d0dc","en","Pre-Google I\u002FO Leak: Gemini Omni May Unify Multimodal Generation","On the eve of Google I\u002FO 2026 (May 19-20), a leak report from TestingCatalog has stirred wide attention in the AI community: Google is testing a new video generation model called Omni, to be opened to consumers as a Gemini built-in feature.\n\nThe leaked screenshots show the Gemini video generation interface's branding silently changed from \"Powered by Veo 3.1\" to \"Powered by Omni.\" This change is far more than a brand name swap: it may signal Google officially bidding farewell to the \"multi-model division of labor\" strategy that has run for two years — Veo handling video, Imagen\u002FNano Banana handling images, each operating independently. Omni is expected to unify video and image generation under the same model architecture.\n\nThis strategic pivot is not hard to understand. OpenAI has long used Sora to unify the video-generation brand exit, while Google has maintained a three-brand four-tool dispersed layout — for ordinary users, it's hard to tell where Gemini's video feature \"lives.\" If Omni lands as leaked, Google will for the first time have a more concise consumer-grade multimodal generation naming system than OpenAI.\n\nBut the current information is still quite limited. The leak only reveals brand attribution, and it's not yet clear whether Omni is a brand-new model based on Veo 4, or an existing combination packaging of Nano Banana series and Veo. Key details like pricing, API opening time, and output duration are also not disclosed.\n\nWith Google I\u002FO approaching, is Omni a big move from Google in the multimodal generation space, or just a product-level brand integration? The answer is about to be revealed.","gemini-omni-leak-google-io-video","2026-05-06T02:10:00Z","2026-05-06T10:07:19.500761Z","2026-08-19T02:08:40.142862Z",true,"agent",159,[37,46],{"slug":38,"tag_slug":38,"title_zh":39,"title_en":40,"intro_zh":41,"intro_en":42,"id":43,"is_active":33,"created_at":44,"modified_at":45},"ai-for-science","AI for Science 2026：从 UniPert 到 GPT-Rosalind 的硬核进化","AI for Science 2026: from UniPert to GPT-Rosalind","生命科学、化学材料、物理世界模型——AI 正在从\"语言工具\"变成\"实验伙伴\"。本专题收录 AI 在三大科学方向的关键节点：UniPert 统一基因与化学扰动空间、GPT-Rosalind 端到端生命科学推理、达摩院 AI 智能体 28 小时找到 4 种超导新材料、Anthropic Claude Science 把工作台做成标准品。","From language tool to lab partner — AI is reshaping life sciences, chemistry\u002Fmaterials, and physical world models. This topic covers the key milestones: UniPert unifying genetic-chemical perturbation spaces, GPT-Rosalind's end-to-end life-sciences reasoning, DAMO's AI agent discovering 4 superconducting materials in 28 hours, and Anthropic's Claude Science workbench going mainstream.","988a4300-5fab-41c4-b5d8-63711a2dc757","2026-09-10T01:34:15.296649Z","2026-09-10T01:34:15.296663Z",{"slug":47,"tag_slug":47,"title_zh":48,"title_en":49,"intro_zh":50,"intro_en":51,"id":52,"is_active":33,"created_at":53,"modified_at":54},"h3-series","MiniMax H3 系列：从开源权重到 35 倍吞吐","MiniMax H3 Series: from open weights to 35x throughput","MiniMax H3 自 2026 年 8 月开源以来节奏密集：官方把生成、参考与编辑收回一个模型；ComfyUI 当天压进 RTX 3060；摩尔线程 3 小时完成国产 GPU 适配；fal 后训练版把吞吐拉到 35 倍；FastH3 蒸馏再砍推理成本。本专题持续追踪 H3 的发布—开源—蒸馏—部署全链路。","Since MiniMax open-sourced H3 in August 2026 the pace has been relentless: one unified omni-modal model, same-day ComfyUI support down to an RTX 3060, a 3-hour Day-0 port to Moore Threads GPUs, fal's post-trained H3 Max at 35x throughput, and FastH3 distillation cutting inference cost further. This topic tracks the full H3 chain — release, open weights, distillation, deployment.","83ef0daa-3c31-4cb3-86ed-e5ee58654d5f","2026-09-08T07:33:19.942193Z","2026-09-08T07:33:19.942209Z",{"items":56},[57,62,67,72,77,82],{"id":58,"title":59,"news_slug":60,"published_at":61},"a33fff56-d6b1-47b4-a65e-2250d73c8875","Gemini Omni Flash 体验：多模态视频生成正在越过令人不安的边界","gemini-omni-flash-deepfake-uncomfortable","2026-05-23T08:10:00+00:00",{"id":63,"title":64,"news_slug":65,"published_at":66},"9e0dd6e3-920b-4802-9b8d-23b118c371a1","Gemini Omni 落地 YouTube Shorts：多模态 AI 从技术秀场走向大众创作工具","gemini-omni-youtube-shorts-remix-mass","2026-05-21T01:01:00+00:00",{"id":68,"title":69,"news_slug":70,"published_at":71},"8303f420-b3f1-485a-aa6f-7775256c84a7","Gemini 3.8 Audio 双发:Live 和 Extended Thinking 把思考+说话压到近实时","gemini-3-8-audio-live-extended-thinking","2026-09-16T03:00:00+00:00",{"id":73,"title":74,"news_slug":75,"published_at":76},"039ff515-68e7-4f11-866a-1da97e26eb45","Gemini 3.8 Live 拿下 S2S 实时语音榜第一","gemini-3-8-live-voice-s2s-number-one","2026-09-15T17:00:00+00:00",{"id":78,"title":79,"news_slug":80,"published_at":81},"1ba7e499-d93d-4566-bbeb-0b762904c0ab","Google 把 Lyria 3.5 装进 Gemini 与公开 API:音乐生成从独立工具变成默认选项","lyria-3-5-gemini-app-api","2026-09-06T23:06:32+00:00",{"id":83,"title":84,"news_slug":85,"published_at":86},"9af3dd83-6ed9-498d-9da0-547d917f3e19","语音转文字有了专用模型:Gemini 3.5 Transcribe 上线,出稿快 70%","gemini-35-transcribe-dedicated-asr","2026-08-31T13:30:00+00:00"]