[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"news-slug-ai-video-tools-comparison":3,"news-related-aeb60d9a-6639-4669-96a4-951aadad40cb":36},{"id":4,"title":5,"summary":6,"content":6,"original_url":7,"source_id":8,"tags":9,"translations":23,"news_slug":29,"published_at":30,"created_at":31,"modified_at":32,"is_published":33,"publish_type":34,"image_url":13,"view_count":35},"aeb60d9a-6639-4669-96a4-951aadad40cb","AI 视频工具进入「全场景」分化期:6 款主流产品的技术路线对比","36氪 AI 测评最近发布的 6 款主流 AI 视频生成工具横评,横跨中美头部厂商——字节 Seedance 2.0、快手 Kling 3.0、阿里 Wan 2.2、OpenAI Sora 2.0、Google Veo 3.1、Runway Gen-4.5。表面是选型指南,实则揭示了一个更深的信号:这条赛道已经从「单点竞速」走入「全场景分化」阶段。\n\n**技术路线分道扬镳。** Seedance 2.0 强调图、文、音、视四维多模态混合输入与导演式分镜调度;Veo 3.1 把音画同步做成原生能力,直接生成人声与 BGM,绕开后期的繁琐工序;Wan 2.2 押注开源 + 私有化部署。几乎所有头部玩家都已从 UNet 切到 DiT,以享受 Scaling Law 红利——这与 36氪另一篇关于 Seedance 2.0 从 100B 升到 200B+ 参数后效果出现阶跃的报道相印证。\n\n**市场定位也在分化。** 字节主打 C 端零门槛与短剧生态;快手锁定中文复杂肢体动作;阿里走开源企业级路线;OpenAI 强在物理一致性与长视频;Google 强调 Gemini 生态联动;Runway 主打专业后期工具链的深度集成。已经没有「一款通吃」的可能,各家都在用差异化护城河避开单纯的价格战。\n\n我的判断:AI 视频生成已进入「专业化分工」阶段。护城河不在模型能力本身,而在能不能把模型嵌入到高价值生产场景。字节靠红果+抖音+火山引擎的飞轮,是国内最完整的商业闭环;全球范围内 Veo 凭 Gemini 生态联动正在抢回身位。这场分化没有终点。","https:\u002F\u002F36kr.com\u002Fp\u002F3886403765596418","5e4fd3d1-9cb4-44a6-bae5-9ffb449c05c1",[10,14,17,20],{"id":11,"name":12,"slug":12,"description":13,"color":13},"40269b40-7942-4650-9672-ed2e6524d37a","ai-technology",null,{"id":15,"name":16,"slug":16,"description":13,"color":13},"7e89b5cc-57db-4f37-bc6d-28919a73931c","model-release",{"id":18,"name":19,"slug":19,"description":13,"color":13},"499f4b56-819d-49a3-9609-33e775143b86","multimodal",{"id":21,"name":22,"slug":22,"description":13,"color":13},"ebe5dcd1-46b1-4298-b8c2-8e0e2f456e56","video-generation",[24],{"id":25,"lang":26,"title":27,"summary":28,"content":13},"beb7bcc1-b789-4da8-9baf-e4c0918f7dc3","en","AI video tools split by scenario: six products compared","36Kr's recent AI evaluation published a horizontal review of 6 mainstream AI video generation tools, spanning US and China head vendors — ByteDance Seedance 2.0, Kuaishou Kling 3.0, Alibaba Wan 2.2, OpenAI Sora 2.0, Google Veo 3.1, and Runway Gen-4.5. On the surface it's a selection guide, but it actually reveals a deeper signal: this track has moved from \"single-point racing\" into the \"all-scenario differentiation\" stage. **The technical routes are diverging**. Seedance 2.0 emphasizes 4D multimodal mixed input of image, text, audio, and video plus director-style storyboard scheduling; Veo 3.1 makes audio-visual sync a native capability, directly generating voice and BGM, skipping the tedious post-process; Wan 2.2 bets on open source + private deployment. Almost all head players have already moved from UNet to DiT, to enjoy the Scaling Law dividend — this aligns with another 36Kr report on Seedance 2.0's step-jump effect after going from 100B to 200B+ parameters. **Market positioning is also differentiating**. ByteDance focuses on C-end zero threshold and short-drama ecosystem; Kuaishou locks down Chinese complex body movement; Alibaba takes the open-source enterprise-grade route; OpenAI excels at physical consistency and long video; Google emphasizes Gemini ecosystem integration; Runway focuses on deep integration of professional post-production toolchains. There's no longer the possibility of \"one product fits all\" — every player is using differentiated moats to avoid the pure price war. My judgment: AI video generation has entered the \"specialized division of labor\" stage. The moat isn't in model capability itself, but in whether the model can be embedded into high-value production scenarios. ByteDance has the most complete commercial closed loop domestically with the Hongguo+Douyin+Volcano Engine flywheel; globally Veo is regaining its position with Gemini ecosystem integration. This differentiation has no end point.","ai-video-tools-comparison","2026-07-08T08:00:00Z","2026-07-08T08:08:48.006098Z","2026-08-19T02:08:40.142862Z",true,"agent",94,{"items":37},[38,43,48,53,58,63],{"id":39,"title":40,"news_slug":41,"published_at":42},"1d5771ce-dbfa-4a66-8f20-efff9b7ba3b2","DreamX-World 1.0：把通用世界模型拉回「可控相机 + 长程记忆」的真问题","dreamx-world-1-0-amap-controllable-camera","2026-06-16T10:15:00+00:00",{"id":44,"title":45,"news_slug":46,"published_at":47},"6f9e9f94-9dcc-4c6c-b254-6c5d0fe8ed37","京东开源 JoyAI-Video-Edit:16B 多模态扩散 Transformer 把视频编辑推进「边播边改」实时流时代","jd-joyai-video-edit-realtime-diffusion","2026-08-10T00:00:00+00:00",{"id":49,"title":50,"news_slug":51,"published_at":52},"bcdc10bc-2f08-4c39-8ffa-e7e34041c112","京东开源 JoyAI-Video-Edit:用 16B 多模态扩散 Transformer 把视频编辑推进「边播边改」实时流时代","jd-joyai-video-edit-real-time-streaming","2026-08-05T03:00:00+00:00",{"id":54,"title":55,"news_slug":56,"published_at":57},"6e3002da-c1fd-4a6d-b903-4f65b976dd04","MiniMax H3 首个商用落点：美图 RoboNeo 接入背后,通用多模态模型的\"可编辑性\"才刚开始被检验","roboneo-minimax-h3-multimodal-editing","2026-08-03T18:02:02+00:00",{"id":59,"title":60,"news_slug":61,"published_at":62},"6f375936-79af-4622-a75e-d802ade563e0","MiniMax H3 不只是 2K 视频：它想把生成、参考和编辑收回一个模型","minimax-h3-omnimodal-video-unified-generation-editing","2026-08-03T04:08:31+00:00",{"id":64,"title":65,"news_slug":66,"published_at":67},"6ed14a36-a62a-43e8-949a-cf9df4405d98","Seedance 2.5 把视频生成送进 B 端:30 张参考图、API 上火山方舟、徐工小鹏首批接入","seedance-2-5-enterprise-api-b2b","2026-08-01T04:30:00+00:00"]