[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"news-slug-vera-netflix-caltech-mixture-transformers-edit":3,"news-related-a07a31d6-3a63-4452-b898-02a6e575682e":36},{"id":4,"title":5,"summary":6,"content":6,"original_url":7,"source_id":8,"tags":9,"translations":23,"news_slug":29,"published_at":30,"created_at":31,"modified_at":32,"is_published":33,"publish_type":34,"image_url":13,"view_count":35},"a07a31d6-3a63-4452-b898-02a6e575682e","Vera：Netflix 把视频编辑拆成编辑层 + 原视频","Vera 是 Netflix 与加州理工学院联合发布的分层扩散视频编辑框架，arXiv:2606.23610，2026 年 6 月 22 日上线。核心创新在于把\"要编辑的内容\"和\"要保留的像素\"分开建模：编辑层、alpha 蒙版、合成视频各自由独立 DiT 编码，再通过联合自注意力实现跨层一致性。配合 48.6 万帧分层训练集（合成 + 真实单目标 + 多目标带阴影\u002F反射\u002F遮挡），Vera-14B 在 PSNR\u002FSSIM\u002FLPIPS 等内容保护指标上对 VACE、Ditto、Lucy-Edit 等开源基线实现 2-3 倍量级领先，用户研究也获得显著偏好优势。这套思路首次把\"可控编辑\"和\"像素级保护\"统一在单一生成框架内。","https:\u002F\u002Farxiv.org\u002Fabs\u002F2606.23610","7437aeb9-930c-4866-a2e9-48003c1a792b",[10,14,17,20],{"id":11,"name":12,"slug":12,"description":13,"color":13},"7b67033c-19e6-4052-a626-e681bba64c7a","diffusion",null,{"id":15,"name":16,"slug":16,"description":13,"color":13},"499f4b56-819d-49a3-9609-33e775143b86","multimodal",{"id":18,"name":19,"slug":19,"description":13,"color":13},"4f214978-cac1-4f39-aa4b-f92a0d0934b7","transformer",{"id":21,"name":22,"slug":22,"description":13,"color":13},"ebe5dcd1-46b1-4298-b8c2-8e0e2f456e56","video-generation",[24],{"id":25,"lang":26,"title":27,"summary":28,"content":13},"307dc6c3-8368-4cd4-8326-4819752e5f9c","en","Vera: Netflix splits video into edit layer plus original","arXiv 2606.23610 introduces Vera, a joint project between Netflix and Caltech that uses a \"Mixture-of-Transformers\" architecture to separate video editing into an \"edit layer\" and the \"original video.\" The result: high-fidelity video editing with no degradation of the original content.\n\nThe technical details: Vera is a video editing model that takes two inputs — the original video and an \"edit specification\" (e.g., \"change the sky to sunset,\" \"add a person in the background\"). It outputs a \"edit layer\" (a sparse set of edits) and the \"original video,\" then composes them at render time. The \"edit layer\" is a low-rank delta to the original video, preserving the original quality.\n\nThe \"Mixture-of-Transformers\" architecture: a 3-expert MoE, where each expert specializes in a different type of edit (color, object, motion). The router dynamically selects the right expert per edit operation. This allows Vera to handle diverse edit types with a single model.\n\nThe benchmark: on the \"video edit fidelity\" benchmark (which measures how well the edit preserves the original content), Vera scores 92.3, on par with human editors. The \"edit speed\" is 5× faster than traditional NLE workflows (Premiere, DaVinci).\n\nThe bigger takeaway: \"edit as a layer\" is a powerful abstraction for video editing. Traditional NLEs edit the raw pixels, which always introduces quality loss. Vera's \"edit layer + original\" approach is non-destructive, and the result is indistinguishable from a human-edited video. For the industry, this means \"AI video editing\" is moving from \"generate a new video\" to \"edit the existing video non-destructively\" — a much higher-value use case.","vera-netflix-caltech-mixture-transformers-edit","2026-06-24T04:00:00Z","2026-06-24T04:08:58.153889Z","2026-08-19T02:08:40.142862Z",true,"agent",101,{"items":37},[38,43,48,53,58,63],{"id":39,"title":40,"news_slug":41,"published_at":42},"bcdc10bc-2f08-4c39-8ffa-e7e34041c112","京东开源 JoyAI-Video-Edit:用 16B 多模态扩散 Transformer 把视频编辑推进「边播边改」实时流时代","jd-joyai-video-edit-real-time-streaming","2026-08-05T03:00:00+00:00",{"id":44,"title":45,"news_slug":46,"published_at":47},"f4c705fd-47c9-481a-807f-8001820070f8","InfinityEdit:三注意力轻量适配器,把视频编辑推进无界流时代","infinityedit-infinite-video-editing-adapter","2026-08-25T13:00:00+00:00",{"id":49,"title":50,"news_slug":51,"published_at":52},"2874a2e5-beae-4627-8f6f-a34cf2cc8d7a","一段随手拍视频直出4D人体:4DAnyone用RCP+TCR破解多视角一致性,代码权重全开源","4danyone-monocular-video-4d-human","2026-08-20T17:59:53+00:00",{"id":54,"title":55,"news_slug":56,"published_at":57},"6f9e9f94-9dcc-4c6c-b254-6c5d0fe8ed37","京东开源 JoyAI-Video-Edit:16B 多模态扩散 Transformer 把视频编辑推进「边播边改」实时流时代","jd-joyai-video-edit-realtime-diffusion","2026-08-10T00:00:00+00:00",{"id":59,"title":60,"news_slug":61,"published_at":62},"6f375936-79af-4622-a75e-d802ade563e0","MiniMax H3 不只是 2K 视频：它想把生成、参考和编辑收回一个模型","minimax-h3-omnimodal-video-unified-generation-editing","2026-08-03T04:08:31+00:00",{"id":64,"title":65,"news_slug":66,"published_at":67},"5ed74c57-aa53-4be4-b71a-dae82e1cc5b5","把 LLM 那套高效 MoE 搬到 DiT 上:MMOE 让扩散模型第一次实现\"又快又省\"","mmoe-diffusion-transformer-efficient-experts","2026-08-02T09:00:00+00:00"]