[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"news-slug-videorae-frozen-video-generator":3,"news-related-a818c807-2131-4950-8f51-62847a57db41":36},{"id":4,"title":5,"summary":6,"content":6,"original_url":7,"source_id":8,"tags":9,"translations":23,"news_slug":29,"published_at":30,"created_at":31,"modified_at":32,"is_published":33,"publish_type":34,"image_url":13,"view_count":35},"a818c807-2131-4950-8f51-62847a57db41","VideoRAE 把 frozen 视频基础模型改造成生成器 latent:UCF-101 gFVD 40\u002F93,收敛提速 5×","港中文(深圳)、华中科大与中科大联合发布 VideoRAE,首次证明 V-JEPA 2、VideoMAEv2 等「理解型」视频基础模型的 frozen 表示可直接改造为生成友好 latent:UCF-101 类条件视频生成 AR\u002FDiT 双通路 gFVD 分别达 40 与 93,2B 文生视频对照中替换 LTX-VAE 后 VBench 三维度同步提升且收敛提速约 5×,代码开源。","https:\u002F\u002Farxiv.org\u002Fabs\u002F2607.14088","7437aeb9-930c-4866-a2e9-48003c1a792b",[10,14,17,20],{"id":11,"name":12,"slug":12,"description":13,"color":13},"40269b40-7942-4650-9672-ed2e6524d37a","ai-technology",null,{"id":15,"name":16,"slug":16,"description":13,"color":13},"7b67033c-19e6-4052-a626-e681bba64c7a","diffusion",{"id":18,"name":19,"slug":19,"description":13,"color":13},"b9bd9039-fcdb-41a8-b85b-fc1587def2b9","open-source",{"id":21,"name":22,"slug":22,"description":13,"color":13},"ebe5dcd1-46b1-4298-b8c2-8e0e2f456e56","video-generation",[24],{"id":25,"lang":26,"title":27,"summary":28,"content":28},"487323f7-ea7e-4071-aa83-9340e8181feb","en","VideoRAE turns frozen video models into generator latents","A joint team from CUHK-Shenzhen, Huazhong University of Science and Technology and USTC released VideoRAE, demonstrating for the first time that frozen representations from \"understanding\" video foundation models such as V-JEPA 2 and VideoMAEv2 can be directly adapted into generation-friendly latents. On UCF-101 class-conditional video generation, the AR\u002FDiT two-track gFVD scores reach 40 and 93 respectively. In a 2B text-to-video comparison, replacing LTX-VAE yields simultaneous improvements on all three VBench dimensions and ~5× faster convergence. Code is open-sourced.","videorae-frozen-video-generator","2026-07-20T04:15:00Z","2026-07-20T04:15:17.540809Z","2026-08-19T02:08:40.142862Z",true,"agent",112,{"items":37},[38,43,48,53,58,63],{"id":39,"title":40,"news_slug":41,"published_at":42},"18d2aa73-7244-4b10-b611-46475e17327e","ForgeWM开源:一步去噪72FPS的可玩世界模型,8张卡复现全流程","forgewm-few-step-playable-world-model","2026-08-24T21:10:00+00:00",{"id":44,"title":45,"news_slug":46,"published_at":47},"2874a2e5-beae-4627-8f6f-a34cf2cc8d7a","一段随手拍视频直出4D人体:4DAnyone用RCP+TCR破解多视角一致性,代码权重全开源","4danyone-monocular-video-4d-human","2026-08-20T17:59:53+00:00",{"id":49,"title":50,"news_slug":51,"published_at":52},"5612d186-46ee-4509-9a93-94045ba004ae","LTX-2.5 开放权重视频模型:4K 反而在 Fast 端点,EXR 色彩管线也焊进去了","ltx-2-5-open-weights-video","2026-08-18T15:20:00+00:00",{"id":54,"title":55,"news_slug":56,"published_at":57},"6f9e9f94-9dcc-4c6c-b254-6c5d0fe8ed37","京东开源 JoyAI-Video-Edit:16B 多模态扩散 Transformer 把视频编辑推进「边播边改」实时流时代","jd-joyai-video-edit-realtime-diffusion","2026-08-10T00:00:00+00:00",{"id":59,"title":60,"news_slug":61,"published_at":62},"bcdc10bc-2f08-4c39-8ffa-e7e34041c112","京东开源 JoyAI-Video-Edit:用 16B 多模态扩散 Transformer 把视频编辑推进「边播边改」实时流时代","jd-joyai-video-edit-real-time-streaming","2026-08-05T03:00:00+00:00",{"id":64,"title":65,"news_slug":66,"published_at":67},"75d2f385-00ef-4074-80bf-ee47ba05a4a4","NVIDIA × HF：Diffusers 微调上 H100，Wan 2.2\u002FFLUX.2 打通","nvidia-huggingface-diffusers-h100","2026-07-17T14:00:00+00:00"]