[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"news-slug-alibaba-happyoyster-1-0":3,"topics-all":36,"news-related-c75518c3-da86-45a3-80fd-004e4650c06b":55},{"id":4,"title":5,"summary":6,"content":6,"original_url":7,"source_id":8,"tags":9,"translations":23,"news_slug":29,"published_at":30,"created_at":31,"modified_at":32,"is_published":33,"publish_type":34,"image_url":13,"view_count":35},"c75518c3-da86-45a3-80fd-004e4650c06b","阿里把「世界模型」搬上百炼:HappyOyster 1.0 用一句话生成可实时交互的开放世界","7月19日，阿里云百炼正式上线开放式世界模型 HappyOyster 1.0（快乐生蚝）。和此前的文本生视频模型不同，HappyOyster 把\"世界\"作为基本生成单位——用户输入一句话或一张图，模型就能实时构建可探索、可交互的开放世界，音画同步并能持续响应 1 分钟以上的交互指令。\n\n技术上 HappyOyster 1.0 提供两个核心模式：世界探索（Adventure）让用户在生成世界里自由行动、改变状态；实时导演（Directing）则把控制权交给用户，用文字指令实时调度镜头、角色与剧情走向，可暂停、改写、回溯。同时提供 Android \u002F iOS \u002F Web 三端 SDK 与 Open API——这意味着它不是 demo 视频，而是面向企业开发者的可生产工具。\n\n从行业角度看，这代表了\"世界模型\"在国内云厂的落地尝试。商汤 U1 Pro 强调长程智能体基座，蚂蚁 LingBot-Video 聚焦具身视频，而阿里这次选择了把\"生成式世界\"做成可实时交互的产品形态，定位于数字人陪伴、互动剧、POV 沉浸式体验等场景。\n\n真正的看点不是\"能不能生成画面\"，而是生成画面之后能不能保持状态一致、响应长时序交互。如果 HappyOyster 真的把\"持续 1 分钟以上主体动作+环境交互+音画同出\"做到产品级，意味着世界模型从论文概念走进开发者生态——这是从 Sora 时代向 Genie 时代过渡的关键信号。","https:\u002F\u002Fwww.ithome.com\u002F0\u002F978\u002F644.htm","ff2b84cb-b1ac-40ce-9da6-d9ef6b82d382",[10,14,17,20],{"id":11,"name":12,"slug":12,"description":13,"color":13},"e676a5cf-1f24-472f-a765-86fa21a1bc3c","ai-model",null,{"id":15,"name":16,"slug":16,"description":13,"color":13},"a8002d98-9df1-4ab9-94d4-a7625af634c4","china-ai",{"id":18,"name":19,"slug":19,"description":13,"color":13},"7e89b5cc-57db-4f37-bc6d-28919a73931c","model-release",{"id":21,"name":22,"slug":22,"description":13,"color":13},"499f4b56-819d-49a3-9609-33e775143b86","multimodal",[24],{"id":25,"lang":26,"title":27,"summary":28,"content":28},"b78010f8-a04b-4b93-8af9-489b29f5044b","en","HappyOyster 1.0: one prompt, one interactive open world","On July 19, Alibaba Cloud Bailian officially launched the open-style world model HappyOyster 1.0 (Happy Oyster). Unlike previous text-to-video models, HappyOyster takes \"the world\" as the basic unit of generation — input a sentence or image, and the model can build an explorable, interactive open world in real time, with synchronized audio-video and the ability to continuously respond to over a minute of interactive commands. Technically HappyOyster 1.0 offers two core modes: World Exploration (Adventure) lets users move freely and change state in the generated world; Real-time Directing hands control to the user, using text commands to schedule camera, characters, and plot in real time — pause, rewrite, roll back. It also provides Android \u002F iOS \u002F Web SDKs and an Open API — meaning this isn't a demo video, but a production-ready tool for enterprise developers. From an industry perspective, this represents domestic cloud vendors' attempt to land \"world models\". SenseTime's U1 Pro emphasizes long-horizon agent base, Ant Group's LingBot-Video focuses on embodied video, and Alibaba has chosen to turn \"generative world\" into a product form for real-time interaction, targeting scenarios like digital-human companionship, interactive drama, and POV immersive experiences. The real story isn't \"can it generate frames\" but whether the generated frames can keep state consistent and respond to long-time-series interaction. If HappyOyster actually delivers \"1+ minute of continuous subject motion + environment interaction + audio-video output\" at product level, it means world models are moving from paper concept to developer ecosystem — a key signal of the transition from the Sora era to the Genie era.","alibaba-happyoyster-1-0","2026-07-19T22:00:00Z","2026-07-19T22:06:22.946015Z","2026-08-19T02:08:40.142862Z",true,"agent",237,[37,46],{"slug":38,"tag_slug":38,"title_zh":39,"title_en":40,"intro_zh":41,"intro_en":42,"id":43,"is_active":33,"created_at":44,"modified_at":45},"ai-for-science","AI for Science 2026：从 UniPert 到 GPT-Rosalind 的硬核进化","AI for Science 2026: from UniPert to GPT-Rosalind","生命科学、化学材料、物理世界模型——AI 正在从\"语言工具\"变成\"实验伙伴\"。本专题收录 AI 在三大科学方向的关键节点：UniPert 统一基因与化学扰动空间、GPT-Rosalind 端到端生命科学推理、达摩院 AI 智能体 28 小时找到 4 种超导新材料、Anthropic Claude Science 把工作台做成标准品。","From language tool to lab partner — AI is reshaping life sciences, chemistry\u002Fmaterials, and physical world models. This topic covers the key milestones: UniPert unifying genetic-chemical perturbation spaces, GPT-Rosalind's end-to-end life-sciences reasoning, DAMO's AI agent discovering 4 superconducting materials in 28 hours, and Anthropic's Claude Science workbench going mainstream.","988a4300-5fab-41c4-b5d8-63711a2dc757","2026-09-10T01:34:15.296649Z","2026-09-10T01:34:15.296663Z",{"slug":47,"tag_slug":47,"title_zh":48,"title_en":49,"intro_zh":50,"intro_en":51,"id":52,"is_active":33,"created_at":53,"modified_at":54},"h3-series","MiniMax H3 系列：从开源权重到 35 倍吞吐","MiniMax H3 Series: from open weights to 35x throughput","MiniMax H3 自 2026 年 8 月开源以来节奏密集：官方把生成、参考与编辑收回一个模型；ComfyUI 当天压进 RTX 3060；摩尔线程 3 小时完成国产 GPU 适配；fal 后训练版把吞吐拉到 35 倍；FastH3 蒸馏再砍推理成本。本专题持续追踪 H3 的发布—开源—蒸馏—部署全链路。","Since MiniMax open-sourced H3 in August 2026 the pace has been relentless: one unified omni-modal model, same-day ComfyUI support down to an RTX 3060, a 3-hour Day-0 port to Moore Threads GPUs, fal's post-trained H3 Max at 35x throughput, and FastH3 distillation cutting inference cost further. This topic tracks the full H3 chain — release, open weights, distillation, deployment.","83ef0daa-3c31-4cb3-86ed-e5ee58654d5f","2026-09-08T07:33:19.942193Z","2026-09-08T07:33:19.942209Z",{"items":56},[57,62,67,72,77,82],{"id":58,"title":59,"news_slug":60,"published_at":61},"d055ddb8-4d82-4523-99b7-39c5f77e2ff7","PhysBrain 1.5 开源：8B 具身基座 28 项评测均分 72.5，官方称追平 GPT-6-Astra","physbrain-1-5-open-embodied-base","2026-09-16T21:07:24+00:00",{"id":63,"title":64,"news_slug":65,"published_at":66},"17006864-46a5-405c-a8cc-24507bbc5e37","YuE2-3B 开源:乐谱可编辑的音乐生成,官方基准反超 Suno v5","yue2-3b-editable-music-generation","2026-09-10T13:20:00+00:00",{"id":68,"title":69,"news_slug":70,"published_at":71},"6062d551-9068-4a9f-ae8e-4e99269cd838","Muse Voice Transcribe 发布:流式转写、20+ 说话人分离、端点检测,Meta 全塞进一个模型","meta-muse-voice-transcribe-streaming-asr","2026-09-05T13:11:00+00:00",{"id":73,"title":74,"news_slug":75,"published_at":76},"7ef479ae-66af-463a-802f-07a84ade93b1","商汤开源 SenseNova-U1.5-8B：原生多模态通吃生成编辑，短板全写进模型卡","sensenova-u1-5-8b-open-source-multimodal","2026-08-25T19:30:00+00:00",{"id":78,"title":79,"news_slug":80,"published_at":81},"ad10985b-425c-4af1-9495-c63792a2b593","腾讯混元把语音识别打到 3% WER：Hy ASR 3.0 preview 让 ASR 从“逐字”走向“读语境”","tencent-hunyuan-hy-asr-3-0-preview-context-aware","2026-08-05T00:00:00+00:00",{"id":83,"title":84,"news_slug":85,"published_at":86},"6ed14a36-a62a-43e8-949a-cf9df4405d98","Seedance 2.5 把视频生成送进 B 端:30 张参考图、API 上火山方舟、徐工小鹏首批接入","seedance-2-5-enterprise-api-b2b","2026-08-01T04:30:00+00:00"]