[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"news-slug-gpt-5-6-launch":3,"news-related-88944bec-d33f-4383-aece-0d5207a06eab":36},{"id":4,"title":5,"summary":6,"content":6,"original_url":7,"source_id":8,"tags":9,"translations":23,"news_slug":29,"published_at":30,"created_at":31,"modified_at":32,"is_published":33,"publish_type":34,"image_url":13,"view_count":35},"88944bec-d33f-4383-aece-0d5207a06eab","GPT-5.6 全面开放:Ultra 把 4 agent 并行写进 API,程序化工具调用把 token 效率再压一档","7 月 9 日,OpenAI 把 GPT-5.6 三个档位 Sol\u002FTerra\u002FLuna 一起推向 GA。比起半年前 preview,真正落地的,是两个把「能力-成本曲线」再压一档的工程选择。\n\nUltra 模式——4 agent 并行编排。传统 scaling 在 thinking time 上做加法;GPT-5.6 走另一条路:4 个 agent 各管一段,主 agent 汇总。BrowseComp 单 agent 90.4%,Ultra 跑到 92.2%;Terminal-Bench 2.1 从 88.8% 抬到 91.9%。OpenAI 把它写进 Responses API 的 multi-agent beta,开发者可直接调,不用自建调度。\n\nProgrammatic Tool Calling(PTC)——工具调用从「透传」变「可编程」。过去每个 tool response 都要塞回模型下一轮;PTC 让模型写一段小程序,内部过滤聚合,只把下一动作送回主循环。早期客户数字:PlayCo 场景构建 token 砍 63.5%、turn 砍 50%;Clio 法律文档 prompt token 砍 38%、质量持平;Rogo 金融研究 output token 砍 24%、完成快 28%。\n\n更长上下文不是新闻(1M 早已 preview),但 GA 在 MRCR v2 8-needle 256K-512K 段跑到 91.5%,GraphWalks 1M BFS 77.1%——这一批旗舰里第一个把 1M 上下文从 benchmark 口号变成工程现实。配合 PTC,长上下文 agent 的成本模型第一次有了支撑点。\n\n拉远看,Sol\u002FTerra\u002FLuna 被定成「durable capability tiers」,会按各自节奏迭代,不再绑定某一代 release。配合三档定价,OpenAI 不再卖「最强模型」,而是卖「按预算可调度的算力梯度」——把「能力」切成可组合的运行时选项。\n\n对搭 agent 平台的团队,直接信号是:多 agent 编排从实验特性变 API 一档;工具调用从透传变可编程。这两个能力,第一次有了头部厂商的稳定接口。","https:\u002F\u002Fopenai.com\u002Findex\u002Fgpt-5-6\u002F","15975962-b5fe-49e5-ae68-687ba6cb7015",[10,14,17,20],{"id":11,"name":12,"slug":12,"description":13,"color":13},"baf131c1-687a-49f4-87f6-4dd87c1c692f","gpt",null,{"id":15,"name":16,"slug":16,"description":13,"color":13},"01598627-1ea6-4b27-a5d8-874971571a71","llm",{"id":18,"name":19,"slug":19,"description":13,"color":13},"7e89b5cc-57db-4f37-bc6d-28919a73931c","model-release",{"id":21,"name":22,"slug":22,"description":13,"color":13},"42e59a88-7795-47dc-a334-ef1e72c24347","openai",[24],{"id":25,"lang":26,"title":27,"summary":28,"content":13},"af61bd2f-116a-4905-ab90-7a79bd8638eb","en","GPT-5.6 opens wide: Ultra writes 4-agent parallelism into API","On July 9, OpenAI pushed all three GPT-5.6 tiers — Sol\u002FTerra\u002FLuna — to GA together. Compared with the half-year-old preview, what really landed are two engineering choices that compress the \"capability-cost curve\" by another tier. **Ultra mode — 4-agent parallel orchestration**. Traditional scaling does addition on thinking time; GPT-5.6 takes a different path: 4 agents each handle a segment, and a main agent consolidates. BrowseComp single-agent is 90.4%, Ultra reaches 92.2%; Terminal-Bench 2.1 climbs from 88.8% to 91.9%. OpenAI writes this into the Responses API's multi-agent beta, so developers can call it directly, no need to build their own scheduler. **Programmatic Tool Calling (PTC) — tool calls from \"passthrough\" to \"programmable\"**. In the past, every tool response had to be stuffed back into the model for the next turn; PTC lets the model write a small program, internally filter and aggregate, and only return the next action to the main loop. Early-customer numbers: PlayCo scenario-construction tokens cut 63.5%, turns cut 50%; Clio legal-document prompt tokens cut 38%, quality flat; Rogo financial-research output tokens cut 24%, completion 28% faster. Longer context isn't news (1M was previewed long ago), but GA's MRCR v2 8-needle 256K-512K segment runs at 91.5%, GraphWalks 1M BFS at 77.1% — the first flagship to take 1M context from benchmark slogan to engineering reality. Paired with PTC, the cost model of long-context Agents has a support point for the first time. Stepping back, Sol\u002FTerra\u002FLuna are positioned as \"durable capability tiers\" and will iterate at their own cadence, no longer tied to a particular release. With three-tier pricing, OpenAI is no longer selling \"the strongest model\", but \"a budget-schedulable compute gradient\" — turning \"capability\" into composable runtime options. For teams building Agent platforms, the direct signal is: multi-agent orchestration moves from experimental feature to one API tier; tool calls from passthrough to programmable. These two capabilities have a first-tier vendor's stable interface for the first time.","gpt-5-6-launch","2026-07-10T06:03:00Z","2026-07-10T06:16:49.796795Z","2026-08-19T02:08:40.142862Z",true,"agent",89,{"items":37},[38,43,48,53,58,63],{"id":39,"title":40,"news_slug":41,"published_at":42},"d95940eb-69c1-467e-9d60-5886ab71d985","GPT-5.6-Cyber 上线、Daybreak 分层、Astra 推迟:OpenAI 把\"网络安全模型\"做成一个独立产品线","openai-gpt-5-6-cyber-daybreak-astra-2026","2026-08-11T04:00:00+00:00",{"id":44,"title":45,"news_slug":46,"published_at":47},"69613959-04c5-43d0-97ec-9311473d8d93","GPT-5.6 三档齐发:用 1\u002F3 token 追平 Mythos,OpenAI 把效率-能力前沿压到新位置","gpt-5-6-three-tiers-1-3-tokens-mythos","2026-06-27T04:00:00+00:00",{"id":49,"title":50,"news_slug":51,"published_at":52},"c801be4d-6309-4d29-958e-5c3b5d38924d","GPT-5.5 Instant 更新与 o3\u002FGPT-4.5 退役：OpenAI 模型策略的重大转向","openai-gpt-5-5-instant-o3-gpt-4-5-retire","2026-05-29T02:06:00+00:00",{"id":54,"title":55,"news_slug":56,"published_at":57},"982093ec-46fb-422a-a201-acb67169015e","GPT-5.5 Instant 成为 ChatGPT 默认模型：幻觉率大幅降低，上下文管理能力显著提升","gpt-5-5-instant-default-chatgpt-low-hallucination","2026-05-05T19:00:00+00:00",{"id":59,"title":60,"news_slug":61,"published_at":62},"d1c7b405-fe4e-40f9-9249-a12e2bba6913","GPT-5.6 八月更新：把「推理强度滑块」下放给 Plus\u002FPro，同时把免费用户拉进 Luna 时代","openai-gpt-5-6-august-update-reasoning-slider","2026-08-10T20:00:00+00:00",{"id":64,"title":65,"news_slug":66,"published_at":67},"c2ee2a09-d001-4740-9820-21fb672eee8b","Copilot 默认模型切到 GPT-5.6 Sol：tokenmaxxing 终结","microsoft-gpt5-6-default-token-budget","2026-08-08T08:00:00+00:00"]