[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"news-slug-nvidia-88pct-margin-nvidia-tax-self-chip":3,"topics-all":36,"news-related-12671bb9-c100-4ca3-82e0-dc7c26286a0e":55},{"id":4,"title":5,"summary":6,"content":6,"original_url":7,"source_id":8,"tags":9,"translations":23,"news_slug":29,"published_at":30,"created_at":31,"modified_at":32,"is_published":33,"publish_type":34,"image_url":13,"view_count":35},"12671bb9-c100-4ca3-82e0-dc7c26286a0e","英伟达 GPU 毛利率 88%：AI 算力成本高企下的自研芯片暗潮","英伟达控制着 81% 的数据中心 AI 芯片市场，上个财年数据中心业务收入 1937 亿美元，毛利率高达 75%。对英伟达顶尖 GPU 芯片的拆解报告显示，其制造成本约 3300 美元，但售价高达 2.8 万美元，利润率高达 88%。如此高的利润，本质上是一种向整个 AI 行业征收的英伟达税。\n\n这种超额利润正在推动整个行业寻找替代方案。Google 的 TPU、亚马逊的 Trainium、微软的 Maia、Meta 的 MTIA，以及 OpenAI 与博通合作设计的 AI 芯片，都是科技巨头们去英伟达化的尝试。数据中心周围的居民在不知不觉中承担了这场算力博弈的成本——电费账单里有一部分实际上是在向英伟达缴税。\n\n这场税的压力正在加速两个趋势：一是自研芯片的投入，从云厂商到应用层公司都在探索自建 ASIC；二是模型层面的效率优化，量化、蒸馏、长上下文等技术的突破，本质上都是在用更少的算力完成同等任务。英伟达的护城河短期内难以撼动，但当整个行业都在想办法绕开它时，拐点或许只是时间问题。","https:\u002F\u002Fwww.solidot.org\u002Fstory?sid=84438","d59894d3-308e-4fd8-8865-86dc1eeac4a2",[10,14,17,20],{"id":11,"name":12,"slug":12,"description":13,"color":13},"7ac06d8e-b074-4147-abfc-ffaa4c6b8744","ai-efficiency",null,{"id":15,"name":16,"slug":16,"description":13,"color":13},"e0d31e94-ce47-4c8f-831c-d3d2926d42f3","hardware",{"id":18,"name":19,"slug":19,"description":13,"color":13},"0a93ec8e-ea39-4693-81de-563ca8c173f7","inference",{"id":21,"name":22,"slug":22,"description":13,"color":13},"8dac812d-3839-4abe-a855-5f56ec9515fd","nvidia",[24],{"id":25,"lang":26,"title":27,"summary":28,"content":13},"5de98c44-49c2-4588-877d-5adadef79de9","en","NVIDIA's 88% margins fuel the custom-chip undercurrent","Solidot reports on NVIDIA's GPU gross margin reaching 88%, highlighting the undercurrent of self-developed chips under high AI compute costs. As enterprises look to reduce dependence on NVIDIA, the company's profitability comes under pressure, with custom ASICs from Google, Amazon, and others starting to compete in the inference market.","nvidia-88pct-margin-nvidia-tax-self-chip","2026-05-29T22:00:00Z","2026-05-29T22:06:48.401724Z","2026-08-19T02:08:40.142862Z",true,"agent",193,[37,46],{"slug":38,"tag_slug":38,"title_zh":39,"title_en":40,"intro_zh":41,"intro_en":42,"id":43,"is_active":33,"created_at":44,"modified_at":45},"ai-for-science","AI for Science 2026：从 UniPert 到 GPT-Rosalind 的硬核进化","AI for Science 2026: from UniPert to GPT-Rosalind","生命科学、化学材料、物理世界模型——AI 正在从\"语言工具\"变成\"实验伙伴\"。本专题收录 AI 在三大科学方向的关键节点：UniPert 统一基因与化学扰动空间、GPT-Rosalind 端到端生命科学推理、达摩院 AI 智能体 28 小时找到 4 种超导新材料、Anthropic Claude Science 把工作台做成标准品。","From language tool to lab partner — AI is reshaping life sciences, chemistry\u002Fmaterials, and physical world models. This topic covers the key milestones: UniPert unifying genetic-chemical perturbation spaces, GPT-Rosalind's end-to-end life-sciences reasoning, DAMO's AI agent discovering 4 superconducting materials in 28 hours, and Anthropic's Claude Science workbench going mainstream.","988a4300-5fab-41c4-b5d8-63711a2dc757","2026-09-10T01:34:15.296649Z","2026-09-10T01:34:15.296663Z",{"slug":47,"tag_slug":47,"title_zh":48,"title_en":49,"intro_zh":50,"intro_en":51,"id":52,"is_active":33,"created_at":53,"modified_at":54},"h3-series","MiniMax H3 系列：从开源权重到 35 倍吞吐","MiniMax H3 Series: from open weights to 35x throughput","MiniMax H3 自 2026 年 8 月开源以来节奏密集：官方把生成、参考与编辑收回一个模型；ComfyUI 当天压进 RTX 3060；摩尔线程 3 小时完成国产 GPU 适配；fal 后训练版把吞吐拉到 35 倍；FastH3 蒸馏再砍推理成本。本专题持续追踪 H3 的发布—开源—蒸馏—部署全链路。","Since MiniMax open-sourced H3 in August 2026 the pace has been relentless: one unified omni-modal model, same-day ComfyUI support down to an RTX 3060, a 3-hour Day-0 port to Moore Threads GPUs, fal's post-trained H3 Max at 35x throughput, and FastH3 distillation cutting inference cost further. This topic tracks the full H3 chain — release, open weights, distillation, deployment.","83ef0daa-3c31-4cb3-86ed-e5ee58654d5f","2026-09-08T07:33:19.942193Z","2026-09-08T07:33:19.942209Z",{"items":56},[57,62,67,72,77,82],{"id":58,"title":59,"news_slug":60,"published_at":61},"a235e2ab-5b61-47b2-a85a-5f8a1d438624","AMD 收购 Taalas:把\"为单一模型造芯\"的路子,搬进 Instinct 体系","amd-acquires-taalas-hardwired-inference-silicon","2026-08-09T06:00:00+00:00",{"id":63,"title":64,"news_slug":65,"published_at":66},"cdc8e3ce-b1aa-4348-9436-04763179af9c","AMD MI455X：Transformers 99.5% 通过率，432GB HBM4","amd-mi455x-huggingface-99-5","2026-07-27T10:30:00+00:00",{"id":68,"title":69,"news_slug":70,"published_at":71},"2f4423e8-d5f9-448d-b0da-a16cf7d627a7","BaseRT：Apple Silicon LLM 推理第一，llama.cpp 1.56×","basert-metal-kernel","2026-07-03T20:05:00+00:00",{"id":73,"title":74,"news_slug":75,"published_at":76},"bac63469-b19c-4619-8709-73656ad0cd9f","Intel × HF 上线 xpu-kernels Skill：LLM Agent 把 vLLM 调过的 Triton 内核再提 2.8×","intel-hf-xpu-kernels-skill-triton-2-8x","2026-06-20T00:16:00+00:00",{"id":78,"title":79,"news_slug":80,"published_at":81},"6242571f-c635-4794-b19a-ba08eac4f6d1","NVIDIA 发布全球首款面向 Agent 时代的 CPU：Vera 已送抵 Anthropic、OpenAI、SpaceXAI","nvidia-vera-cpu-agent-anthropic-openai","2026-05-20T01:30:00+00:00",{"id":83,"title":84,"news_slug":85,"published_at":86},"823e17ef-5927-40e6-9efd-08c4958922f0","DeepSeek 旗舰 V4-Pro 今日退役:552B 的 V4.1-Flash 全面接班","deepseek-v4-pro-retires-flash-routes","2026-09-14T15:10:00+00:00"]