[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"news-slug-nvidia-rtx-spark-120b-laptop":3,"news-related-f4aad332-1f97-4cb2-96a4-37d8d2980728":36},{"id":4,"title":5,"summary":6,"content":6,"original_url":7,"source_id":8,"tags":9,"translations":23,"news_slug":29,"published_at":30,"created_at":31,"modified_at":32,"is_published":33,"publish_type":34,"image_url":13,"view_count":35},"f4aad332-1f97-4cb2-96a4-37d8d2980728","英伟达BW首秀RTX Spark：笔记本本地跑120B大模型","英伟达在Bilibili World首次展出搭载RTX Spark超级芯片的笔记本：Blackwell RTX GPU + 20核Grace CPU通过NVLink-C2C直连，1 Petaflop FP4算力配128GB统一内存，可在本地运行120B参数大模型、上下文拉到100万token——以往只有云端推理集群才碰的配置，现在塞进超薄笔记本。\n\n配套软件栈同样关键：OpenShell运行时把智能体权限拆成可声明策略，NemoClaw负责敏感数据留本地；现场用35B Qwen多模态驱动个人智能体，识别手绘草图后几十秒内本地复刻出完整网页，全程不烧云端token。\"硬件+运行时+隐私层\"三件套标志着NVIDIA把Agent竞争从\"卖卡\"升级成\"卖本地Agent计算平台\"。\n\n同期亮相的DGX Spark桌面超算参数几乎一致（128GB\u002F1 Petaflop），但基于Linux预装全套NVIDIA AI栈，两台ConnectX互联即可处理200B模型，让独立开发者把训练、微调、推理从云端搬回书桌成真。","https:\u002F\u002Fwww.qbitai.com\u002F2026\u002F07\u002F447981.html","3bd971a8-3897-43d9-84ac-43879efd2f94",[10,14,17,20],{"id":11,"name":12,"slug":12,"description":13,"color":13},"fca9258a-9430-455a-b95d-b9fae5e373a8","ai-inference",null,{"id":15,"name":16,"slug":16,"description":13,"color":13},"e0d31e94-ce47-4c8f-831c-d3d2926d42f3","hardware",{"id":18,"name":19,"slug":19,"description":13,"color":13},"01598627-1ea6-4b27-a5d8-874971571a71","llm",{"id":21,"name":22,"slug":22,"description":13,"color":13},"8dac812d-3839-4abe-a855-5f56ec9515fd","nvidia",[24],{"id":25,"lang":26,"title":27,"summary":28,"content":13},"b7d8dfc4-2b17-412a-ba03-373246a2b77c","en","NVIDIA Blackwell RTX Spark: 120B models on laptops","NVIDIA at Bilibili World made its first showing of a laptop equipped with the RTX Spark superchip: Blackwell RTX GPU + 20-core Grace CPU directly connected via NVLink-C2C, 1 Petaflop FP4 compute paired with 128GB unified memory, able to run a 120B-parameter LLM locally with context stretched to 1M tokens — a configuration previously only encountered on cloud inference clusters, now stuffed into an ultra-thin laptop. The accompanying software stack is equally critical: the OpenShell runtime breaks Agent permissions into declarative policies, NemoClaw is responsible for keeping sensitive data local; on-site a 35B Qwen multimodal model drives a personal Agent that, after recognizing hand-drawn sketches, can replicate a complete webpage locally in tens of seconds, with zero cloud tokens burned throughout. The \"hardware + runtime + privacy layer\" trio signals NVIDIA's upgrade of Agent competition from \"selling cards\" to \"selling local Agent compute platforms\". Also unveiled at the same time, the DGX Spark desktop supercomputer has nearly identical specs (128GB \u002F 1 Petaflop), but is Linux-based with the full NVIDIA AI stack pre-installed, and two of them can interconnect via ConnectX to handle a 200B model — making it real for independent developers to move training, fine-tuning, and inference from the cloud back to the desk.","nvidia-rtx-spark-120b-laptop","2026-07-12T12:00:00Z","2026-07-12T12:13:36.773949Z","2026-08-19T02:08:40.142862Z",true,"agent",103,{"items":37},[38,43,48,53,58,63],{"id":39,"title":40,"news_slug":41,"published_at":42},"aadc3e5d-7b81-4996-aed5-12e5a26f6b21","英伟达RTX Spark平台曝光：N2X\u002FN3X路线图初现，端侧推理的芯片博弈","nvidia-rtx-spark-n1x-n2x-n3x-roadmap","2026-06-02T19:00:00+00:00",{"id":44,"title":45,"news_slug":46,"published_at":47},"4aa9534a-778e-4cd7-8194-fdf3097249b8","OpenAI Jalapeño Hot Chips 实测:峰值每瓦 1.9×,延迟压到 1 秒","openai-jalapeno-hot-chips-benchmark-2026","2026-08-26T02:00:00+00:00",{"id":49,"title":50,"news_slug":51,"published_at":52},"a235e2ab-5b61-47b2-a85a-5f8a1d438624","AMD 收购 Taalas:把\"为单一模型造芯\"的路子,搬进 Instinct 体系","amd-acquires-taalas-hardwired-inference-silicon","2026-08-09T06:00:00+00:00",{"id":54,"title":55,"news_slug":56,"published_at":57},"bce0fe8f-14de-4ffc-8c22-2d798e711e73","Kimi K3 上线 48 小时打满集群:开源旗舰正在把推理算力拖进新一轮\"卖方周期\"","kimi-k3-48h-saturate-chinese-compute-supernode","2026-08-02T06:04:11+00:00",{"id":59,"title":60,"news_slug":61,"published_at":62},"30a147aa-e3ed-475c-b15f-9e5ffce6ffc9","英伟达 Vera CPU：DeepInfra 实测 Agent 编排提速 2.2 倍","nvidia-vera-cpu-agent-orchestration","2026-07-22T04:50:00+00:00",{"id":64,"title":65,"news_slug":66,"published_at":67},"39f5dabb-a59e-4672-9caa-446fd6d6b0cd","Tenstorrent 同台刷新三项推理记录：RISC-V + Tensix 把\"GPU = 默认\"撕开一道口子","tenstorrent-risc-v-tensix","2026-06-30T14:05:00+00:00"]