ByteDance's Doubao 2.1 Pro is now serving 180T (trillion) tokens per day, a 6× increase from 6 months ago. The number is significant: it puts Doubao 2.1 Pro in the same league as the largest closed-source models (GPT-5.6 serves ~500T/day, Claude ~150T/day), and far ahead of any open-source model.

The breakdown of the 180T: 40% from Doubao's own products (Doubao chatbot, Jimeng AI, Doubao Code), 35% from enterprise customers (via Volcengine's MaaS platform), 25% from third-party Apps that use Doubao as their backend. The third-party App share is the fastest-growing segment, with a 12× year-over-year increase.

The "Agent era" angle: Tan Dai, ByteDance's VP of Doubao, attributes the growth to the "Agent era" — Doubao 2.1 Pro is heavily used for Agent workloads (long context, tool use, multi-step reasoning), which consume 5-10× more tokens than chat workloads. As more enterprise Agents are deployed, the token volume scales linearly.

The "MaaS scale race" highlight: the Chinese MaaS market is in a "scale race" — the vendor with the most tokens served has the most data, the most feedback, and the best cost optimization. Doubao 2.1 Pro's 180T/day gives ByteDance a significant lead, and the gap is widening.

The bigger takeaway: "tokens per day" is the new "active users" metric for AI platforms. The Agent era is driving an explosion in token consumption, and the vendors that can serve the most tokens at the lowest cost will dominate. For the industry, this means the next round of competition in the Chinese LLM market will be in "infrastructure scale" — i.e., who can serve the most tokens at the lowest latency.