OpenRouter Token Usage Soars 25,000% Since Jan 2025
OpenRouter, a platform that lets developers plug AI models into their products, has seen an explosive rise in token consumption. Weekly usage jumped from 0.5 trillion tokens in January 2025 to 126.2 trillion tokens, a 25,000% increase.
Key points
- Token consumption on OpenRouter jumped 25,000% from 0.5T to 126.2T tokens since Jan 2025.
- GPT 5.6 Luna dominates token use, but high token count may reflect model size, not user base.
- Chinese models Kimi, GLM, DeepSeek see tenfold monthly spending growth in 2026.
The spike is driven largely by reasoning models that generate vast amounts of internal "thinking" tokens before producing an answer. A small uptick in user activity can therefore translate into a massive token surge, especially from unoptimized agentic systems.
OpenAI’s GPT 5.6 Luna currently dominates token use on the platform, though its high token count may reflect the model’s size rather than popularity. On the revenue side, OpenAI’s Astra leads, while Chinese models Kimi, GLM and DeepSeek are experiencing tenfold monthly spending growth in 2026.
The story so far
2 episodes →- OpenRouter Token Usage Soars 25,000% Since Jan 2025 this story
OpenRouter's staggering token chart is the AI bubble debate in a single image
The Decoder · 17 September 2026
OpenRouter's staggering token chart is the AI bubble debate in a single image
On OpenRouter, a platform developers use to plug AI models into their products, weekly token consumption has surged more than 25,000 percent since January 2025, from 0.5 trillion to 126.2 trillion tokens. Tokens are the basic unit of AI processing, like gallons of gas for a car. But this chart says less about booming AI adoption than about how inflated token metrics have become.
The surge shouldn't be confused with a matching jump in actual usage or business value. Reasoning models generate massive amounts of "thinking" tokens before producing an answer, inflating the count dramatically. A small uptick in usage can mean a huge spike in token consumption, especially from unoptimized agentic AI systems that burn through tokens at staggering rates.
OpenAI's GPT 5.6 Luna recently dominated token consumption on OpenRouter, though that doesn't necessarily mean more people are using it. The model may simply generate more tokens per prompt. On the revenue side, OpenAI's Astra leads. Chinese models like Kimi, GLM, and DeepSeek are growing fast too, with monthly spending up tenfold in 2026, though from a much smaller base.
This text was published by The Decoder and written by Matthias Bastian. It is reproduced here with attribution so you can read it in full; the rights remain with the publisher. Read it at the source ↗
The headline, key points and digest above were generated by Digest AI's editorial model from the linked sources. Automated summaries can contain errors: the sources are the record. Spotted a mistake? Tell us.
More in Generative AI & Models
All →- Anthropic's Invisible Text Watermarking in Claude AI · 1 src
- Z.ai Deploys GLM‑5.3‑Flash on 100,000‑Chip Chinese Cluster · 1 src
- AI Chatbots Lag Behind on Recent Information · 1 src
- OpenAI's GPT-6 Astra Drives Enterprise Spend Ahead of Anthropic · 5 src
- GPU Cache Placement: Insights for Efficient Sessions · 1 src
Comments
via GitHub Discussions