Skip to content
Model Season every model has a season
Market intelligence · Language models

What matters in AI today

Who leads, what changed and what language models are being used for, measured on real token traffic.

Daily data through Sep 10, 2026
History: Jan 1, 2025 to Sep 10, 2026
Full weeks: 85 · Models: 403
Data license: CC BY 4.0

Now

data through Sep 10, 2026

Outside the filters, on purpose. This is the market as of the latest day the source published: who leads, who is rising, who debuted, what it costs and how concentrated it is. A 7-day window (Sep 4, 2026 to Sep 10, 2026) and a 30-day window.

What stands out today · automatic rule

A model launched 13 days ago already takes 15.5% of traffic.

Hy4 preview (Tencent) went from 10.1% to 15.5% in one week. 5 models debuted in the last seven days.

daily volume · 30 days · 9.6T to 19T

Top 5 · last 7 days

123T · 61 active models

  1. 1Hy4 previewTencent · China15.5%3
    19T
  2. 2GPT-5.6 LunaOpenAI · US/Canada11.5%
    14T
  3. 3DeepSeek V4 Flash 0731DeepSeek · China10.0%
    12T
  4. 4GLM 5.3 FlashZ.ai (GLM) · China10.0%3
    12T
  5. 5MiMo-V2.5Xiaomi · China4.2%3
    5.2T

Top 5 · last 30 days

451T · 81 active models

  1. 1DeepSeek V4 Flash 0731DeepSeek · China11.3%
    51T
  2. 2GPT-5.6 LunaOpenAI · US/Canada8.4%
    38T
  3. 3Hy4 previewTencent · China6.7%new
    30T
  4. 4MiMo-V2.5Xiaomi · China6.3%
    29T
  5. 5Hy3Tencent · China6.2%
    28T

Who moved this week

Last 7 days vs. the previous 7

up

Hy4 preview+5.5 ppTencent · 10.1%15.5%
Gemini 3.8 Flash+1.8 ppGoogle · 0.3%2.1%
Muse Spark 1.3 Contributor+1.4 ppMeta · 0.1%1.5%

down

MiniMax M3−2.8 ppMiniMax · 6.0%3.2%
Hy3−2.1 ppTencent · 4.8%2.7%
Gemini 3.7 Flash−1.6 ppGoogle · 2.4%0.8%
Volume · 7 days
123T
+13% vs. prior 7 days
Effective price
$1.07
per 1M tokens · 94% priced
Concentration
51.2%
top 5 · HHI 754
Chinese labs
60.8%
open weights 68.2%
Debuts
5
new models in 7 days

Debuted in the last 7 days

first volume recorded in the window

DeepSeek V4.1 FlashDeepSeekSep 10, 20260.55%
GPT-6 AstraOpenAISep 5, 20260.49%
Qwen3.8 Max (0902)QwenSep 5, 20260.29%
Ling 3.0 Flash Sante (free)InclusionAISep 5, 20260.23%
Nex-N2.5-Pro (free)Nex AGISep 9, 20260.09%

Age of the leaders

days since entering the catalog, for the most used models

Hy4 preview13 days15.5%
GPT-5.6 Luna63 days11.5%
DeepSeek V4 Flash 073141 days10.0%
GLM 5.3 Flash15 days10.0%
MiMo-V2.5141 days4.2%
DeepSeek V4 Flash 0423139 days3.9%

2 of the 4 largest entered the catalog in the last 31 days.

Leaders by criterion

Three kinds of leadership, each with its own label: benchmark performance, observed usage and fit for a scenario. There is no “best AI” on a single metric.

Overall evaluationbenchmark performance
Claude Fable 5.1 and Qwen3.8 Max53.4 · tied on the published valueHighest Artificial Analysis Intelligence Index among evaluated models. Artificial Analysis, via OpenRouter, Sep 11, 2026, 96 models evaluated.
Codingbenchmark performance
Claude Fable 5.181.6 · 2nd: Claude Opus 5 (78)Highest Artificial Analysis Coding Index. Artificial Analysis, via OpenRouter, Sep 11, 2026, 147 models evaluated.
Agentsbenchmark performance
Claude Fable 5.158 · 2nd: Claude Opus 5 (56.2)Highest Artificial Analysis Agentic Index. Artificial Analysis, via OpenRouter, Sep 11, 2026, 101 models evaluated.
Usageobserved usage
Hy4 preview15.5% · 2nd: GPT-5.6 Luna (11.5%)Highest token share over the last 7 full days. OpenRouter, daily rankings, Sep 10, 2026.
Share gainobserved usage
Hy4 preview+5.5 ppBiggest share gain in percentage points, last 7 days against the previous 7. OpenRouter, daily rankings, Sep 10, 2026.
Value pickfit for a scenario
GLM 5.3$2.15 / 1M · index 44.9Lowest price per 1M tokens (75% input, 25% output blend) among the 10 highest Intelligence Index scores. Artificial Analysis, via OpenRouter, catalog prices, Sep 11, 2026.

What changed since last week

Up to three changes, each with the evidence and what to watch. The rule detects association; causation takes investigation.

Hy4 preview gained ground

Share went from 10.1% to 15.5% between the last two 7-day windows (+5.5 pp).

What to watch. Is one week a spike or a trend? The rule only calls it sustained when the gain repeats in non-overlapping windows.

MiniMax M3 lost ground

Share went from 6.0% to 3.2% (−2.8 pp).

What to watch. A falling share can mean the others are growing. Check absolute volume on the model page before drawing a conclusion.

DeepSeek V4.1 Flash debuted

First volume recorded on Sep 10, 2026, already with 0.6% of the week's traffic.

What to watch. A debut on the router is not a launch date, and debut share often includes test traffic.

History

From here down, everything responds to the window, grouping and filters. Full weeks, since January 2025.

WindowGroup by403 models, unfiltered
01

What the market uses it for

Snapshot of the source's rolling 7-day window, ending Sep 10, 2026. It is the same for any filtered view: this section does not respond to the History window or filters. The source keeps no history, so this section's series exists only because the site archives one snapshot a day.

What the traffic is used for new

Share of classified tokens by use case, over the 7 days through Sep 10, 2026

By tokens, Code leads with 37.1%; by requests, General use leads, with 49.9%. Among the 12 largest tasks, the heaviest workload is Frontend and UI: 2.9% of tokens from 0.80% of requests (3.6×). The lightest is Classification, with 21.7% of requests and 5.4% of tokens.

See all 29 tasks
TaskMacro% tokens% req.tok÷reqLeader
Workflow executionAgents24.6%9.2%2.7×Hy4 preview
Code generationCode13.1%4.6%2.8×GPT-5.6 Luna
DebuggingCode7.3%2.2%3.3×GPT-5.6 Luna
Multi-step planningAgents6.6%2.7%2.4×GLM 5.3 Flash
File read and writeCode5.9%1.9%3.1×Hy4 preview
ClassificationGeneral use5.4%21.7%0.2×DeepSeek V4 Flash 0731
Data extractionData4.6%14.1%0.3×DeepSeek V4 Flash 0731
Roleplay and fictionGeneral use3.2%7.8%0.4×DeepSeek V4 Flash 0423
ConversationGeneral use3.1%2.5%1.2×Hy4 preview
Frontend and UICode2.9%0.80%3.6×GLM 5.3 Flash
Shell executionCode2.6%1.1%2.4×DeepSeek V4 Flash 0731
Code reviewCode2.5%1.0%2.5×GPT-5.6 Luna
Content writingGeneral use2.3%4.3%0.5×DeepSeek V4 Flash 0731
Q&A and knowledgeGeneral use2.3%2.8%0.8×DeepSeek V4 Flash 0731
Data transformationData1.9%7.9%0.2×GPT-5.6 Luna
Memory extractionAgents1.4%0.90%1.6×DeepSeek V4 Flash 0731
Repo scanningCode1.4%0.50%2.8×GPT-5.6 Luna
Tool callingAgents1.2%2.3%0.5×GPT-5.6 Luna
Research and reportsGeneral use1.2%0.60%2.0×Hy4 preview
SummarizationGeneral use1.1%2.0%0.6×DeepSeek V4 Flash 0731
Customer supportGeneral use0.90%1.6%0.6×GPT-5.6 Luna
DevOps configurationCode0.90%0.20%4.5×GPT-5.6 Luna
Web searchAgents0.80%0.50%1.6×GLM 5.3 Flash
Security auditGeneral use0.80%0.40%2.0×GLM 5.3
SQL and databasesCode0.50%0.20%2.5×DeepSeek V4 Flash 0731
DevOpsGeneral use0.50%0.20%2.5×DeepSeek V4 Flash 0731
Finance and tradingGeneral use0.40%0.40%1.0×DeepSeek V4 Flash 0731
TranslationGeneral use0.30%5.0%0.1×GPT-5.6 Luna
MathGeneral use0.20%0.50%0.4×DeepSeek V4 Flash 0731

Who leads each task new

Click a task in the adjacent card, or pick one here

Agents24.6% of tokens9.2% of requestssource: Workflow Execution

Hy4 preview (Tencent) takes 22.8% of the tokens in workflow execution. The 10 models listed add up to 71.0%; the rest is split among models the source does not list.

Share within the task: the model's tokens in this task divided by all tokens classified in it. The source lists only the largest models in each task.

archive2 snapshots archived so far. The use-case time series builds up here as the archive grows.

A sample classified by the source: only traffic it can attribute to a use case counts, and the other bucket stays out of the denominator. Rolling 7-day windows overlap, so two consecutive snapshots share almost all of their volume.

02

Market size and concentration

How much the router processes, and how widely that volume is spread across models. It is the denominator of every share from here down.

Tokens per week

Trillions of tokens processed each week. Named models plus the source's aggregate line

050T100T150TApr '25JulOctJan '26AprJul
See the numbers
WeekTokens/week
Aug 31, 2026115T
Aug 24, 2026113T
Aug 17, 202693T
Aug 10, 202675T
Aug 3, 202669T
Jul 27, 202657T
Jul 20, 202658T
Jul 13, 202663T
Jul 6, 202653T
Jun 29, 202647T
Jun 22, 202647T
Jun 15, 202647T
Jun 8, 202645T
Jun 1, 202636T
May 25, 202632T
May 18, 202629T
May 11, 202627T
May 4, 202626T
Apr 27, 202624T
Apr 20, 202622T
Apr 13, 202620T
Apr 6, 202621T
Mar 30, 202627T
Mar 23, 202623T
Mar 16, 202620T
Mar 9, 202617T
Mar 2, 202615T
Feb 23, 202614T
Feb 16, 202614T
Feb 9, 202613T
Feb 2, 20269.8T
Jan 26, 20268.3T
Jan 19, 20267.5T
Jan 12, 20267.6T
Jan 5, 20266.4T
Dec 29, 20255.6T
Dec 22, 20255.7T
Dec 15, 20255.8T
Dec 8, 20255.8T
Dec 1, 20256.2T
Nov 24, 20257.2T
Nov 17, 20256.3T
Nov 10, 20255.9T
Nov 3, 20255.8T
Oct 27, 20255.5T
Oct 20, 20254.9T
Oct 13, 20254.6T
Oct 6, 20255.0T
Sep 29, 20255.4T
Sep 22, 20255.3T
Sep 15, 20254.9T
Sep 8, 20254.9T
Sep 1, 20254.6T
Aug 25, 20253.7T
Aug 18, 20253.2T
Aug 11, 20253.2T
Aug 4, 20253.4T
Jul 28, 20253.4T
Jul 21, 20253.0T
Jul 7, 20252.4T
Jun 30, 20252.1T
Jun 23, 20252.4T
Jun 16, 20252.3T
Jun 2, 20252.4T
May 26, 20252.5T
May 19, 20252.4T
May 12, 20252.2T
May 5, 20252.1T
Apr 28, 20251.8T
Apr 21, 20251.8T
Apr 14, 20251.8T
Apr 7, 20252.2T
Mar 31, 20252.0T
Mar 24, 20251.6T
Mar 17, 20251.5T
Mar 10, 20251.4T
Mar 3, 20251.3T
Feb 24, 20251.1T
Feb 17, 20250.95T
Feb 10, 20250.84T
Feb 3, 20250.74T
Jan 27, 20250.57T
Jan 20, 20250.53T
Jan 13, 20250.52T
Jan 6, 20250.50T

Weekly volume went from 0.50T (week of Jan 6, 2025) to 115T (week of Aug 31, 2026): up 231× in 84 weeks.

The largest jump between two consecutive periods came in the week of Aug 24, 2026: +20T, from 93T to 113T. The curve climbs in steps, not in a straight line.

Comparing absolute volume across distant points calls for care: the source itself grew along with the market, and any share at the start is worth far fewer tokens.

Concentration: share of the top 5 models

% of weekly volume captured by the 5 most-used models; the tooltip shows who they are and the HHI

0%20%40%60%80%100%Apr '25JulOctJan '26AprJul
See the numbers and who the five are
WeekTop 5HHI1st2nd3rd4th5th
Aug 31, 202649.8%700Hy4 preview 12.7%GPT-5.6 Luna 11.2%GLM 5.3 Flash 10.7%DeepSeek V4 Flash 0731 10.7%DeepSeek V4 Flash 0423 4.5%
Aug 24, 202645.6%649ox-alpha 13.9%DeepSeek V4 Flash 0731 10.9%MiMo-V2.5 8.1%GPT-5.6 Luna 6.9%Hy3 5.9%
Aug 17, 202650.0%709DeepSeek V4 Flash 0731 12.4%ox-alpha 12.4%MiMo-V2.5 10.6%Hy3 8.8%DeepSeek V4 Flash 0423 5.8%
Aug 10, 202647.3%704DeepSeek V4 Flash 0731 14.8%Hy3 13.2%GPT-5.6 Luna 7.0%DeepSeek V4 Flash 0423 6.4%GLM 5.2 5.8%
Aug 3, 202647.2%655DeepSeek V4 Flash 0731 12.8%Hy3 11.7%DeepSeek V4 Flash 0423 8.5%MiMo-V2.5 7.8%GPT-5.6 Luna 6.4%
Jul 27, 202643.2%597DeepSeek V4 Flash 0423 12.7%MiMo-V2.5 11.1%Hy3 8.5%DeepSeek V4 Pro 0423 5.8%GLM 5.2 5.1%
Jul 20, 202646.9%737MiMo-V2.5 18.0%DeepSeek V4 Flash 0423 11.0%Hy3 6.8%GLM 5.2 5.7%DeepSeek V4 Pro 0423 5.5%
Jul 13, 202653.7%883Hy3 (free) 18.4%MiMo-V2.5 14.8%DeepSeek V4 Flash 0423 8.6%MiniMax M3 6.1%GLM 5.2 5.8%
Jul 6, 202647.0%649Hy3 (free) 11.7%MiMo-V2.5 11.3%DeepSeek V4 Flash 0423 9.9%MiniMax M3 8.1%GLM 5.2 6.1%
Jun 29, 202641.8%563DeepSeek V4 Flash 0423 11.4%MiMo-V2.5 9.4%MiniMax M3 8.8%Hy3 preview 6.7%GLM 5.2 5.5%
Jun 22, 202642.2%557DeepSeek V4 Flash 0423 10.0%MiMo-V2.5 9.6%MiniMax M3 8.0%owl-alpha 7.4%Hy3 preview 7.2%
Jun 15, 202640.4%532DeepSeek V4 Flash 0423 10.6%MiMo-V2.5 8.4%MiniMax M3 8.1%Hy3 preview 7.8%owl-alpha 5.5%
Jun 8, 202642.4%562DeepSeek V4 Flash 0423 9.9%MiniMax M3 9.7%Hy3 preview 9.3%MiMo-V2.5 8.1%owl-alpha 5.5%
Jun 1, 202636.7%481DeepSeek V4 Flash 0423 10.2%Hy3 preview 8.1%MiniMax M3 6.9%MiMo-V2.5 6.1%owl-alpha 5.4%
May 25, 202637.9%495DeepSeek V4 Flash 0423 9.8%Hy3 preview 9.5%Claude Opus 4.7 7.3%Claude Sonnet 4.6 6.0%owl-alpha 5.2%
May 18, 202639.7%551DeepSeek V4 Flash 0423 11.8%Hy3 preview 10.6%Claude Opus 4.7 6.7%Claude Sonnet 4.6 6.5%owl-alpha 4.0%
May 11, 202633.3%439Hy3 preview 9.9%DeepSeek V4 Flash 0423 7.7%Claude Sonnet 4.6 5.8%Claude Opus 4.7 5.7%Gemini 3 Flash Preview 4.3%
May 4, 202631.4%420Hy3 preview (free) 10.4%Kimi K2.6 6.3%Claude Sonnet 4.6 5.6%Claude Opus 4.7 4.8%DeepSeek V4 Flash 0423 4.3%
Apr 27, 202634.0%487Hy3 preview (free) 12.7%Kimi K2.6 7.6%Claude Sonnet 4.6 5.7%Gemini 3 Flash Preview 4.1%Claude Opus 4.7 3.9%
Apr 20, 202629.2%367Kimi K2.6 7.2%Claude Sonnet 4.6 6.2%DeepSeek V3.2 5.8%Claude Opus 4.7 5.3%Gemini 3 Flash Preview 4.7%
Apr 13, 202630.1%390Claude Sonnet 4.6 6.8%DeepSeek V3.2 6.3%Claude Opus 4.6 6.0%mimo-v2-pro 5.6%Gemini 3 Flash Preview 5.6%
Apr 6, 202630.8%406Qwen3.6 Plus (free) 7.9%DeepSeek V3.2 6.0%Claude Opus 4.6 5.7%MiniMax M2.7 5.7%Claude Sonnet 4.6 5.5%
Mar 30, 202643.6%691Qwen3.6 Plus (free) 17.0%mimo-v2-pro 11.4%qwen3.6-plus-preview (free) 6.1%Step 3.5 Flash (free) 4.7%MiniMax M2.7 4.4%
Mar 23, 202639.7%642mimo-v2-pro 17.4%Step 3.5 Flash (free) 6.6%MiniMax M2.7 5.7%DeepSeek V3.2 5.5%Claude Sonnet 4.6 4.6%
Mar 16, 202631.7%420mimo-v2-pro 7.3%Step 3.5 Flash (free) 7.3%MiniMax M2.5 6.4%DeepSeek V3.2 5.6%Claude Sonnet 4.6 5.1%
Mar 9, 202635.7%468MiniMax M2.5 10.3%Step 3.5 Flash (free) 7.9%DeepSeek V3.2 6.1%Gemini 3 Flash Preview 6.0%Claude Sonnet 4.6 5.3%
Mar 2, 202635.5%512MiniMax M2.5 12.6%Gemini 3 Flash Preview 7.0%DeepSeek V3.2 5.6%Claude Opus 4.6 5.2%Step 3.5 Flash (free) 5.0%
Feb 23, 202635.1%498MiniMax M2.5 11.9%Gemini 3 Flash Preview 7.5%DeepSeek V3.2 5.9%Kimi K2.5 5.1%Claude Sonnet 4.6 4.7%
Feb 16, 202643.1%697MiniMax M2.5 18.4%Kimi K2.5 7.4%Gemini 3 Flash Preview 6.2%GLM 5 5.8%DeepSeek V3.2 5.3%
Feb 9, 202637.6%509MiniMax M2.5 11.1%Kimi K2.5 10.0%Gemini 3 Flash Preview 6.0%DeepSeek V3.2 5.8%GLM 5 4.8%
Feb 2, 202638.1%509Kimi K2.5 11.8%Gemini 3 Flash Preview 7.8%Claude Sonnet 4.5 7.2%DeepSeek V3.2 6.8%MiniMax M2.1 4.5%
Jan 26, 202633.9%446Claude Sonnet 4.5 9.4%Gemini 3 Flash Preview 8.3%DeepSeek V3.2 6.2%grok-code-fast-1 5.1%Gemini 2.5 Flash 4.9%
Jan 19, 202636.8%485Claude Sonnet 4.5 9.5%mimo-v2-flash (free) 7.7%grok-code-fast-1 7.2%Gemini 3 Flash Preview 7.2%DeepSeek V3.2 5.3%
Jan 12, 202635.3%467Claude Opus 4.5 8.2%Claude Sonnet 4.5 8.1%mimo-v2-flash (free) 7.1%grok-code-fast-1 6.1%Gemini 3 Flash Preview 5.9%
Jan 5, 202632.7%432Claude Sonnet 4.5 8.3%grok-code-fast-1 6.4%mimo-v2-flash (free) 6.2%Gemini 3 Flash Preview 6.0%Claude Opus 4.5 5.8%
Dec 29, 202532.6%407grok-code-fast-1 7.3%Gemini 2.5 Flash 6.3%DeepSeek V3.2 6.3%Claude Sonnet 4.5 6.3%mimo-v2-flash (free) 6.3%
Dec 22, 202534.0%415grok-code-fast-1 8.5%mimo-v2-flash (free) 6.7%Gemini 2.5 Flash 6.7%Claude Sonnet 4.5 6.4%DeepSeek V3.2 5.8%
Dec 15, 202530.9%400grok-code-fast-1 8.9%Gemini 2.5 Flash 7.5%Claude Sonnet 4.5 7.1%gpt-oss-120b 3.8%Claude Opus 4.5 3.6%
Dec 8, 202535.3%451grok-code-fast-1 10.7%Gemini 2.5 Flash 7.8%Claude Sonnet 4.5 7.5%gpt-oss-120b 5.6%Claude Opus 4.5 3.7%
Dec 1, 202540.4%565grok-code-fast-1 14.2%grok-4.1-fast (free) 9.5%Claude Sonnet 4.5 6.9%Gemini 2.5 Flash 6.5%Claude Opus 4.5 3.4%
Nov 24, 202549.6%914grok-4.1-fast (free) 22.0%grok-code-fast-1 13.6%Claude Sonnet 4.5 6.0%Gemini 2.5 Flash 5.3%grok-4.1-fast 2.8%
Nov 17, 202543.1%720grok-code-fast-1 19.9%Claude Sonnet 4.5 7.9%grok-4.1-fast 6.2%Gemini 2.5 Flash 5.8%MiniMax M2 3.2%
Nov 10, 202548.6%969grok-code-fast-1 24.5%Claude Sonnet 4.5 10.2%Gemini 2.5 Flash 6.3%MiniMax M2 3.9%grok-4-fast 3.7%
Nov 3, 202550.6%946grok-code-fast-1 23.4%Claude Sonnet 4.5 11.5%Gemini 2.5 Flash 6.5%MiniMax M2 (free) 5.8%grok-4-fast 3.5%
Oct 27, 202552.3%1109grok-code-fast-1 26.6%Claude Sonnet 4.5 11.9%Gemini 2.5 Flash 5.8%MiniMax M2 (free) 4.4%Gemini 2.5 Pro 3.7%
Oct 20, 202549.4%1052grok-code-fast-1 26.0%Claude Sonnet 4.5 10.7%Gemini 2.5 Flash 6.0%Gemini 2.5 Pro 3.5%grok-4-fast 3.2%
Oct 13, 202550.0%1032grok-code-fast-1 25.3%Claude Sonnet 4.5 10.9%Gemini 2.5 Flash 6.8%Claude Sonnet 4 3.5%grok-4-fast 3.5%
Oct 6, 202548.8%914grok-code-fast-1 23.6%Claude Sonnet 4.5 9.6%Gemini 2.5 Flash 6.7%Claude Sonnet 4 5.0%grok-4-fast 4.0%
Sep 29, 202549.7%809grok-code-fast-1 19.6%grok-4-fast (free) 12.0%Gemini 2.5 Flash 6.8%Claude Sonnet 4 5.9%Claude Sonnet 4.5 5.5%
Sep 22, 202559.2%1056grok-code-fast-1 20.3%grok-4-fast (free) 17.4%Claude Sonnet 4 11.0%Gemini 2.5 Flash 6.6%DeepSeek V3.1 (free) 4.0%
Sep 15, 202550.2%947grok-code-fast-1 23.3%Claude Sonnet 4 11.9%Gemini 2.5 Flash 6.6%sonoma-sky-alpha 4.6%gemini-2.0-flash-001 3.8%
Sep 8, 202549.2%953grok-code-fast-1 23.8%Claude Sonnet 4 11.6%Gemini 2.5 Flash 6.6%gemini-2.0-flash-001 3.8%sonoma-sky-alpha 3.5%
Sep 1, 202553.0%1002grok-code-fast-1 23.9%Claude Sonnet 4 11.8%Gemini 2.5 Flash 9.6%gemini-2.0-flash-001 4.0%DeepSeek V3 0324 3.6%
Aug 25, 202545.2%642Claude Sonnet 4 15.0%grok-code-fast-1 10.6%Gemini 2.5 Flash 9.2%gemini-2.0-flash-001 5.4%Gemini 2.5 Pro 4.9%
Aug 18, 202540.5%598Claude Sonnet 4 16.1%Gemini 2.5 Flash 7.9%gemini-2.0-flash-001 6.9%DeepSeek V3 0324 5.1%Gemini 2.5 Pro 4.5%
Aug 11, 202542.4%635Claude Sonnet 4 15.9%Gemini 2.5 Flash 8.4%gemini-2.0-flash-001 8.2%DeepSeek V3 0324 5.0%claude-3-7-sonnet 4.8%
Aug 4, 202541.4%599Claude Sonnet 4 15.2%gemini-2.0-flash-001 8.0%Gemini 2.5 Flash 7.6%horizon-beta 5.5%DeepSeek V3 0324 (free) 5.2%
Jul 28, 202545.4%693Claude Sonnet 4 17.7%Gemini 2.5 Flash 8.8%gemini-2.0-flash-001 8.0%DeepSeek V3 0324 (free) 6.2%Gemini 2.5 Pro 4.8%
Jul 21, 202551.2%795Claude Sonnet 4 19.1%Gemini 2.5 Flash 10.6%gemini-2.0-flash-001 8.7%DeepSeek V3 0324 (free) 6.9%Gemini 2.5 Pro 6.0%
Jul 7, 202550.6%744Claude Sonnet 4 14.8%gemini-2.0-flash-001 10.4%gemini-2.5-flash-preview-05-20 9.8%DeepSeek V3 0324 (free) 7.9%Gemini 2.5 Flash 7.7%
Jun 30, 202548.4%720Claude Sonnet 4 15.3%gemini-2.0-flash-001 11.2%gemini-2.5-flash-preview-05-20 8.3%Gemini 2.5 Pro 7.1%DeepSeek V3 0324 (free) 6.3%
Jun 23, 202545.7%663Claude Sonnet 4 14.1%gemini-2.0-flash-001 11.1%gemini-2.5-flash-preview-05-20 9.8%DeepSeek V3 0324 (free) 5.5%DeepSeek V3 0324 5.1%
Jun 16, 202547.3%704Claude Sonnet 4 15.9%gemini-2.0-flash-001 11.2%gemini-2.0-flash-lite-001 7.2%gemini-2.5-flash-preview-05-20 7.2%claude-3-7-sonnet 5.8%
Jun 2, 202547.6%677GPT-4o-mini 12.9%Claude Sonnet 4 10.3%gemini-2.0-flash-001 9.8%claude-3-7-sonnet 7.4%gemini-2.5-flash-preview-05-20 7.1%
May 26, 202555.0%857GPT-4o-mini 18.9%Claude Sonnet 4 10.9%gemini-2.0-flash-001 8.8%Gemini 2.5 Pro Preview 05-06 8.6%claude-3-7-sonnet 7.8%
May 19, 202556.3%934GPT-4o-mini 20.3%claude-3-7-sonnet 13.6%gemini-2.0-flash-001 8.9%Gemini 2.5 Pro Preview 05-06 7.5%gemini-2.5-flash-preview-04-17 5.9%
May 12, 202559.4%1001GPT-4o-mini 20.4%claude-3-7-sonnet 14.9%gemini-2.0-flash-001 9.7%gemini-2.5-flash-preview-04-17 7.7%Gemini 2.5 Pro Preview 05-06 6.5%
May 5, 202552.5%791GPT-4o-mini 15.0%claude-3-7-sonnet 14.1%gemini-2.0-flash-001 10.7%gemini-2.5-flash-preview-04-17 7.2%gemini-2.5-pro-exp-03-25 5.6%
Apr 28, 202546.6%729claude-3-7-sonnet 17.1%gemini-2.0-flash-001 11.7%gemini-2.5-pro-exp-03-25 6.2%GPT-4o-mini 5.9%gemini-2.5-flash-preview-04-17 5.8%
Apr 21, 202547.2%757claude-3-7-sonnet 19.1%gemini-2.0-flash-001 10.9%gemini-2.5-flash-preview-04-17 5.8%gemini-2.5-pro-exp-03-25 (free) 5.8%DeepSeek V3 0324 (free) 5.5%
Apr 14, 202552.1%897claude-3-7-sonnet 21.9%gemini-2.0-flash-001 12.0%gemini-2.5-pro-exp-03-25 (free) 8.5%DeepSeek V3 0324 (free) 5.2%Gemini 2.5 Pro Preview 05-06 4.5%
Apr 7, 202554.5%799claude-3-7-sonnet 16.7%gemini-2.0-flash-001 11.9%GPT-4o-mini 9.8%quasar-alpha 9.6%gemini-2.5-pro-exp-03-25 (free) 6.6%
Mar 31, 202556.3%840claude-3-7-sonnet 15.8%gemini-2.0-flash-001 13.9%GPT-4o-mini 12.2%gemini-2.5-pro-exp-03-25 (free) 8.3%DeepSeek V3 0324 (free) 6.1%
Mar 24, 202554.0%925claude-3-7-sonnet 20.3%gemini-2.0-flash-001 15.4%Llama 3.3 70B Instruct 6.6%gemini-2.5-pro-exp-03-25 (free) 6.4%DeepSeek V3 0324 (free) 5.3%
Mar 17, 202555.1%1017claude-3-7-sonnet 22.3%gemini-2.0-flash-001 16.7%Llama 3.3 70B Instruct 5.5%R1 (free) 5.4%claude-3-7-sonnet (thinking) 5.3%
Mar 10, 202558.0%1180claude-3-7-sonnet 23.5%gemini-2.0-flash-001 20.2%R1 (free) 5.3%claude-3-7-sonnet (thinking) 4.8%claude-3.5-sonnet 4.2%
Mar 3, 202557.2%1158gemini-2.0-flash-001 23.3%claude-3-7-sonnet 19.9%claude-3.5-sonnet 5.0%R1 (free) 4.6%claude-3-7-sonnet (thinking) 4.4%
Feb 24, 202559.9%1228gemini-2.0-flash-001 25.8%claude-3-7-sonnet 17.8%claude-3.5-sonnet 8.0%R1 (free) 4.5%claude-3.5-sonnet (beta) 3.8%
Feb 17, 202564.8%1446gemini-2.0-flash-001 28.1%claude-3.5-sonnet 18.6%claude-3.5-sonnet (beta) 11.6%gemini-flash-1.5 3.3%R1 (free) 3.3%
Feb 10, 202561.3%1275claude-3.5-sonnet 26.8%gemini-2.0-flash-001 14.1%claude-3.5-sonnet (beta) 12.7%Mistral Nemo 4.1%gemini-flash-1.5 3.7%
Feb 3, 202560.7%1314claude-3.5-sonnet 25.6%claude-3.5-sonnet (beta) 20.4%gemini-flash-1.5-8b 5.4%gemini-2.0-flash-001 5.2%gemini-flash-1.5 4.1%
Jan 27, 202562.7%1576claude-3.5-sonnet (beta) 26.2%claude-3.5-sonnet 25.9%gemini-flash-1.5 4.3%GPT-4o-mini 3.5%Mistral Nemo 2.8%
Jan 20, 202565.1%1349claude-3.5-sonnet (beta) 25.6%claude-3.5-sonnet 20.1%gemini-flash-1.5 8.8%gemini-flash-1.5-8b 7.6%GPT-4o-mini 3.0%
Jan 13, 202566.5%1408claude-3.5-sonnet (beta) 28.0%claude-3.5-sonnet 17.8%gemini-flash-1.5 9.0%gemini-flash-1.5-8b 8.6%DeepSeek V3 3.1%
Jan 6, 202566.1%1447claude-3.5-sonnet (beta) 30.2%claude-3.5-sonnet 15.4%gemini-flash-1.5 8.5%gemini-flash-1.5-8b 8.0%DeepSeek V3 4.1%

The top 5 models went from 66.1% to 49.8% of volume between the weeks of Jan 6, 2025 and Aug 31, 2026. The model-level HHI went from 1447 to 700, which keeps it “unconcentrated” on the conventional antitrust scale.

Concentration peaked within the window at 66.5%, in the week of Jan 13, 2025.

In the latest period, the largest is Hy4 preview, with 12.7% of volume on its own.

Routing implication: there are close substitutes at the top, so being locked into one specific model is a choice, not an inevitability.

03

Share by lab

Who is gaining and losing ground in traffic, and which models are at the top now. Color follows the lab or the origin, never the rank.

Weekly token share by lab

DeepSeek, OpenAI, Google and Anthropic in fixed colors, the same across the page; all others combined in Other, which the tooltip breaks down

0%20%40%60%80%100%Apr '25JulOctJan '26AprJul
See the numbers
WeekDeepSeekOpenAIGoogleAnthropicOther
Aug 31, 202618.0%14.6%5.7%4.5%57.2%
Aug 24, 202618.6%10.4%6.9%3.9%60.3%
Aug 17, 202621.8%9.5%6.8%4.9%57.1%
Aug 10, 202626.3%11.7%8.4%8.3%45.4%
Aug 3, 202625.6%11.1%8.8%6.7%47.8%
Jul 27, 202621.5%8.2%7.9%7.7%54.7%
Jul 20, 202617.1%5.5%7.8%8.7%60.9%
Jul 13, 202613.5%5.8%6.4%11.4%62.9%
Jul 6, 202616.0%5.7%7.4%13.3%57.5%
Jun 29, 202617.4%6.6%8.7%14.9%52.4%
Jun 22, 202615.9%6.5%8.3%13.6%55.8%
Jun 15, 202618.1%6.8%8.4%13.8%52.8%
Jun 8, 202616.9%5.5%9.3%15.0%53.3%
Jun 1, 202618.1%6.3%11.3%14.2%50.1%
May 25, 202617.0%7.5%12.3%17.1%46.2%
May 18, 202619.1%8.4%14.1%16.1%42.2%
May 11, 202615.2%7.6%13.8%15.0%48.4%
May 4, 202610.9%7.7%13.1%14.0%54.3%
Apr 27, 20268.5%8.2%13.5%13.6%56.3%
Apr 20, 20266.9%9.2%15.2%16.9%51.9%
Apr 13, 20266.7%10.1%16.6%17.0%49.6%
Apr 6, 20266.6%10.3%15.5%13.8%53.7%
Mar 30, 20265.0%6.3%10.2%9.9%68.6%
Mar 23, 20266.1%8.3%12.4%12.3%60.8%
Mar 16, 20266.3%8.2%14.4%13.7%57.4%
Mar 9, 20266.9%10.9%16.8%15.0%50.5%
Mar 2, 20266.5%12.3%18.5%15.6%47.1%
Feb 23, 20267.0%9.5%19.0%15.8%48.8%
Feb 16, 20266.5%9.4%16.2%13.2%54.7%
Feb 9, 20267.1%10.2%16.7%12.2%53.8%
Feb 2, 20268.6%11.3%21.0%14.7%44.5%
Jan 26, 20268.8%13.1%23.2%16.4%38.7%
Jan 19, 20268.0%11.7%23.0%16.5%40.8%
Jan 12, 20267.6%11.1%25.8%18.9%36.7%
Jan 5, 20268.4%9.4%23.8%16.6%41.8%
Dec 29, 202510.3%8.1%22.9%12.3%46.5%
Dec 22, 20259.7%8.8%22.0%12.6%47.0%
Dec 15, 20257.3%13.2%22.6%14.1%42.8%
Dec 8, 20256.3%14.5%22.6%15.0%41.6%
Dec 1, 20255.9%10.0%20.5%13.7%49.8%
Nov 24, 20254.4%8.3%17.6%10.2%59.5%
Nov 17, 20254.7%8.8%19.5%11.3%55.7%
Nov 10, 20255.2%8.0%18.0%14.3%54.5%
Nov 3, 20255.9%7.7%18.8%15.7%52.0%
Oct 27, 20255.2%8.2%19.2%16.4%50.9%
Oct 20, 20255.8%9.6%18.4%17.1%49.0%
Oct 13, 20256.7%10.4%18.4%16.8%47.7%
Oct 6, 20257.1%12.2%18.3%16.3%46.1%
Sep 29, 20259.4%11.1%18.2%13.1%48.1%
Sep 22, 20259.2%10.9%16.9%13.0%50.0%
Sep 15, 202510.8%10.5%17.9%14.4%46.4%
Sep 8, 202512.0%11.0%17.7%14.1%45.1%
Sep 1, 202510.8%10.0%21.0%14.7%43.5%
Aug 25, 202513.9%9.4%24.0%18.9%33.8%
Aug 18, 202516.6%10.1%23.9%22.1%27.4%
Aug 11, 202515.4%9.5%24.7%22.8%27.6%
Aug 4, 202515.1%7.9%23.0%21.1%32.8%
Jul 28, 202514.8%4.8%27.3%23.6%29.4%
Jul 21, 202517.4%5.5%30.4%25.3%21.3%
Jul 7, 202519.4%5.4%41.4%20.3%13.5%
Jun 30, 202517.9%6.3%39.1%21.4%15.4%
Jun 23, 202515.5%6.7%43.3%20.9%13.6%
Jun 16, 202516.2%6.2%40.8%24.2%12.6%
Jun 2, 202513.9%17.3%35.8%20.5%12.5%
May 26, 202511.4%23.2%31.8%21.7%12.0%
May 19, 202510.1%24.4%30.9%22.1%12.5%
May 12, 202510.3%24.7%31.8%19.6%13.6%
May 5, 202511.5%18.8%34.9%19.2%15.6%
Apr 28, 202512.9%9.7%36.0%23.4%18.0%
Apr 21, 202512.5%6.0%39.3%25.9%16.3%
Apr 14, 202511.7%7.3%35.7%29.1%16.2%
Apr 7, 202510.6%10.6%27.9%23.0%27.9%
Mar 31, 202513.9%12.9%29.8%24.0%19.5%
Mar 24, 202514.9%4.8%30.2%31.3%18.8%
Mar 17, 202510.0%4.1%28.6%39.3%18.0%
Mar 10, 20259.2%4.3%32.5%38.7%15.3%
Mar 3, 20258.3%3.7%37.6%34.5%15.9%
Feb 24, 20258.3%3.8%37.3%36.0%14.5%
Feb 17, 20257.8%5.0%39.9%31.5%15.8%
Feb 10, 20256.6%4.4%25.9%41.3%21.8%
Feb 3, 20255.6%4.7%17.4%48.7%23.6%
Jan 27, 20256.1%5.4%8.6%55.0%24.8%
Jan 20, 20254.2%5.1%18.7%48.0%24.0%
Jan 13, 20253.1%4.5%19.5%48.2%24.7%
Jan 6, 20254.1%4.2%18.2%48.1%25.4%

At the start of the window (week of Jan 6, 2025), the largest was Anthropic, with 48.1%. At the end (week of Aug 31, 2026), the leader is DeepSeek (18.0%), followed by Tencent (16.2%) and Z.ai (GLM) (15.4%).

The biggest gainer was Tencent, +16.2 pp (from 0.0% to 16.2%). The biggest loser was Anthropic, −43.6 pp (from 48.1% to 4.5%).

Labs with no meaningful share at the start of the window: Tencent, Z.ai (GLM), MiniMax, NVIDIA and Xiaomi. Today they add up to 48.3%.

Other adds up to 57.2% in the latest period. The largest inside it: Tencent (16.2%), Z.ai (GLM) (15.4%) and MiniMax (6.2%). Hover over the chart to see who was in Other in each period.

Within the window, the lead changed hands 17 times among 8 labs. DeepSeek has been on top for 6 straight weeks; Google had the longest streak, at 20 weeks.

Top 15 for the latest full week

Week of Aug 31, 2026. Share among named models; bar color = lab origin

The 15 highest-volume models in the latest period, with origin, license, share, tokens and rank 12 weeks ago
#ModelShare barShareTokens/wk12 wk ago
1Hy4 previewChinaOpen weights13.5%15Tdebuted later
2GPT-5.6 LunaUS/CanadaProprietary11.9%13Tdebuted later
3GLM 5.3 FlashChinaOpen weights11.4%12Tdebuted later
4DeepSeek V4 Flash 0731ChinaOpen weights11.4%12Tdebuted later
5DeepSeek V4 Flash 0423ChinaOpen weights4.8%5.2T#14
6MiniMax M3 (free)ChinaOpen weights4.6%5.0Tdebuted later
7Hy3ChinaOpen weights3.7%4.0Tdebuted later
8Nemotron 3 Ultra (free)US/CanadaOpen weights3.4%3.6T#124
9GLM 5.3ChinaOpen weights2.8%3.0Tdebuted later
10MiMo-V2.5ChinaOpen weights2.2%2.4T#46
11GLM 5.2ChinaOpen weights2.2%2.4Tdebuted later
12Gemini 3.7 FlashUS/CanadaProprietary1.9%2.1Tdebuted later
13Kimi K3ChinaOpen weights1.8%2.0Tdebuted later
14GPT-5.6 SolUS/CanadaProprietary1.7%1.9Tdebuted later
15Claude Opus 5US/CanadaProprietary1.6%1.8Tdebuted later
US/CanadaChina

Of the 10 most used in the latest week, 8 come from Chinese labs and 9 have open weights. US proprietary models: 1.

12 of the 15 were not in the top 15 in the week of Jun 8, 2026, 12 weeks earlier, and none of them existed yet. Each model's season at the top is short.

The most-used Anthropic model (Claude Opus 5) is at #15. The metric here is tokens, not revenue or quality: frontier models show up with smaller share and higher-value usage.

04

Origin, license and billing

Where the model comes from, whether its weights are open, how much of the traffic is free and where the providers serving inference are based.

Lab origin

% of weekly volume by headquarters of the lab that trained the model

0%20%40%60%80%100%Apr '25JulOctJan '26AprJul
See the numbers
WeekChinaEuropeKoreaOtherUnknownUS/Canada
Aug 31, 202661.3%0.1%1.1%0.0%6.0%31.6%
Aug 24, 202651.3%0.0%0.6%0.0%19.0%29.1%
Aug 17, 202650.8%0.2%0.6%0.0%18.1%30.3%
Aug 10, 202657.3%0.1%0.4%0.0%6.5%35.8%
Aug 3, 202659.5%0.1%0.0%0.0%6.1%34.3%
Jul 27, 202660.0%0.2%0.0%0.0%7.4%32.4%
Jul 20, 202663.5%0.4%0.0%0.0%7.1%29.0%
Jul 13, 202663.4%0.4%0.0%0.0%5.5%30.7%
Jul 6, 202660.6%0.5%0.0%0.0%6.0%32.8%
Jun 29, 202655.9%0.6%0.0%0.0%7.9%35.7%
Jun 22, 202653.7%0.5%0.0%0.0%13.2%32.5%
Jun 15, 202655.6%0.5%0.0%0.0%11.5%32.4%
Jun 8, 202653.5%0.4%0.0%0.0%12.2%34.0%
Jun 1, 202650.5%0.4%0.0%0.0%12.9%36.1%
May 25, 202644.9%0.5%0.0%0.0%13.6%41.1%
May 18, 202643.7%0.5%0.0%0.0%13.8%42.1%
May 11, 202644.8%0.5%0.0%0.0%13.6%41.1%
May 4, 202646.0%0.5%0.0%0.0%12.3%41.2%
Apr 27, 202647.5%0.5%0.0%0.0%10.9%41.1%
Apr 20, 202639.7%0.6%0.0%0.0%11.2%48.5%
Apr 13, 202636.1%0.6%0.0%0.0%12.4%50.8%
Apr 6, 202643.4%0.7%0.0%0.0%9.0%47.0%
Mar 30, 202661.8%0.4%0.0%0.0%5.7%32.1%
Mar 23, 202653.2%0.4%0.0%0.0%6.7%39.7%
Mar 16, 202643.1%0.4%0.0%0.0%13.0%43.5%
Mar 9, 202635.6%0.5%0.0%0.0%12.7%51.2%
Mar 2, 202636.2%0.5%0.0%0.0%7.5%55.8%
Feb 23, 202637.6%0.8%0.0%0.0%7.0%54.5%
Feb 16, 202644.7%0.8%0.0%0.0%5.9%48.6%
Feb 9, 202641.9%1.0%0.0%0.0%7.1%50.0%
Feb 2, 202631.1%2.2%0.0%0.0%7.6%59.1%
Jan 26, 202623.3%3.1%0.0%0.0%8.1%65.6%
Jan 19, 202622.4%4.5%0.0%0.0%8.0%65.1%
Jan 12, 202620.5%4.3%0.0%0.0%7.0%68.1%
Jan 5, 202623.7%5.0%0.0%0.0%8.7%62.6%
Dec 29, 202527.5%5.8%0.0%0.0%8.9%57.8%
Dec 22, 202526.5%5.9%0.0%0.0%8.8%58.9%
Dec 15, 202519.5%5.4%0.0%0.0%9.6%65.6%
Dec 8, 202516.6%5.0%0.0%0.0%9.0%69.4%
Dec 1, 202514.9%3.8%0.0%0.0%8.8%72.5%
Nov 24, 202513.2%2.5%0.0%0.1%7.3%76.9%
Nov 17, 202515.7%2.7%0.0%0.0%10.8%70.8%
Nov 10, 202516.5%3.0%0.0%0.0%10.9%69.6%
Nov 3, 202518.8%3.1%0.0%0.0%7.7%70.5%
Oct 27, 202515.7%3.2%0.0%0.0%6.3%74.9%
Oct 20, 202513.5%3.3%0.0%0.0%7.1%76.1%
Oct 13, 202512.9%3.4%0.0%0.0%7.0%76.7%
Oct 6, 202515.2%2.7%0.0%0.0%6.1%75.9%
Sep 29, 202514.3%1.4%0.0%0.0%6.1%78.2%
Sep 22, 202512.9%1.2%0.0%0.0%5.4%80.4%
Sep 15, 202515.7%1.9%0.0%0.0%11.3%71.1%
Sep 8, 202518.8%1.8%0.0%0.0%10.6%68.8%
Sep 1, 202518.6%2.5%0.0%0.0%7.0%72.0%
Aug 25, 202523.8%3.7%0.0%0.0%6.0%66.5%
Aug 18, 202529.7%3.8%0.0%0.0%6.3%60.1%
Aug 11, 202531.0%3.4%0.0%0.0%6.1%59.5%
Aug 4, 202531.1%2.8%0.0%0.0%11.1%55.0%
Jul 28, 202527.7%2.7%0.0%0.0%9.5%60.0%
Jul 21, 202525.6%3.6%0.0%0.0%4.8%66.0%
Jul 7, 202520.1%2.8%0.0%0.0%5.7%71.4%
Jun 30, 202518.6%2.7%0.0%0.1%6.1%72.5%
Jun 23, 202516.6%2.1%0.0%0.0%5.3%76.0%
Jun 16, 202516.8%2.0%0.0%0.0%5.5%75.6%
Jun 2, 202514.7%2.3%0.0%0.1%5.1%77.8%
May 26, 202512.3%2.1%0.0%0.1%4.7%80.8%
May 19, 202511.7%2.1%0.0%0.2%4.8%81.3%
May 12, 202512.1%1.7%0.0%0.2%5.4%80.6%
May 5, 202513.5%1.8%0.0%0.2%5.6%79.0%
Apr 28, 202514.9%2.4%0.0%0.3%6.0%76.4%
Apr 21, 202513.8%2.4%0.0%0.3%5.1%78.4%
Apr 14, 202512.8%2.5%0.0%0.3%6.3%78.2%
Apr 7, 202511.7%1.8%0.0%0.3%18.1%68.1%
Mar 31, 202515.3%2.2%0.0%0.3%7.1%75.0%
Mar 24, 202516.2%3.0%0.0%0.4%3.5%76.8%
Mar 17, 202511.8%2.9%0.0%0.5%3.6%81.2%
Mar 10, 202510.9%2.9%0.0%0.5%3.4%82.3%
Mar 3, 202510.4%3.3%0.0%0.6%3.5%82.2%
Feb 24, 20259.7%3.6%0.0%0.7%3.5%82.6%
Feb 17, 20259.6%3.2%0.0%1.2%4.1%81.9%
Feb 10, 20258.0%7.2%0.0%1.1%4.8%79.0%
Feb 3, 20257.8%7.0%0.0%1.5%4.8%79.0%
Jan 27, 20258.0%4.2%0.0%1.9%4.5%81.4%
Jan 20, 20255.8%4.0%0.0%3.4%3.8%83.1%
Jan 13, 20254.4%4.0%0.0%3.6%3.6%84.4%
Jan 6, 20256.1%4.1%0.0%3.4%3.6%82.8%

Chinese labs went from 6.1% to 61.3% of volume; US and Canadian labs, from 82.8% to 31.6%, between the weeks of Jan 6, 2025 and Aug 31, 2026. The gap between the two is now 29.6 points, in favor of China.

Outside the two poles, the largest slice is Korea, with 1.1% (+1.1 pp over the window).

This is the headquarters of whoever trained the model. Where the providers serving it are based is in the providers card, further down.

Weights license

% of weekly volume in open-weights, proprietary and unknown models

0%20%40%60%80%100%Apr '25JulOctJan '26AprJul
See the numbers
WeekOpen weightsUnknownProprietary
Aug 31, 202669.2%6.0%24.9%
Aug 24, 202660.1%19.0%20.9%
Aug 17, 202661.0%18.1%20.9%
Aug 10, 202665.8%6.5%27.8%
Aug 3, 202668.3%6.1%25.6%
Jul 27, 202669.7%7.4%22.9%
Jul 20, 202671.2%7.1%21.7%
Jul 13, 202671.0%5.5%23.5%
Jul 6, 202668.0%6.0%26.0%
Jun 29, 202661.7%7.9%30.5%
Jun 22, 202658.1%13.2%28.6%
Jun 15, 202659.9%11.5%28.7%
Jun 8, 202658.4%12.2%29.4%
Jun 1, 202655.9%12.9%31.2%
May 25, 202649.6%13.6%36.8%
May 18, 202648.7%13.8%37.5%
May 11, 202649.7%13.6%36.7%
May 4, 202651.4%12.3%36.4%
Apr 27, 202652.4%10.9%36.7%
Apr 20, 202643.0%11.2%45.8%
Apr 13, 202635.9%12.4%51.7%
Apr 6, 202638.4%9.0%52.6%
Mar 30, 202630.0%5.7%64.3%
Mar 23, 202639.1%6.7%54.1%
Mar 16, 202641.6%13.0%45.4%
Mar 9, 202643.2%12.7%44.1%
Mar 2, 202644.8%7.5%47.7%
Feb 23, 202646.2%7.0%46.8%
Feb 16, 202652.2%5.9%41.9%
Feb 9, 202649.9%7.1%43.0%
Feb 2, 202640.6%7.6%51.8%
Jan 26, 202632.6%8.1%59.3%
Jan 19, 202632.2%8.0%59.9%
Jan 12, 202629.6%7.0%63.4%
Jan 5, 202632.3%8.7%58.9%
Dec 29, 202537.7%8.9%53.4%
Dec 22, 202536.9%8.8%54.3%
Dec 15, 202531.0%9.6%59.4%
Dec 8, 202529.9%9.0%61.1%
Dec 1, 202522.9%8.8%68.3%
Nov 24, 202518.0%7.3%74.7%
Nov 17, 202520.8%10.8%68.4%
Nov 10, 202522.6%10.9%66.5%
Nov 3, 202526.0%7.7%66.3%
Oct 27, 202523.7%6.3%70.0%
Oct 20, 202521.4%7.1%71.5%
Oct 13, 202521.1%7.0%71.9%
Oct 6, 202525.1%6.1%68.8%
Sep 29, 202522.2%6.1%71.7%
Sep 22, 202519.3%5.4%75.3%
Sep 15, 202521.3%11.3%67.5%
Sep 8, 202525.7%10.6%63.7%
Sep 1, 202524.7%7.0%68.3%
Aug 25, 202531.4%6.0%62.5%
Aug 18, 202537.5%6.3%56.1%
Aug 11, 202538.0%6.1%55.8%
Aug 4, 202537.3%11.1%51.5%
Jul 28, 202534.1%9.5%56.4%
Jul 21, 202532.4%4.8%62.8%
Jul 7, 202526.8%5.7%67.5%
Jun 30, 202526.2%6.1%67.7%
Jun 23, 202522.8%5.3%71.8%
Jun 16, 202522.4%5.5%72.1%
Jun 2, 202521.2%5.1%73.6%
May 26, 202518.7%4.7%76.6%
May 19, 202517.4%4.8%77.8%
May 12, 202518.1%5.4%76.5%
May 5, 202521.3%5.6%73.2%
Apr 28, 202524.5%6.0%69.4%
Apr 21, 202523.5%5.1%71.4%
Apr 14, 202521.3%6.3%72.4%
Apr 7, 202520.8%18.1%61.1%
Mar 31, 202526.5%7.1%66.3%
Mar 24, 202529.9%3.5%66.5%
Mar 17, 202524.4%3.6%72.0%
Mar 10, 202520.8%3.4%75.8%
Mar 3, 202519.8%3.5%76.7%
Feb 24, 202518.7%3.5%77.8%
Feb 17, 202518.8%4.1%77.1%
Feb 10, 202522.5%4.8%72.8%
Feb 3, 202523.2%4.8%72.1%
Jan 27, 202523.4%4.5%72.1%
Jan 20, 202522.8%3.8%73.5%
Jan 13, 202522.1%3.6%74.2%
Jan 6, 202523.2%3.6%73.2%

Open-weights models went from 23.2% to 69.2% of volume; proprietary models, from 73.2% to 24.9%. 6.0% of the latest period has no identified license, and ignoring that band would inflate the other two.

At the end of the window, open weights and Chinese labs are 7.9 points apart (69.2% vs. 61.3%); at the start, the gap was 17.1. The curves converged over the window and now move almost together, because most relevant open-weights models are Chinese: it is the same phenomenon seen from two angles, not two independent trends.

Traffic on free endpoints

% of volume in models served by a no-charge endpoint, marked with the :free suffix in the slug. It is the same model as the paid endpoint, not a different model.

0%10%20%30%40%Apr '25JulOctJan '26AprJul
See the numbers
WeekFree share
Aug 31, 202611.3%
Aug 24, 202610.0%
Aug 17, 20268.5%
Aug 10, 20266.7%
Aug 3, 20268.1%
Jul 27, 20269.8%
Jul 20, 20269.5%
Jul 13, 202624.9%
Jul 6, 202617.6%
Jun 29, 20265.1%
Jun 22, 20264.4%
Jun 15, 20265.1%
Jun 8, 20265.6%
Jun 1, 20264.5%
May 25, 20264.9%
May 18, 20263.9%
May 11, 20265.5%
May 4, 202615.3%
Apr 27, 202618.3%
Apr 20, 20267.6%
Apr 13, 20263.9%
Apr 6, 202614.6%
Mar 30, 202630.0%
Mar 23, 20269.8%
Mar 16, 202611.4%
Mar 9, 202612.3%
Mar 2, 20269.0%
Feb 23, 20267.7%
Feb 16, 20266.5%
Feb 9, 20266.2%
Feb 2, 20265.9%
Jan 26, 20264.4%
Jan 19, 202611.3%
Jan 12, 202610.6%
Jan 5, 202611.1%
Dec 29, 202513.1%
Dec 22, 202514.6%
Dec 15, 20259.8%
Dec 8, 20256.2%
Dec 1, 202512.7%
Nov 24, 202524.3%
Nov 17, 20254.5%
Nov 10, 20253.5%
Nov 3, 20257.9%
Oct 27, 20256.5%
Oct 20, 20253.1%
Oct 13, 20253.3%
Oct 6, 20253.8%
Sep 29, 202517.6%
Sep 22, 202522.3%
Sep 15, 202510.3%
Sep 8, 20258.7%
Sep 1, 20257.6%
Aug 25, 20258.4%
Aug 18, 202512.2%
Aug 11, 202514.7%
Aug 4, 202512.9%
Jul 28, 202515.6%
Jul 21, 202515.2%
Jul 7, 202512.5%
Jun 30, 202510.3%
Jun 23, 20259.3%
Jun 16, 202510.0%
Jun 2, 20259.1%
May 26, 20257.7%
May 19, 20257.1%
May 12, 20257.1%
May 5, 20258.2%
Apr 28, 202510.5%
Apr 21, 202516.9%
Apr 14, 202518.7%
Apr 7, 202515.6%
Mar 31, 202520.5%
Mar 24, 202520.4%
Mar 17, 202511.9%
Mar 10, 202511.7%
Mar 3, 202511.8%
Feb 24, 202510.7%
Feb 17, 202510.2%
Feb 10, 20258.1%
Feb 3, 20254.2%
Jan 27, 20252.0%
Jan 20, 20251.8%
Jan 13, 20251.3%
Jan 6, 20250.9%

Traffic on free endpoints went from 0.9% to 11.3% of volume over the window, peaking at 30.0% in the week of Mar 30, 2026.

Free volume inflates adoption without signaling willingness to pay. Every share on this page includes that subsidized demand, which has not yet been tested against price.

Where inference providers are based new

Snapshot from Sep 11, 2026. Does not respond to the window or filters

106 inference providers listed, by declared headquarters country

United States55 · 52%China6 · 6%Singapore6 · 6%Israel2 · 2%5 other countries5 · 5%Headquarters not reported32 · 30%

Zero data retention (ZDR)

847
endpoints with zero retention
305
models with at least one of those endpoints
50
of 106 providers offer the option

This is the headquarters of whoever serves the model, the inference provider that receives the request. That is a different thing from the origin chart, which shows the headquarters of whoever trained it: a Chinese model served by a US provider counts as China there and as the United States here.

55 of the 106 providers (52%) declare headquarters in the United States. 32 declare no headquarters at all, and for anyone with jurisdiction constraints this is the group that needs case-by-case checking.

Zero retention is a per-endpoint policy, not a per-model one: 50 of the 106 providers offer it on at least one endpoint, and the same model can have endpoints with and without it.

05

What workload the traffic demands

Share of each week's volume by an attribute of the model that served the traffic: maximum context window, input modality and declared reasoning. It responds to the window, the grouping and the filters.

Context window new

% of volume by the model's maximum context window

Up to 32k33k to 128k129k to 400kOver 400kUnknown
0%20%40%60%80%100%Apr '25JulOctJan '26AprJul
Over 400k rose from 0.2% to 82.2% of volume between Jan 6, 2025 and Aug 31, 2026; within identified volume alone, from 0.8% to 87.6%. 33k to 128k fell from 9.5% to 0.1%. Unknown at the end of the window: 6.1%.
See the numbers
WeekUnknownOver 400k129k to 400k33k to 128kUp to 32k
Aug 31, 20266.1%82.2%11.6%0.1%0.0%
Aug 24, 202619.3%65.9%14.7%0.2%0.0%
Aug 17, 202618.5%62.3%19.0%0.3%0.0%
Aug 10, 20267.1%70.3%22.4%0.3%0.0%
Aug 3, 20266.8%70.1%22.8%0.3%0.0%
Jul 27, 20268.6%66.4%24.6%0.4%0.0%
Jul 20, 20268.1%66.6%24.9%0.4%0.0%
Jul 13, 20266.8%61.8%31.0%0.5%0.0%
Jul 6, 20267.6%64.6%27.3%0.6%0.0%
Jun 29, 202610.1%67.4%21.5%0.9%0.0%
Jun 22, 202615.0%62.0%22.2%0.8%0.0%
Jun 15, 202614.3%60.7%24.1%1.0%0.0%
Jun 8, 202615.4%58.1%25.5%0.9%0.0%
Jun 1, 202615.4%56.0%27.5%1.1%0.0%
May 25, 202616.1%53.0%29.4%1.5%0.0%
May 18, 202614.9%48.9%34.8%1.4%0.0%
May 11, 202618.1%42.5%38.5%0.9%0.0%
May 4, 202618.1%36.9%44.0%0.9%0.0%
Apr 27, 202617.7%34.0%47.4%0.9%0.0%
Apr 20, 202621.3%35.0%42.7%1.1%0.0%
Apr 13, 202624.1%35.8%39.0%1.1%0.0%
Apr 6, 202617.6%39.2%40.6%2.6%0.0%
Mar 30, 202630.0%37.6%31.1%1.3%0.0%
Mar 23, 202632.5%24.8%41.8%0.9%0.0%
Mar 16, 202628.8%27.6%42.6%1.0%0.0%
Mar 9, 202623.0%30.8%45.0%1.3%0.0%
Mar 2, 202619.7%31.5%47.7%1.1%0.0%
Feb 23, 202620.4%30.7%47.9%1.0%0.0%
Feb 16, 202618.8%25.0%55.4%0.8%0.0%
Feb 9, 202621.9%23.2%53.8%1.0%0.0%
Feb 2, 202625.6%26.9%46.0%1.5%0.0%
Jan 26, 202628.9%29.5%40.1%1.5%0.0%
Jan 19, 202636.6%28.9%32.7%1.8%0.0%
Jan 12, 202634.1%29.2%34.9%1.8%0.0%
Jan 5, 202637.0%27.6%33.6%1.8%0.0%
Dec 29, 202539.9%24.7%33.6%1.8%0.0%
Dec 22, 202541.1%24.1%33.2%1.5%0.0%
Dec 15, 202539.7%24.5%34.2%1.7%0.0%
Dec 8, 202539.8%25.0%33.6%1.6%0.0%
Dec 1, 202549.9%21.6%26.2%1.8%0.6%
Nov 24, 202559.0%18.5%21.1%1.3%0.1%
Nov 17, 202553.0%22.6%23.0%1.3%0.1%
Nov 10, 202548.5%25.7%24.4%1.5%0.0%
Nov 3, 202543.5%27.6%27.4%1.4%0.1%
Oct 27, 202545.0%28.3%25.1%1.4%0.1%
Oct 20, 202546.3%27.5%23.6%2.3%0.2%
Oct 13, 202547.5%28.4%22.2%1.8%0.2%
Oct 6, 202543.5%28.8%26.1%1.5%0.2%
Sep 29, 202547.2%27.6%23.9%1.3%0.0%
Sep 22, 202550.9%24.9%22.7%1.4%0.1%
Sep 15, 202546.5%27.8%23.8%1.8%0.1%
Sep 8, 202543.1%27.7%27.3%1.9%0.0%
Sep 1, 202540.7%30.9%26.1%2.1%0.1%
Aug 25, 202531.4%33.7%32.2%2.5%0.2%
Aug 18, 202524.8%33.5%38.5%3.1%0.2%
Aug 11, 202524.7%33.2%39.1%2.8%0.2%
Aug 4, 202527.9%31.7%37.4%2.6%0.4%
Jul 28, 202528.7%37.2%30.7%3.0%0.4%
Jul 21, 202525.4%41.4%29.2%3.3%0.7%
Jul 7, 202538.2%33.5%24.3%3.5%0.4%
Jun 30, 202538.8%32.4%22.7%5.4%0.7%
Jun 23, 202545.4%28.2%20.1%5.5%0.8%
Jun 16, 202546.6%27.9%19.5%5.3%0.6%
Jun 2, 202541.8%23.8%18.1%15.5%0.8%
May 26, 202538.6%23.3%15.5%21.8%0.9%
May 19, 202546.5%15.9%13.0%23.6%1.0%
May 12, 202551.0%10.9%13.2%23.7%1.3%
May 5, 202555.5%8.2%15.3%18.6%2.4%
Apr 28, 202562.0%8.6%17.7%10.0%1.7%
Apr 21, 202565.9%10.1%16.4%6.1%1.6%
Apr 14, 202567.3%10.1%14.8%6.4%1.4%
Apr 7, 202566.1%4.5%13.8%13.6%2.0%
Mar 31, 202561.0%0.9%17.9%17.4%2.8%
Mar 24, 202566.1%0.1%19.8%11.3%2.6%
Mar 17, 202573.0%0.3%12.0%11.7%3.1%
Mar 10, 202576.5%0.3%9.1%11.5%2.7%
Mar 3, 202577.9%0.3%8.7%9.9%3.3%
Feb 24, 202578.6%0.6%8.2%10.2%2.4%
Feb 17, 202577.0%0.8%9.4%9.9%2.9%
Feb 10, 202575.4%0.5%11.7%9.3%3.1%
Feb 3, 202575.0%0.8%11.3%8.9%4.1%
Jan 27, 202571.8%1.1%12.1%10.4%4.6%
Jan 20, 202573.8%1.1%12.2%9.3%3.7%
Jan 13, 202575.5%0.6%11.6%8.9%3.4%
Jan 6, 202574.3%0.2%11.5%9.5%4.5%

Input modality new

% of volume in models that accept text only or also image, audio or file

Text onlyMultimodalUnknown
0%20%40%60%80%100%Apr '25JulOctJan '26AprJul
Multimodal rose from 4.4% to 46.6% of volume between Jan 6, 2025 and Aug 31, 2026; within identified volume alone, from 17.2% to 49.6%. Text only rose from 21.3% to 47.3%. Unknown at the end of the window: 6.1%.
See the numbers
WeekUnknownMultimodalText only
Aug 31, 20266.1%46.6%47.3%
Aug 24, 202619.3%40.6%40.1%
Aug 17, 202618.5%36.0%45.5%
Aug 10, 20267.1%38.8%54.1%
Aug 3, 20266.8%40.9%52.3%
Jul 27, 20268.6%43.8%47.6%
Jul 20, 20268.1%50.3%41.6%
Jul 13, 20266.8%46.7%46.5%
Jul 6, 20267.6%48.7%43.7%
Jun 29, 202610.1%52.4%37.5%
Jun 22, 202615.0%50.1%34.9%
Jun 15, 202614.3%48.6%37.1%
Jun 8, 202615.4%49.6%34.9%
Jun 1, 202615.4%47.0%37.7%
May 25, 202616.1%43.0%40.9%
May 18, 202614.9%41.6%43.5%
May 11, 202618.1%41.3%40.6%
May 4, 202618.1%41.9%40.0%
Apr 27, 202617.7%43.8%38.5%
Apr 20, 202621.3%49.1%29.6%
Apr 13, 202624.1%44.6%31.3%
Apr 6, 202617.6%49.1%33.3%
Mar 30, 202630.0%43.8%26.1%
Mar 23, 202632.5%32.5%35.0%
Mar 16, 202628.8%35.8%35.4%
Mar 9, 202623.0%42.5%34.5%
Mar 2, 202619.7%45.7%34.6%
Feb 23, 202620.4%43.7%35.9%
Feb 16, 202618.8%40.9%40.3%
Feb 9, 202621.9%42.6%35.4%
Feb 2, 202625.6%51.2%23.2%
Jan 26, 202628.9%47.7%23.4%
Jan 19, 202636.6%43.3%20.1%
Jan 12, 202634.1%47.7%18.2%
Jan 5, 202637.0%42.6%20.3%
Dec 29, 202539.9%37.0%23.1%
Dec 22, 202541.1%37.4%21.5%
Dec 15, 202539.7%40.4%19.9%
Dec 8, 202539.8%38.0%22.2%
Dec 1, 202549.9%32.9%17.2%
Nov 24, 202559.0%27.0%13.9%
Nov 17, 202553.0%29.7%17.3%
Nov 10, 202548.5%33.1%18.4%
Nov 3, 202543.5%35.4%21.2%
Oct 27, 202545.0%35.9%19.1%
Oct 20, 202546.3%37.0%16.6%
Oct 13, 202547.5%35.6%17.0%
Oct 6, 202543.5%35.9%20.6%
Sep 29, 202547.2%33.9%18.9%
Sep 22, 202550.9%32.5%16.6%
Sep 15, 202546.5%35.2%18.3%
Sep 8, 202543.1%34.1%22.8%
Sep 1, 202540.7%37.7%21.6%
Aug 25, 202531.4%40.6%28.0%
Aug 18, 202524.8%40.5%34.6%
Aug 11, 202524.7%39.5%35.8%
Aug 4, 202527.9%36.6%35.5%
Jul 28, 202528.7%40.9%30.4%
Jul 21, 202525.4%45.5%29.1%
Jul 7, 202538.2%37.5%24.3%
Jun 30, 202538.8%37.6%23.5%
Jun 23, 202545.4%33.7%20.9%
Jun 16, 202546.6%32.7%20.8%
Jun 2, 202541.8%38.7%19.5%
May 26, 202538.6%44.5%16.9%
May 19, 202546.5%37.9%15.7%
May 12, 202551.0%32.7%16.3%
May 5, 202555.5%25.1%19.4%
Apr 28, 202562.0%16.1%21.9%
Apr 21, 202565.9%13.8%20.4%
Apr 14, 202567.3%14.3%18.4%
Apr 7, 202566.1%15.9%17.9%
Mar 31, 202561.0%14.5%24.5%
Mar 24, 202566.1%5.6%28.2%
Mar 17, 202573.0%5.1%21.9%
Mar 10, 202576.5%4.8%18.8%
Mar 3, 202577.9%4.1%18.0%
Feb 24, 202578.6%4.5%16.9%
Feb 17, 202577.0%5.5%17.5%
Feb 10, 202575.4%5.2%19.4%
Feb 3, 202575.0%5.5%19.4%
Jan 27, 202571.8%7.0%21.2%
Jan 20, 202573.8%6.4%19.9%
Jan 13, 202575.5%5.2%19.3%
Jan 6, 202574.3%4.4%21.3%

Reasoning new

% of volume in models that declare reasoning support in the catalog

No reasoningReasoningUnknown
0%20%40%60%80%100%Apr '25JulOctJan '26AprJul
Reasoning rose from 0.0% to 93.7% of volume between Jan 6, 2025 and Aug 31, 2026; within identified volume alone, from 0.0% to 99.8%. No reasoning fell from 25.7% to 0.2%. Unknown at the end of the window: 6.1%. In identified volume, 99.8% already runs on models that declare reasoning: the attribute has saturated and no longer separates one kind of traffic from another.
See the numbers
WeekUnknownReasoningNo reasoning
Aug 31, 20266.1%93.7%0.2%
Aug 24, 202619.3%80.5%0.2%
Aug 17, 202618.5%81.1%0.5%
Aug 10, 20267.1%92.5%0.4%
Aug 3, 20266.8%92.8%0.4%
Jul 27, 20268.6%90.8%0.6%
Jul 20, 20268.1%91.0%0.9%
Jul 13, 20266.8%92.3%0.9%
Jul 6, 20267.6%91.3%1.1%
Jun 29, 202610.1%88.2%1.7%
Jun 22, 202615.0%83.4%1.6%
Jun 15, 202614.3%84.1%1.7%
Jun 8, 202615.4%82.8%1.8%
Jun 1, 202615.4%81.8%2.8%
May 25, 202616.1%81.2%2.7%
May 18, 202614.9%82.2%2.8%
May 11, 202618.1%79.3%2.7%
May 4, 202618.1%79.5%2.4%
Apr 27, 202617.7%79.8%2.5%
Apr 20, 202621.3%76.0%2.7%
Apr 13, 202624.1%72.5%3.4%
Apr 6, 202617.6%77.2%5.2%
Mar 30, 202630.0%67.2%2.8%
Mar 23, 202632.5%64.4%3.2%
Mar 16, 202628.8%67.7%3.4%
Mar 9, 202623.0%72.4%4.6%
Mar 2, 202619.7%75.4%4.9%
Feb 23, 202620.4%74.8%4.8%
Feb 16, 202618.8%76.5%4.7%
Feb 9, 202621.9%72.8%5.2%
Feb 2, 202625.6%67.1%7.2%
Jan 26, 202628.9%62.4%8.7%
Jan 19, 202636.6%52.1%11.3%
Jan 12, 202634.1%55.6%10.3%
Jan 5, 202637.0%52.1%10.8%
Dec 29, 202539.9%47.2%12.8%
Dec 22, 202541.1%46.5%12.3%
Dec 15, 202539.7%48.3%12.1%
Dec 8, 202539.8%48.2%12.0%
Dec 1, 202549.9%39.6%10.5%
Nov 24, 202559.0%32.1%8.9%
Nov 17, 202553.0%35.6%11.3%
Nov 10, 202548.5%41.5%10.1%
Nov 3, 202543.5%44.8%11.7%
Oct 27, 202545.0%42.0%13.0%
Oct 20, 202546.3%39.7%13.9%
Oct 13, 202547.5%39.9%12.7%
Oct 6, 202543.5%40.6%15.9%
Sep 29, 202547.2%37.9%14.9%
Sep 22, 202550.9%36.0%13.1%
Sep 15, 202546.5%37.8%15.6%
Sep 8, 202543.1%39.9%17.0%
Sep 1, 202540.7%41.4%17.9%
Aug 25, 202531.4%49.6%19.1%
Aug 18, 202524.8%48.7%26.5%
Aug 11, 202524.7%45.5%29.8%
Aug 4, 202527.9%44.3%27.7%
Jul 28, 202528.7%39.8%31.4%
Jul 21, 202525.4%43.0%31.6%
Jul 7, 202538.2%34.8%26.9%
Jun 30, 202538.8%33.7%27.5%
Jun 23, 202545.4%29.9%24.7%
Jun 16, 202546.6%30.2%23.2%
Jun 2, 202541.8%25.0%33.3%
May 26, 202538.6%24.3%37.2%
May 19, 202546.5%15.5%38.0%
May 12, 202551.0%10.0%39.0%
May 5, 202555.5%8.7%35.8%
Apr 28, 202562.0%9.2%28.8%
Apr 21, 202565.9%9.0%25.1%
Apr 14, 202567.3%7.5%25.1%
Apr 7, 202566.1%6.2%27.6%
Mar 31, 202561.0%5.5%33.5%
Mar 24, 202566.1%6.3%27.6%
Mar 17, 202573.0%7.7%19.4%
Mar 10, 202576.5%7.1%16.5%
Mar 3, 202577.9%6.2%15.9%
Feb 24, 202578.6%6.3%15.1%
Feb 17, 202577.0%5.8%17.2%
Feb 10, 202575.4%4.9%19.6%
Feb 3, 202575.0%3.7%21.3%
Jan 27, 202571.8%4.0%24.2%
Jan 20, 202573.8%1.9%24.3%
Jan 13, 202575.5%0.0%24.5%
Jan 6, 202574.3%0.0%25.7%

What is measured here. Declared capability, counted model by model in the catalog, saturates: when nearly every model advertises tool or reasoning support, counting models separates nothing. What still moves is the workload, measured by volume: how much traffic goes to models with long context windows, multimodal input or reasoning. The gray band is the same in all three charts: it is volume in models with no catalog record, 74.3% (Jan 6, 2025) and 6.1% (Aug 31, 2026). When it shrinks, the others grow for that reason alone, which is why the reading also compares identified volume only. The catalog says whether a model supports reasoning, not whether reasoning is mandatory, so the third card measures support weighted by volume. And all three describe the model, not the request: traffic on a model with a window over 400k does not mean prompts that long.

06

Where models are used

Which apps the traffic runs through, and what an agent session costs. Both are snapshots from the source, which keeps no history of them, and do not respond to the window or the History filters.

Top apps of the day new

Tokens processed per app on Sep 10, 2026, the last day published by the source

Trending the source's criterion

1Hermes Agent1.64T
2Codex0.216T
3DeepSeek Harness0.168T
4Open WebUI0.116T
5pi0.287T
6Kilo Code0.412T
7omp0.222T
8pre.dev-agentoutside the top 150.038T

Ranked by recent growth, as defined by the source; the number is the day's volume.

Hermes Agent leads in the overall ranking, with 1.64T on the day, 28.8% of the volume of the top 100 apps (5.68T). The top three add up to 49.9%. Per request, the heaviest workload on the list is Hello Minds, powered by Ethoswarm, at 208K tokens per call; the lightest is Open WebUI, at 1.9K.

Only apps that identify themselves to the source are included; direct API calls are left out. An app can appear in more than one category.

Median cost per session new

Source median per model, aggregated by harness and turn range; window of 30 days through Sep 6, 2026

$0.001$0.01$0.1$1$101 turn2 to 910 to 4950 or more0.410.00030.910.00234.160.02526.530.170.0190.0470.322.51
See the numbers
HarnessTurnsModelsMedianMinimumMaximum
Claude Code1 turn28$0.019$0.0003$0.41
Claude Code2 to 9 turns38$0.047$0.0023$0.91
Claude Code10 to 49 turns35$0.32$0.025$4.16
Claude Code50 turns or more27$2.51$0.17$26.53
Codex1 turn18$0.0081$0.0005$0.11
Codex2 to 9 turns11$0.054$0.0025$0.29
Codex10 to 49 turns11$0.35$0.017$2.81
Codex50 turns or more9$1.10$0.029$5.31
Hermes Agent1 turn88$0.0054$0.0001$0.31
Hermes Agent2 to 9 turns111$0.015$0.0005$0.81
Hermes Agent10 to 49 turns97$0.13$0.01$4.22
Hermes Agent50 turns or more59$1.04$0.06$40.98
Kilo Code1 turn15$0.004$0.0003$0.062
Kilo Code2 to 9 turns19$0.04$0.0054$0.53
Kilo Code10 to 49 turns20$0.24$0.019$2.22
Kilo Code50 turns or more14$1.07$0.1$13.24

Log scale. The light band spans the cheapest to the priciest model in the selected harness; the bold number is the median across models.

Claude Code, 10 to 49 turns: 35 models, median $0.32, from $0.025 to $4.16

Cheapest

MiMo-V2.5Xiaomi$0.025
GPT-5.6 LunaOpenAI$0.037
Qwen3.7 FlashQwen$0.038

Priciest

Claude Fable 5Anthropic$4.16
Claude Opus 4.8Anthropic$1.96
Claude Opus 5Anthropic$1.80
Claude Opus 4.6Anthropic$1.58
Claude Opus 4.7Anthropic$1.32

At 50 turns or more, the median runs from $1.04 on Hermes Agent to $2.51 on Claude Code. Within Claude Code, in the same range, cost runs from $0.17 to $26.53 depending on the model, a 153-fold spread, versus 2.4-fold across harnesses: the choice of model weighs more than the choice of tool.

A gap between harnesses does not prove savings. Each tool serves different tasks, codebases and users, and the median across models is not weighted by each one's volume.

07

Where the money goes

Tokens multiplied by each model's list price. An estimate, not a measurement: the source does not split prompt from completion, and the price is today's.

Estimated weekly spend

Millions of dollars per week, at list prices. The band runs from "all prompt" to "all completion"; the line assumes 75% prompt.

EstimateFloor-to-ceiling range
0M100M200M300MApr '25JulOctJan '26AprJul
At list prices, traffic in the last week of the window (Aug 31, 2026) comes to $117M per week, within a range of $65M to $275M. In the first week (Jan 6, 2025) it was $0.06M. The effective price went from $0.11 to $1.01 per million tokens. Much of the rise at the start of the series is coverage, not spend: 74.3% of the initial volume came from models with no catalog price, which count as zero, versus 6.1% at the end. Traffic on free endpoints (11.3% at the end) also costs zero.
Today's prices applied to the past. The catalog keeps only the current price, so the whole history is recalculated with today's price list. A model that got cheaper has its past spend overstated, and one that got pricier, understated. Correcting this requires a price history, which does not exist yet.
See the numbers
WeekEstimate ($M)Floor ($M)Ceiling ($M)$ per 1M tokens
Aug 31, 2026117.1864.55275.071.01
Aug 24, 202688.1147.67209.450.78
Aug 17, 202685.5845.31206.380.92
Aug 10, 202698.8551.87239.771.31
Aug 3, 202676.9940.67185.951.11
Jul 27, 202668.3736.47164.071.20
Jul 20, 202680.3642.72193.311.39
Jul 13, 2026100.0052.12243.661.59
Jul 6, 202694.9949.37231.831.80
Jun 29, 202694.9648.78233.502.03
Jun 22, 202687.9445.09216.501.88
Jun 15, 202688.7245.60218.091.90
Jun 8, 202682.1242.07202.271.84
Jun 1, 202662.9932.21155.331.74
May 25, 202664.1732.63158.802.02
May 18, 202656.3928.35140.521.95
May 11, 202649.7325.06123.761.85
May 4, 202644.8422.61111.531.74
Apr 27, 202639.6719.9398.871.67
Apr 20, 202642.5421.29106.281.94
Apr 13, 202638.2319.0195.911.87
Apr 6, 202633.8316.8384.831.61
Mar 30, 202629.6414.8374.061.10
Mar 23, 202631.1015.5677.721.37
Mar 16, 202630.2315.1375.521.49
Mar 9, 202626.3612.9366.651.56
Mar 2, 202624.6312.0562.361.67
Feb 23, 202622.4811.0856.691.65
Feb 16, 202620.7910.3752.061.49
Feb 9, 202618.369.1346.051.41
Feb 2, 202614.847.2537.581.51
Jan 26, 202613.106.3433.391.59
Jan 19, 202611.365.5128.921.52
Jan 12, 202614.206.8136.351.86
Jan 5, 20269.924.8225.221.55
Dec 29, 20256.393.1016.241.15
Dec 22, 20256.613.2016.861.17
Dec 15, 20257.723.6719.881.32
Dec 8, 20257.613.6719.421.32
Dec 1, 20257.363.5618.771.19
Nov 24, 20256.363.0316.330.89
Nov 17, 20255.792.7614.900.92
Nov 10, 20256.533.1316.751.10
Nov 3, 20256.823.2817.421.17
Oct 27, 20256.663.2017.041.21
Oct 20, 20255.872.8315.021.20
Oct 13, 20255.732.7514.681.23
Oct 6, 20256.072.9315.501.21
Sep 29, 20255.502.6514.051.02
Sep 22, 20255.822.7714.961.11
Sep 15, 20256.052.9015.501.23
Sep 8, 20255.762.7714.741.18
Sep 1, 20255.802.7814.861.26
Aug 25, 20255.872.8115.031.57
Aug 18, 20255.512.6714.011.72
Aug 11, 20255.642.7414.351.76
Aug 4, 20255.972.9115.131.74
Jul 28, 20256.163.0115.601.81
Jul 21, 20255.862.8514.871.95
Jul 7, 20253.781.829.631.59
Jun 30, 20253.361.638.561.59
Jun 23, 20253.411.668.671.45
Jun 16, 20253.561.739.041.58
Jun 2, 20253.221.558.251.36
May 26, 20253.381.638.641.35
May 19, 20251.930.924.950.81
May 12, 20250.900.422.340.42
May 5, 20250.730.371.810.34
Apr 28, 20250.550.271.390.30
Apr 21, 20250.590.281.530.34
Apr 14, 20250.590.291.480.33
Apr 7, 20250.420.211.050.19
Mar 31, 20250.230.140.490.11
Mar 24, 20250.150.100.290.09
Mar 17, 20250.120.090.240.09
Mar 10, 20250.120.080.220.08
Mar 3, 20250.110.070.210.08
Feb 24, 20250.100.070.200.09
Feb 17, 20250.110.070.230.12
Feb 10, 20250.090.060.180.11
Feb 3, 20250.090.060.170.12
Jan 27, 20250.080.050.150.14
Jan 20, 20250.070.050.130.13
Jan 13, 20250.060.040.110.11
Jan 6, 20250.060.040.100.11

Volume vs. money, by lab

Token share and estimated spend share in the last week of the window (Aug 31, 2026), for labs with at least 0.5% of tokens. Equal bars mean the market's average price.

Token shareEstimated spend share
0%10%20%30%40%ratioDeepSeek0.20×Tencent0.96×Z.ai (GLM)0.67×OpenAI1.03×MiniMax0.10×Google1.01×Anthropic7.42×NVIDIA0.00×Xiaomi0.19×Moonshot AI4.29×
Anthropic has 4.8% of tokens and 35.7% of the money, a ratio of 7.42×. At the other end, MiniMax has 6.6% of tokens and 0.7% of the money, a ratio of 0.10×. NVIDIA and Poolside show zero spend despite the volume: all of the traffic is on free endpoints or on models with no catalog price. Token share favors whoever charges less; for a contract decision the ratio says more, because it approximates where each lab is priced within real usage.
See the numbers
LabToken shareSpend shareRatio
DeepSeek19.1%3.9%0.20×
Tencent17.2%16.4%0.96×
Z.ai (GLM)16.4%11.0%0.67×
OpenAI15.6%16.1%1.03×
MiniMax6.6%0.7%0.10×
Google6.0%6.1%1.01×
Anthropic4.8%35.7%7.42×
NVIDIA4.6%0.0%0.00×
Xiaomi2.4%0.5%0.19×
Moonshot AI1.9%8.0%4.29×

Spend share by lab new

% of each period's estimated spend, four labs in fixed colors and the rest in Other

DeepSeekOpenAIGoogleAnthropicOther
0%20%40%60%80%100%Apr '25JulOctJan '26AprJul
In the week of Jan 6, 2025, the largest share of spend belonged to OpenAI (39.8%). In the week of Aug 31, 2026, it belongs to Anthropic, with 35.7% of the money and 4.8% of tokens.
See the numbers
WeekDeepSeekOpenAIGoogleAnthropicOther↳ Tencent↳ Z.ai (GLM)↳ Moonshot AI↳ xAI
Aug 31, 20263.9%16.1%6.1%35.7%38.3%16.4%11.0%8.0%0.8%
Aug 24, 20265.8%15.9%10.3%38.4%29.6%6.1%9.9%8.4%1.5%
Aug 17, 20265.9%19.0%8.5%43.7%22.9%2.2%7.2%7.5%1.9%
Aug 10, 20265.7%13.8%7.0%54.1%19.5%2.3%6.5%6.7%1.5%
Aug 3, 20266.1%14.1%8.7%48.3%22.8%2.4%6.7%8.7%1.3%
Jul 27, 20267.3%12.4%6.4%49.4%24.4%1.6%6.4%10.1%1.5%
Jul 20, 20265.7%12.2%6.0%53.0%23.1%1.1%6.3%8.2%2.1%
Jul 13, 20263.9%13.4%4.3%64.3%14.1%0.0%5.8%2.5%1.0%
Jul 6, 20264.3%13.4%4.3%65.8%12.3%0.1%5.5%1.1%0.7%
Jun 29, 20263.7%16.3%4.9%63.7%11.4%0.9%4.5%0.9%0.2%
Jun 22, 20263.6%17.1%5.1%62.6%11.6%1.1%4.2%1.2%0.1%
Jun 15, 20264.4%16.7%5.0%62.8%11.1%1.2%4.0%1.6%0.0%
Jun 8, 20263.9%10.0%6.2%71.0%9.0%1.4%1.2%1.2%0.0%
Jun 1, 20264.4%12.2%7.8%66.9%8.7%1.3%1.1%1.1%0.0%
May 25, 20263.4%12.4%7.7%69.2%7.4%1.3%1.2%1.7%0.1%
May 18, 20263.5%15.0%9.2%65.1%7.2%1.6%1.3%2.4%0.1%
May 11, 20263.3%14.8%8.2%63.7%9.9%1.5%1.7%4.1%0.0%
May 4, 20263.0%15.1%7.8%62.0%12.0%0.5%2.2%6.7%0.0%
Apr 27, 20262.1%13.9%8.9%61.1%14.0%0.0%2.7%8.4%0.0%
Apr 20, 20261.1%10.1%8.7%67.2%13.0%0.0%2.9%7.0%0.0%
Apr 13, 20261.1%11.1%10.7%68.0%9.0%0.0%4.3%0.9%0.0%
Apr 6, 20261.3%11.9%13.0%63.1%10.7%0.0%4.7%1.4%0.1%
Mar 30, 20261.4%10.2%11.2%65.5%11.7%0.0%5.7%1.8%0.0%
Mar 23, 20261.4%11.4%10.5%63.9%12.8%0.0%7.2%1.5%0.0%
Mar 16, 20261.3%9.6%11.2%65.4%12.4%0.0%7.5%1.7%0.0%
Mar 9, 20261.4%11.0%12.8%67.3%7.5%0.0%1.4%1.9%0.0%
Mar 2, 20261.3%10.6%12.3%66.8%9.2%0.0%1.4%2.7%0.0%
Feb 23, 20261.4%8.2%12.3%67.7%10.4%0.0%2.5%2.8%0.0%
Feb 16, 20261.4%7.4%9.9%65.7%15.6%0.0%4.1%4.5%0.0%
Feb 9, 20261.6%8.6%9.0%64.8%16.0%0.0%3.8%6.5%0.0%
Feb 2, 20261.9%9.9%10.8%67.0%10.4%0.0%0.9%7.3%0.0%
Jan 26, 20261.9%12.2%11.6%68.7%5.7%0.0%1.0%3.0%0.0%
Jan 19, 20261.8%9.5%12.4%73.6%2.6%0.0%0.9%0.4%0.0%
Jan 12, 20261.4%7.3%15.6%73.7%1.9%0.0%0.7%0.3%0.0%
Jan 5, 20261.9%9.0%11.5%74.4%3.2%0.0%1.4%0.5%0.0%
Dec 29, 20253.1%8.8%15.4%68.3%4.4%0.0%1.7%0.6%0.0%
Dec 22, 20252.9%10.6%14.5%68.3%3.7%0.0%1.3%0.5%0.0%
Dec 15, 20252.0%14.9%12.5%67.6%3.0%0.0%1.0%0.5%0.0%
Dec 8, 20251.8%10.6%12.5%71.8%3.3%0.0%1.0%0.3%0.0%
Dec 1, 20252.0%8.4%12.6%74.0%3.1%0.0%1.2%0.1%0.0%
Nov 24, 20252.2%10.5%13.6%69.4%4.3%0.0%1.4%0.3%0.0%
Nov 17, 20252.3%12.1%15.2%65.4%5.1%0.0%1.5%0.6%0.0%
Nov 10, 20252.1%8.5%15.0%69.6%4.8%0.0%1.3%0.7%0.0%
Nov 3, 20252.2%6.8%14.7%72.7%3.6%0.0%1.5%0.5%0.0%
Oct 27, 20252.0%6.7%15.0%73.5%2.9%0.0%1.4%0.0%0.0%
Oct 20, 20252.1%7.3%14.8%72.4%3.4%0.0%1.3%0.3%0.0%
Oct 13, 20252.3%8.2%14.2%72.4%2.8%0.0%1.1%0.4%0.0%
Oct 6, 20252.0%8.4%13.3%73.2%3.1%0.0%1.6%0.2%0.0%
Sep 29, 20252.1%9.7%14.8%70.3%3.1%0.0%1.6%0.4%0.0%
Sep 22, 20252.0%12.6%13.5%69.6%2.3%0.0%0.8%0.5%0.0%
Sep 15, 20251.9%12.4%12.7%70.7%2.2%0.0%0.8%0.5%0.0%
Sep 8, 20252.0%10.3%13.8%70.7%3.1%0.0%1.3%0.6%0.0%
Sep 1, 20252.0%10.2%15.5%69.2%3.0%0.0%1.1%0.5%0.0%
Aug 25, 20252.6%8.9%15.9%69.8%2.8%0.0%0.7%0.4%0.0%
Aug 18, 20252.3%7.4%13.1%73.1%4.0%0.0%1.0%1.3%0.0%
Aug 11, 20251.6%6.9%12.6%74.4%4.5%0.0%1.0%1.7%0.0%
Aug 4, 20251.6%5.7%12.4%75.7%4.6%0.0%1.1%1.5%0.0%
Jul 28, 20251.4%3.9%13.3%78.5%2.8%0.0%0.6%0.5%0.0%
Jul 21, 20251.7%4.6%15.2%76.5%2.0%0.0%0.0%0.8%0.0%
Jul 7, 20252.3%5.5%18.4%72.6%1.1%0.0%0.0%0.2%0.0%
Jun 30, 20252.5%5.9%18.4%72.1%1.1%0.0%0.0%0.0%0.0%
Jun 23, 20252.2%6.6%16.5%73.6%1.1%0.0%0.0%0.0%0.0%
Jun 16, 20252.2%6.3%16.4%74.4%0.7%0.0%0.0%0.0%0.0%
Jun 2, 20252.1%9.8%22.7%64.3%1.1%0.0%0.0%0.0%0.0%
May 26, 20251.8%11.2%21.8%64.2%1.0%0.0%0.0%0.0%0.0%
May 19, 20252.5%19.1%31.8%44.6%2.0%0.0%0.0%0.0%0.0%
May 12, 20255.3%36.6%53.9%0.0%4.2%0.0%0.0%0.0%0.0%
May 5, 20258.7%40.0%44.3%0.1%6.9%0.0%0.0%0.0%0.0%
Apr 28, 20258.4%35.6%48.3%0.1%7.7%0.0%0.0%0.0%0.0%
Apr 21, 20256.5%32.3%55.3%0.3%5.5%0.0%0.0%0.0%0.0%
Apr 14, 20256.5%40.4%47.4%0.3%5.3%0.0%0.0%0.0%0.0%
Apr 7, 202510.8%26.9%52.8%0.3%9.2%0.0%0.0%0.0%0.0%
Mar 31, 202522.0%44.1%13.7%0.1%20.2%0.0%0.0%0.0%0.0%
Mar 24, 202529.2%39.2%0.6%0.0%30.9%0.0%0.0%0.0%0.0%
Mar 17, 202528.3%38.2%0.6%0.3%32.7%0.0%0.0%0.0%0.0%
Mar 10, 202525.8%42.3%0.2%0.5%31.2%0.0%0.0%0.0%0.0%
Mar 3, 202521.0%47.8%0.0%0.8%30.3%0.0%0.0%0.0%0.0%
Feb 24, 202520.9%49.7%0.0%1.0%28.5%0.0%0.0%0.0%0.0%
Feb 17, 202520.2%55.1%0.0%0.8%23.9%0.0%0.0%0.0%0.0%
Feb 10, 202521.6%49.9%0.0%1.1%27.5%0.0%0.0%0.0%0.0%
Feb 3, 202523.3%44.7%0.0%1.1%30.9%0.0%0.0%0.0%0.0%
Jan 27, 202529.4%38.0%0.0%1.5%31.2%0.0%0.0%0.0%0.0%
Jan 20, 202524.5%35.8%0.0%1.5%38.3%0.0%0.0%0.0%0.0%
Jan 13, 202513.0%41.4%1.0%1.8%42.9%0.0%0.0%0.0%0.0%
Jan 6, 202516.6%39.8%1.3%1.3%40.9%0.0%0.0%0.0%0.0%
08

Quality vs. adoption

Capability measured by third parties against what traffic actually picks. The first card responds to the window and the filters; the second is a snapshot from the evaluation date.

Intelligence Index vs. token share

Artificial Analysis Intelligence Index, as exposed by the router's API. 37 of 61 models with volume in the last week of the window (Aug 31, 2026) have an index score; the rest are left out.

US/CanadaChina
0%2%4%6%8%10%12%14%1020304050Intelligence Index →share in the last week →index below median, heavily usedindex above median, heavily usedbelow median, lightly usedabove median, lightly usedgpt-5.6-lunaglm-5.3-flashdeepseek-v4-flash Jul 31deepseek-v4-flash Apr 23minimax-m3 (free)

Color: lab origin. Bubble size: tokens in the period. Dashed lines: medians of the models in the chart. Click a point to open the model's page.

Of the 61 models with volume in the last week, 37 have an index score. The median index score is 34.3 and the median share is 0.90%. The clearest case of quality without adoption is anthropic/claude-fable-5.1-20260831: index 53.4, share 0.26%. The reverse, adoption without a high index: deepseek/deepseek-v4-flash-20260423, index 24.8 and share 4.5%. Benchmarks say what a model can do; share says what people chose to run. When they diverge, Price mode helps show whether the gap is cost.
See the numbers
ModelOriginShareTokens/wkIntelligenceElo$/1MContextLaunch
tencent/hy4-preview-20260827China12.71%15T$1.251MAug 28, 2026
openai/gpt-5.6-luna-20260709US/Canada11.21%13T37.5$0.451MJul 9, 2026
z-ai/glm-5.3-flash-20260826China10.73%12T41.91309$0.2371.31MAug 26, 2026
deepseek/deepseek-v4-flash-20260731China10.71%12T34.51239$0.0941.31MJul 31, 2026
deepseek/deepseek-v4-flash-20260423China4.49%5.2T24.81200$0.1111MApr 24, 2026
minimax/minimax-m3-20260531:freeChina4.34%5.0T29.61246$0.5251MMay 31, 2026
tencent/hy3-20260706China3.46%4.0T1186$0.231256kJul 6, 2026
nvidia/nemotron-3-ultra-550b-a55b-20260604:freeUS/Canada3.15%3.6T23.41153$1.25256kJun 4, 2026
z-ai/glm-5.3-20260816China2.62%3.0T44.91332$2.151.31MAug 18, 2026
xiaomi/mimo-v2.5-20260422China2.04%2.4T22.31270$0.1751MApr 22, 2026
z-ai/glm-5.2-20260616China2.03%2.3T1308$1.481MJun 16, 2026
google/gemini-3.7-flash-20260813US/Canada1.82%2.1T39.41321$1.501MAug 13, 2026
moonshotai/kimi-k3-20260715China1.72%2.0T43.81369$4.681MJul 16, 2026
openai/gpt-5.6-sol-20260709US/Canada1.63%1.9T47.1$4.001MJul 9, 2026
anthropic/claude-opus-5-20260723US/Canada1.54%1.8T50.71359$10.001MJul 24, 2026
minimax/minimax-m3-20260531China1.25%1.4T29.61246$0.5251MMay 31, 2026
anthropic/claude-sonnet-5-20260630US/Canada1.15%1.3T38.41290$4.001MJun 30, 2026
poolside/laguna-s-2.1-20260720:freeUS/Canada1.15%1.3T$0.1131MJul 21, 2026
deepseek/deepseek-v4-pro-20260423China1.08%1.3T30.91242$1.191MApr 24, 2026
upstage/solar-pro4-20260810Korea1.05%1.2T1196$0.158512kAug 10, 2026
anthropic/claude-4.6-sonnet-20260217US/Canada1.00%1.2T30.51286$6.001MFeb 17, 2026
nvidia/nemotron-3.5-lightning-20260807:freeUS/Canada0.93%1.1T13.6$0.11256kAug 11, 2026
google/gemini-3.8-flash-20260902US/Canada0.93%1.1T41.21322$1.501MSep 2, 2026
deepseek/deepseek-v4-pro-20260813China0.90%1.0T36.3$0.991MAug 12, 2026
google/gemini-3-flash-preview-20251217US/Canada0.75%0.87T1206$1.131MDec 17, 2025
meta/muse-spark-1.3-contributor-20260902US/Canada0.72%0.83T$0.1251MSep 2, 2026
inclusionai/ling-3.0-flash-fin-20260827:freeChina0.71%0.82T$0.09256kAug 27, 2026
google/gemini-2.5-flash-liteUS/Canada0.61%0.70T$0.1751MJul 22, 2025
minimax/minimax-m2.7-20260318:freeChina0.59%0.69T23.21230$0.525200kMar 18, 2026
openai/gpt-5.6-terra-20260709US/Canada0.49%0.57T42.3$4.501MJul 9, 2026
google/gemini-2.5-flashUS/Canada0.44%0.51T1108$0.851MJun 17, 2025
openai/gpt-oss-120bUS/Canada0.40%0.46T12.3979$0.07128kAug 5, 2025
stepfun/step-3.7-flash-20260528China0.40%0.46T1184$0.438256kMay 28, 2026
deepseek/deepseek-v4-flash-vision-exp-20260821China0.39%0.46T$0.331MAug 21, 2026
deepseek/deepseek-v3.2-20251201China0.39%0.45T1164$0.302160kDec 1, 2025
google/gemini-3.1-flash-lite-20260507US/Canada0.39%0.45T$0.5621MMay 7, 2026
openai/gpt-oss-20bUS/Canada0.36%0.41T9906$0.055128kAug 5, 2025
anthropic/claude-4.8-opus-20260528US/Canada0.35%0.41T421266$10.001MMay 27, 2026
openai/gpt-5.6-luna-pro-20260709US/Canada0.31%0.35T$0.451MJul 9, 2026
google/gemma-4-26b-a4b-it-20260403US/Canada0.31%0.35T$0.086256kApr 3, 2026
google/gemma-4-31b-it-20260402US/Canada0.30%0.35T15.4$0.152256kApr 2, 2026
x-ai/grok-4.6-20260810US/Canada0.29%0.33T44.41308$3.00488kAug 12, 2026
anthropic/claude-fable-5.1-20260831US/Canada0.26%0.29T53.41348$20.001MSep 1, 2026
nvidia/nemotron-3-super-120b-a12b-20230311:freeUS/Canada0.25%0.29T13.6$0.164256kMar 11, 2026
xiaomi/mimo-v2.5-pro-20260422China0.19%0.22T26.41280$0.5441MApr 22, 2026
qwen/qwen3.8-max-20260803China0.17%0.20T
anthropic/claude-4.5-haiku-20251001US/Canada0.17%0.19T17.61126$2.00195kOct 15, 2025
thinkingmachines/inkling-20260715:freeUS/Canada0.16%0.18T25.51190$1.761MJul 17, 2026
qwen/qwen3.8-27b-20260814China0.15%0.17T33.9$1.061MAug 14, 2026
meta/muse-spark-1.2-contributor-20260805US/Canada0.12%0.14T$0.1251MAug 21, 2026
openai/gpt-4o-miniUS/Canada0.12%0.13T$0.262125kJul 18, 2024
openai/gpt-6-astra-20260903US/Canada0.12%0.13T52.8$20.001MSep 4, 2026
google/gemini-3.1-pro-preview-20260219US/Canada0.08%0.09T30.41265$4.501MFeb 19, 2026
anthropic/claude-5-fable-20260609US/Canada0.06%0.07T49.71332$20.001MJun 9, 2026
qwen/qwen3.8-max-20260902China0.06%0.07T40.3$3.001MSep 3, 2026
qwen/qwen3.8-flash-20260826China0.06%0.07T$0.231MAug 26, 2026
google/gemini-3.6-flash-20260721US/Canada0.06%0.07T34.31305$1.501MJul 21, 2026
mistralai/mistral-nemoEurope0.06%0.07T$0.022128kJul 19, 2024
inclusionai/ling-3.0-flash-sante-20260904:freeChina0.06%0.07T$0256kSep 4, 2026
meta/muse-spark-1.3-20260902US/Canada0.04%0.04T1364$2.001MSep 2, 2026
moonshotai/kimi-k2.6-20260420China0.03%0.04T1279$1.71256kApr 20, 2026

Accuracy vs. cost per task new

Evaluations published by the router on Sep 11, 2026. A snapshot of that day: it does not respond to the window or the filters. 130 models with accuracy and cost on GPQA Diamond, from 156 to 5346 tasks per model.

On the efficiency frontierOff the frontier
0%20%40%60%80%100%$0.0001$0.001$0.01$0.1$1average cost per task, $, log scale →accuracy →Mistral NemoGemini 3.1 Pro PreviewGLM 5.3 FlashDeepSeek V4 Flash 0423mimo-v2-flash
On GPQA Diamond, 15 of the 130 models form the frontier: no other model is both cheaper and more accurate than they are. It runs from Mistral Nemo (31.6%, $0.0001 per task) to Gemini 3.1 Pro Preview (94.4%, $0.202): 62.9 more points of accuracy cost 2019 times as much per task. The priciest model off the frontier is Claude Opus 4.5, at $0.842 per task with 86.6%; DeepSeek V4 Flash 0423 scores 86.6% for $0.0036.
Models on the efficiency frontier, from cheapest to most accurate
On the frontierLabAccuracyCost per taskTasks
Mistral NemoMistral AI31.6%$0.00011548
Qwen2.5 7B InstructQwen32.8%$0.0005198
GPT-4.1 NanoOpenAI50.9%$0.0008396
Qwen3 30B A3B Instruct 2507Qwen63.9%$0.00144902
Llama 4 MaverickMeta65.9%$0.0022733
Ling 3.0 FlashInclusionAI73.0%$0.003196
mimo-v2-flashXiaomi81.1%$0.00332970
DeepSeek V4 Flash 0423DeepSeek86.6%$0.00363366
GLM 5.3 FlashZ.ai (GLM)86.7%$0.007396
GPT-5.6 LunaOpenAI87.7%$0.0074396
DeepSeek V4 Flash Vision ExpDeepSeek88.2%$0.023396
MiniMax M3MiniMax90.5%$0.0302172
Gemini 3.7 FlashGoogle94.3%$0.030198
GPT-6 AstraOpenAI94.4%$0.099346
Gemini 3.1 Pro PreviewGoogle94.4%$0.202198

Left out of the selector for having too few models: BrowseComp (4), DSQA (4), HLE with search (2), WideSearch (4).

09

The Big Three: Anthropic, OpenAI and Google

The same breakdown for all three: share against competitors, volume by model family and time in the top 10. Together they hold 24.8% of volume in the week of Aug 31, 2026.

Pick a lab

Anthropic share · Aug 31, 2026
4.5% 43.6 pp in the window
was 48.1% in the week of Jan 6, 2025; peak of 55.0% in the week of Jan 27, 2025
Volume per week
5.2T↑ ×22 in the window
was 0.24T; the market grew 231-fold over the same period
Family that grew most
Sonnet
+2.3T per week in the window, from 0.23T to 2.5T

Share vs. competitors

% of weekly volume; Anthropic highlighted against its three largest competitors and the rest of the Big Three

0%20%40%60%May '25SepJan '26MayDeepSeek 18.0%Tencent 16.2%Z.ai (GLM) 15.4%OpenAI 14.6%Google 5.7%Anthropic 4.5%
See the numbers
WeekAnthropicDeepSeekTencentZ.ai (GLM)OpenAIGoogle
Aug 31, 20264.5%18.0%16.2%15.4%14.6%5.7%
Aug 24, 20263.9%18.6%8.6%9.3%10.4%6.9%
Aug 17, 20264.9%21.8%8.8%4.2%9.5%6.8%
Aug 10, 20268.3%26.3%13.2%5.8%11.7%8.4%
Aug 3, 20266.7%25.6%11.7%5.1%11.1%8.8%
Jul 27, 20267.7%21.5%8.5%5.3%8.2%7.9%
Jul 20, 20268.7%17.1%9.8%6.0%5.5%7.8%
Jul 13, 202611.4%13.5%18.4%6.3%5.8%6.4%
Jul 6, 202613.3%16.0%12.4%6.9%5.7%7.4%
Jun 29, 202614.9%17.4%6.7%6.3%6.6%8.7%
Jun 22, 202613.6%15.9%7.2%5.5%6.5%8.3%
Jun 15, 202613.8%18.1%7.8%5.3%6.8%8.4%
Jun 8, 202615.0%16.9%9.3%1.7%5.5%9.3%
Jun 1, 202614.2%18.1%8.1%1.9%6.3%11.3%
May 25, 202617.1%17.0%9.5%2.7%7.5%12.3%
May 18, 202616.1%19.1%10.6%2.6%8.4%14.1%
May 11, 202615.0%15.2%9.9%3.0%7.6%13.8%
May 4, 202614.0%10.9%13.7%3.2%7.7%13.1%
Apr 27, 202613.6%8.5%12.7%3.5%8.2%13.5%
Apr 20, 202616.9%6.9%1.5%4.5%9.2%15.2%
Apr 13, 202617.0%6.7%0.0%6.7%10.1%16.6%
Apr 6, 202613.8%6.6%0.0%6.1%10.3%15.5%
Mar 30, 20269.9%5.0%0.0%4.3%6.3%10.2%
Mar 23, 202612.3%6.1%0.0%6.3%8.3%12.4%
Mar 16, 202613.7%6.3%0.0%7.0%8.2%14.4%
Mar 9, 202615.0%6.9%0.0%2.8%10.9%16.8%
Mar 2, 202615.6%6.5%0.0%3.1%12.3%18.5%
Feb 23, 202615.8%7.0%0.0%5.4%9.5%19.0%
Feb 16, 202613.2%6.5%0.0%7.4%9.4%16.2%
Feb 9, 202612.2%7.1%0.0%6.7%10.2%16.7%
Feb 2, 202614.7%8.6%0.0%2.5%11.3%21.0%
Jan 26, 202616.4%8.8%0.0%2.7%13.1%23.2%
Jan 19, 202616.5%8.0%0.0%2.1%11.7%23.0%
Jan 12, 202618.9%7.6%0.0%2.1%11.1%25.8%
Jan 5, 202616.6%8.4%0.0%3.0%9.4%23.8%
Dec 29, 202512.3%10.3%0.0%2.6%8.1%22.9%
Dec 22, 202512.6%9.7%0.0%2.0%8.8%22.0%
Dec 15, 202514.1%7.3%0.0%1.8%13.2%22.6%
Dec 8, 202515.0%6.3%0.0%2.1%14.5%22.6%
Dec 1, 202513.7%5.9%0.0%2.1%10.0%20.5%
Nov 24, 202510.2%4.4%0.0%1.9%8.3%17.6%
Nov 17, 202511.3%4.7%0.0%2.4%8.8%19.5%
Nov 10, 202514.3%5.2%0.0%2.5%8.0%18.0%
Nov 3, 202515.7%5.9%0.0%2.9%7.7%18.8%
Oct 27, 202516.4%5.2%0.0%2.8%8.2%19.2%
Oct 20, 202517.1%5.8%0.0%2.9%9.6%18.4%
Oct 13, 202516.8%6.7%0.0%2.6%10.4%18.4%
Oct 6, 202516.3%7.1%0.0%3.1%12.2%18.3%
Sep 29, 202513.1%9.4%0.0%2.4%11.1%18.2%
Sep 22, 202513.0%9.2%0.0%1.2%10.9%16.9%
Sep 15, 202514.4%10.8%0.0%1.4%10.5%17.9%
Sep 8, 202514.1%12.0%0.0%1.9%11.0%17.7%
Sep 1, 202514.7%10.8%0.0%1.9%10.0%21.0%
Aug 25, 202518.9%13.9%0.0%2.0%9.4%24.0%
Aug 18, 202522.1%16.6%0.0%2.8%10.1%23.9%
Aug 11, 202522.8%15.4%0.0%2.9%9.5%24.7%
Aug 4, 202521.1%15.1%0.0%3.1%7.9%23.0%
Jul 28, 202523.6%14.8%0.0%1.5%4.8%27.3%
Jul 21, 202525.3%17.4%0.0%0.0%5.5%30.4%
Jul 7, 202520.3%19.4%0.0%0.0%5.4%41.4%
Jun 30, 202521.4%17.9%0.0%0.0%6.3%39.1%
Jun 23, 202520.9%15.5%0.0%0.0%6.7%43.3%
Jun 16, 202524.2%16.2%0.0%0.0%6.2%40.8%
Jun 2, 202520.5%13.9%0.0%0.0%17.3%35.8%
May 26, 202521.7%11.4%0.0%0.0%23.2%31.8%
May 19, 202522.1%10.1%0.0%0.0%24.4%30.9%
May 12, 202519.6%10.3%0.0%0.0%24.7%31.8%
May 5, 202519.2%11.5%0.0%0.0%18.8%34.9%
Apr 28, 202523.4%12.9%0.0%0.0%9.7%36.0%
Apr 21, 202525.9%12.5%0.0%0.0%6.0%39.3%
Apr 14, 202529.1%11.7%0.0%0.0%7.3%35.7%
Apr 7, 202523.0%10.6%0.0%0.0%10.6%27.9%
Mar 31, 202524.0%13.9%0.0%0.0%12.9%29.8%
Mar 24, 202531.3%14.9%0.0%0.0%4.8%30.2%
Mar 17, 202539.3%10.0%0.0%0.0%4.1%28.6%
Mar 10, 202538.7%9.2%0.0%0.0%4.3%32.5%
Mar 3, 202534.5%8.3%0.0%0.0%3.7%37.6%
Feb 24, 202536.0%8.3%0.0%0.0%3.8%37.3%
Feb 17, 202531.5%7.8%0.0%0.0%5.0%39.9%
Feb 10, 202541.3%6.6%0.0%0.0%4.4%25.9%
Feb 3, 202548.7%5.6%0.0%0.0%4.7%17.4%
Jan 27, 202555.0%6.1%0.0%0.0%5.4%8.6%
Jan 20, 202548.0%4.2%0.0%0.0%5.1%18.7%
Jan 13, 202548.2%3.1%0.0%0.0%4.5%19.5%
Jan 6, 202548.1%4.1%0.0%0.0%4.2%18.2%

Anthropic's share fell from 48.1% in the week of Jan 6, 2025 to 4.5% in the week of Aug 31, 2026, but absolute volume rose from 0.24T to 5.2T per week (×22). The market as a whole expanded 231× over the same period: share fell because the denominator grew faster, not because usage shrank. The peak in the window was 55.0%, in the week of Jan 27, 2025. In the last period, it is the 7th largest lab in the filtered view.

Volume by model family

Trillions of tokens per week, by family: Opus, Sonnet, Haiku and Fable

0.00T2.0T4.0T6.0T8.0TApr '25JulOctJan '26AprJul
See the numbers
WeekOpusSonnetHaikuFableTotal
Aug 31, 20262.2T2.5T0.19T0.37T5.2T
Aug 24, 20262.0T2.0T0.24T0.18T4.4T
Aug 17, 20262.3T1.8T0.23T0.27T4.6T
Aug 10, 20263.9T1.8T0.24T0.27T6.2T
Aug 3, 20262.3T1.9T0.24T0.23T4.7T
Jul 27, 20261.9T2.0T0.24T0.22T4.4T
Jul 20, 20262.7T1.8T0.23T0.33T5.1T
Jul 13, 20264.5T2.0T0.25T0.47T7.2T
Jul 6, 20264.5T1.9T0.24T0.37T7.0T
Jun 29, 20264.5T2.0T0.24T0.18T7.0T
Jun 22, 20264.5T1.6T0.25T0.00T6.3T
Jun 15, 20264.5T1.6T0.26T0.00T6.4T
Jun 8, 20264.0T2.2T0.26T0.24T6.7T
Jun 1, 20263.1T1.9T0.23T0.00T5.1T
May 25, 20263.2T2.0T0.23T0.00T5.4T
May 18, 20262.4T2.0T0.23T0.00T4.7T
May 11, 20262.1T1.7T0.24T0.00T4.0T
May 4, 20261.8T1.6T0.24T0.00T3.6T
Apr 27, 20261.4T1.5T0.24T0.00T3.2T
Apr 20, 20261.8T1.6T0.23T0.00T3.7T
Apr 13, 20261.5T1.7T0.25T0.00T3.5T
Apr 6, 20261.2T1.5T0.23T0.00T2.9T
Mar 30, 20261.1T1.4T0.21T0.00T2.7T
Mar 23, 20261.1T1.4T0.29T0.00T2.8T
Mar 16, 20261.1T1.4T0.28T0.00T2.8T
Mar 9, 20260.90T1.4T0.27T0.00T2.5T
Mar 2, 20260.85T1.3T0.20T0.00T2.3T
Feb 23, 20260.75T1.2T0.17T0.00T2.1T
Feb 16, 20260.80T0.89T0.14T0.00T1.8T
Feb 9, 20260.75T0.70T0.14T0.00T1.6T
Feb 2, 20260.51T0.80T0.14T0.00T1.4T
Jan 26, 20260.38T0.86T0.11T0.00T1.3T
Jan 19, 20260.36T0.78T0.09T0.00T1.2T
Jan 12, 20260.62T0.69T0.13T0.00T1.4T
Jan 5, 20260.37T0.61T0.09T0.00T1.1T
Dec 29, 20250.18T0.43T0.08T0.00T0.68T
Dec 22, 20250.18T0.44T0.09T0.00T0.71T
Dec 15, 20250.21T0.51T0.10T0.00T0.82T
Dec 8, 20250.21T0.55T0.10T0.00T0.86T
Dec 1, 20250.21T0.56T0.08T0.00T0.85T
Nov 24, 20250.11T0.56T0.06T0.00T0.73T
Nov 17, 20250.00T0.64T0.07T0.00T0.71T
Nov 10, 20250.00T0.76T0.08T0.00T0.84T
Nov 3, 20250.00T0.85T0.07T0.00T0.92T
Oct 27, 20250.00T0.84T0.07T0.00T0.91T
Oct 20, 20250.00T0.74T0.10T0.00T0.84T
Oct 13, 20250.00T0.74T0.04T0.00T0.78T
Oct 6, 20250.00T0.82T0.00T0.00T0.82T
Sep 29, 20250.01T0.70T0.00T0.00T0.70T
Sep 22, 20250.02T0.66T0.01T0.00T0.68T
Sep 15, 20250.03T0.67T0.01T0.00T0.71T
Sep 8, 20250.02T0.66T0.02T0.00T0.69T
Sep 1, 20250.03T0.63T0.02T0.00T0.68T
Aug 25, 20250.02T0.66T0.02T0.00T0.71T
Aug 18, 20250.03T0.66T0.02T0.00T0.71T
Aug 11, 20250.04T0.68T0.02T0.00T0.73T
Aug 4, 20250.05T0.66T0.02T0.00T0.72T
Jul 28, 20250.04T0.75T0.02T0.00T0.81T
Jul 21, 20250.04T0.69T0.03T0.00T0.76T
Jul 7, 20250.02T0.44T0.02T0.00T0.48T
Jun 30, 20250.02T0.43T0.01T0.00T0.45T
Jun 23, 20250.02T0.47T0.01T0.00T0.49T
Jun 16, 20250.02T0.53T0.00T0.00T0.55T
Jun 2, 20250.02T0.46T0.00T0.00T0.49T
May 26, 20250.02T0.52T0.00T0.00T0.54T
May 19, 20250.01T0.52T0.00T0.00T0.53T
May 12, 20250.00T0.42T0.00T0.00T0.42T
May 5, 20250.00T0.40T0.01T0.00T0.41T
Apr 28, 20250.00T0.42T0.01T0.00T0.42T
Apr 21, 20250.00T0.45T0.01T0.00T0.46T
Apr 14, 20250.00T0.51T0.01T0.00T0.52T
Apr 7, 20250.00T0.49T0.01T0.00T0.50T
Mar 31, 20250.00T0.47T0.02T0.00T0.49T
Mar 24, 20250.00T0.50T0.01T0.00T0.51T
Mar 17, 20250.00T0.53T0.05T0.00T0.58T
Mar 10, 20250.00T0.52T0.03T0.00T0.55T
Mar 3, 20250.00T0.44T0.01T0.00T0.44T
Feb 24, 20250.00T0.39T0.01T0.00T0.40T
Feb 17, 20250.00T0.29T0.01T0.00T0.30T
Feb 10, 20250.00T0.34T0.01T0.00T0.35T
Feb 3, 20250.00T0.35T0.01T0.00T0.36T
Jan 27, 20250.00T0.31T0.01T0.00T0.31T
Jan 20, 20250.00T0.25T0.01T0.00T0.25T
Jan 13, 20250.00T0.24T0.01T0.00T0.25T
Jan 6, 20250.00T0.23T0.01T0.00T0.24T

In the week of Aug 31, 2026, Sonnet accounts for 47.5% of Anthropic's 5.2T. The mix: Sonnet 47.5%, Opus 41.8%, Fable 7.1%, Haiku 3.6%. In the week of Jan 6, 2025, the largest was Sonnet, at 97.5%. The biggest mix shift was Sonnet, from 97.5% to 47.5% of the lab's volume.

Weeks in the top 10 by model

Anthropic models that reached the top 10, most weeks first. Full history since Jan 2025: does not respond to the window or the filters

OpusSonnetHaiku

● still above half its own peak in the latest week: the count for these models can still grow.

See the numbers
ModelFamilyWeeks in the top 10DebutPeak share
claude-3-7-sonnetSonnet25Feb 24, 202523.5%
claude-4-sonnetSonnet24May 19, 202519.1%
claude-4.5-sonnetSonnet22Sep 29, 202511.9%
claude-4.6-sonnetSonnet19Feb 16, 20266.8%
claude-4.7-opusOpus13Apr 13, 20267.3%
claude-3.5-sonnetSonnet12Jan 6, 202526.8%
claude-4.6-opusOpus11Feb 2, 20266.0%
claude-4.5-opusOpus10Nov 24, 20258.2%
claude-4.8-opusOpus8May 25, 20264.5%
claude-3-5-haikuHaiku1Jan 6, 20253.1%
claude-opus-5Opus1Jul 20, 20263.5%

claude-3-7-sonnet is Anthropic's longest-running model at the top, with 25 weeks among the 10 most used. The most recent debut, claude-opus-5 (Jul 2026), has 1 week in the top 10. None of the 11 is above half its own peak today.

10

Model lifecycle

How fast a model rises, peaks and loses ground. Turnover and the age of the leaders respond to the window and the grouping; the cohorts and the shape of a season use the full history since Jan 2025.

Top-10 turnover

How many of the 10 most used models were not in the top 10 four weeks earlier. From 0 (frozen top) to 10 (fully replaced top)

0246810Apr '25JulOctJan '26AprJul
See the numbers
WeekNew in the top 10
Aug 31, 20264
Aug 24, 20265
Aug 17, 20264
Aug 10, 20265
Aug 3, 20265
Jul 27, 20264
Jul 20, 20264
Jul 13, 20263
Jul 6, 20264
Jun 29, 20262
Jun 22, 20263
Jun 15, 20264
Jun 8, 20263
Jun 1, 20264
May 25, 20265
May 18, 20264
May 11, 20266
May 4, 20266
Apr 27, 20266
Apr 20, 20264
Apr 13, 20263
Apr 6, 20264
Mar 30, 20264
Mar 23, 20264
Mar 16, 20265
Mar 9, 20263
Mar 2, 20264
Feb 23, 20265
Feb 16, 20265
Feb 9, 20265
Feb 2, 20263
Jan 26, 20262
Jan 19, 20261
Jan 12, 20262
Jan 5, 20265
Dec 29, 20255
Dec 22, 20254
Dec 15, 20255
Dec 8, 20253
Dec 1, 20254
Nov 24, 20254
Nov 17, 20254
Nov 10, 20252
Nov 3, 20253
Oct 27, 20253
Oct 20, 20253
Oct 13, 20254
Oct 6, 20254
Sep 29, 20253
Sep 22, 20254
Sep 15, 20255
Sep 8, 20255
Sep 1, 20254
Aug 25, 20253
Aug 18, 20252
Aug 11, 20252
Aug 4, 20253
Jul 28, 20253
Jul 21, 20255
Jul 7, 20253
Jun 30, 20252
Jun 23, 20253
Jun 16, 20254
Jun 2, 20253
May 26, 20253
May 19, 20254
May 12, 20252
May 5, 20254
Apr 28, 20254
Apr 21, 20254
Apr 14, 20255
Apr 7, 20255
Mar 31, 20254
Mar 24, 20253
Mar 17, 20255
Mar 10, 20254
Mar 3, 20255
Feb 24, 20255
Feb 17, 20253
Feb 10, 20254
Feb 3, 20253
Jan 27, 2025
Jan 20, 2025
Jan 13, 2025
Jan 6, 2025

In the window, a median of 4 of the 10 most used models are new compared with four weeks earlier. In the last period, there were 4. The peak was 6, in the week of Apr 27, 2026. The top 10 in the week of Aug 31, 2026, by volume: hy4-previewnew, gpt-5.6-luna, glm-5.3-flashnew, deepseek-v4-flash · Jul 2026, deepseek-v4-flash · Apr 2026, minimax-m3 (free)new, hy3, nemotron-3-ultra-550b-a55b (free), glm-5.3new, mimo-v2.5.

Median age of the top 10

Weeks since first appearance in the ranking (first 12 weeks censored by the start of the dataset)

05101520Apr '25JulOctJan '26AprJul
See the numbers
WeekMedian age (weeks)
Aug 31, 20265.5 wk
Aug 24, 20264.5 wk
Aug 17, 20267.5 wk
Aug 10, 20266.5 wk
Aug 3, 20265.5 wk
Jul 27, 20268 wk
Jul 20, 20267.5 wk
Jul 13, 20266.5 wk
Jul 6, 20266 wk
Jun 29, 20268 wk
Jun 22, 20267.5 wk
Jun 15, 20266.5 wk
Jun 8, 20266.5 wk
Jun 1, 20265.5 wk
May 25, 20265 wk
May 18, 20264.5 wk
May 11, 20263.5 wk
May 4, 20262.5 wk
Apr 27, 20266.5 wk
Apr 20, 202610 wk
Apr 13, 202613.5 wk
Apr 6, 20267.5 wk
Mar 30, 20266.5 wk
Mar 23, 20266.5 wk
Mar 16, 20265.5 wk
Mar 9, 20265.5 wk
Mar 2, 20265 wk
Feb 23, 20264 wk
Feb 16, 20266 wk
Feb 9, 20265 wk
Feb 2, 20269.5 wk
Jan 26, 202613.5 wk
Jan 19, 202612.5 wk
Jan 12, 202611.5 wk
Jan 5, 20266.5 wk
Dec 29, 20259.5 wk
Dec 22, 20258.5 wk
Dec 15, 202513.5 wk
Dec 8, 202516.5 wk
Dec 1, 202512 wk
Nov 24, 202510.5 wk
Nov 17, 20257.5 wk
Nov 10, 202513.5 wk
Nov 3, 202512.5 wk
Oct 27, 202516.5 wk
Oct 20, 202518 wk
Oct 13, 202515.5 wk
Oct 6, 202512.5 wk
Sep 29, 202512.5 wk
Sep 22, 202511.5 wk
Sep 15, 202513 wk
Sep 8, 202514 wk
Sep 1, 202513 wk
Aug 25, 202512 wk
Aug 18, 202512.5 wk
Aug 11, 202511.5 wk
Aug 4, 202511 wk
Jul 28, 20259.5 wk
Jul 21, 202511.5 wk
Jul 7, 20257 wk
Jun 30, 202512.5 wk
Jun 23, 20257.5 wk
Jun 16, 202512 wk
Jun 2, 20259.5 wk
May 26, 20258.5 wk
May 19, 20257.5 wk
May 12, 20259 wk
May 5, 20258 wk
Apr 28, 20257 wk
Apr 21, 20256 wk
Apr 14, 20255 wk
Apr 7, 20254 wk
Mar 31, 20256.5 wk
Mar 24, 2025
Mar 17, 2025
Mar 10, 2025
Mar 3, 2025
Feb 24, 2025
Feb 17, 2025
Feb 10, 2025
Feb 3, 2025
Jan 27, 2025
Jan 20, 2025
Jan 13, 2025
Jan 6, 2025

In the week of Aug 31, 2026, the 10 most used models have a median age of 5.5 wk since their ranking debut. In the window, it ranged from 2.5 wk (May 4, 2026) to 18 wk (Oct 20, 2025). The first periods are left out because they fall in the censored weeks at the start of the series.

Time to peak and half-life, by launch quarter

Median in weeks, models that reached 2% of weekly volume, by quarter of ranking debut. Full history since Jan 2025: does not respond to the window or the grouping

Weeks to peakHalf-life after peak
0246Q1 '25n=25Q2 '25n=14Q3 '25n=20Q4 '25n=16Q1 '26n=18Q2 '26n=18Q3 '26n=153342.53.5222.5135321*

* Dashed bar: in the Q3 '26 cohort, half or more of the models are still above half their peak. Their half-life only counts models that have already fallen, so it tends to rise.

See the numbers
QuarternWks to peakHalf-lifeWith half-lifeAbove half of peak
Q1 '252533250%
Q2 '251442.5140%
Q3 '25203.52200%
Q4 '251622.5160%
Q1 '261813180%
Q2 '261853176%
Q3 '261521753%

Across the 126 models, the median is 2 weeks to peak and 2 weeks from peak to losing half of it. The Q1 '25 cohort took 3 weeks to peak and had a half-life of 3; the Q2 '26 cohort, the most recent in which most models have already fallen, 5 and 3. The Q1 '25 cohort includes the models that already existed when the series begins, in Jan 2025: for them, the count to peak starts from that date, not from launch.

The shape of a season

126 models aligned at their debut week, each normalized to its own peak. The thick line is the median; the band, the 1st to 3rd quartile. Full history, does not respond to the filtered view

Median of the 1261st to 3rd quartile
0%25%50%75%100%debut5 wk10 wk15 wk20 wk25 wk30 wk35 wk40 wktwo quartersmedian peak: week 1half of peak: week 80% of peak
See all 126 models
ModelOriginCohortPeak shareWks to peakHalf-lifeStatus
claude-3.5-sonnet (beta)US/CanadaQ1 '2530.2%05 wkbelow half its peak
gemini-2.0-flash-001US/CanadaQ1 '2528.1%26 wkbelow half its peak
claude-3.5-sonnetUS/CanadaQ1 '2526.8%52 wkbelow half its peak
grok-code-fast-1US/CanadaQ3 '2526.6%96 wkbelow half its peak
claude-3-7-sonnetUS/CanadaQ1 '2523.5%211 wkbelow half its peak
grok-4.1-fast (free)US/CanadaQ4 '2522.0%11 wkbelow half its peak
gpt-4o-miniUS/CanadaQ1 '2520.4%185 wkbelow half its peak
claude-4-sonnetUS/CanadaQ2 '2519.1%910 wkbelow half its peak
minimax-m2.5ChinaQ1 '2618.4%14 wkbelow half its peak
hy3 (free)ChinaQ3 '2618.4%11 wkbelow half its peak
mimo-v2.5ChinaQ2 '2618.0%112 wkbelow half its peak
mimo-v2-proChinaQ1 '2617.4%12 wkbelow half its peak
grok-4-fast (free)US/CanadaQ3 '2517.4%12 wkbelow half its peak
qwen3.6-plus-04-02 (free)ChinaQ1 '2617.0%01 wkbelow half its peak
deepseek-v4-flash · Jul 2026ChinaQ3 '2614.8%2above half its peak
ox-alphaUnknownQ3 '2613.9%11 wkbelow half its peak
hy3ChinaQ3 '2613.2%32 wkbelow half its peak
hy3-preview (free)ChinaQ2 '2612.7%12 wkbelow half its peak
deepseek-v4-flash · Apr 2026ChinaQ2 '2612.7%143 wkbelow half its peak
hy4-previewChinaQ3 '2612.7%1above half its peak
claude-4.5-sonnetUS/CanadaQ3 '2511.9%415 wkbelow half its peak
kimi-k2.5-0127ChinaQ1 '2611.8%13 wkbelow half its peak
gpt-5.6-lunaUS/CanadaQ3 '2611.2%8above half its peak
glm-5.3-flashChinaQ3 '2610.7%1above half its peak
hy3-previewChinaQ2 '2610.6%27 wkbelow half its peak
gemini-2.5-flashUS/CanadaQ2 '2510.6%518 wkbelow half its peak
gemini-2.5-flash-preview-05-20US/CanadaQ2 '259.8%54 wkbelow half its peak
minimax-m3ChinaQ2 '269.7%16 wkbelow half its peak
quasar-alphaUnknownQ1 '259.6%11 wkbelow half its peak
gemini-flash-1.5US/CanadaQ1 '259.0%12 wkbelow half its peak
gemini-flash-1.5-8bUS/CanadaQ1 '258.6%12 wkbelow half its peak
gemini-2.5-pro-preview-03-25US/CanadaQ1 '258.6%83 wkbelow half its peak
gemini-2.5-pro-exp-03-25 (free)US/CanadaQ1 '258.5%32 wkbelow half its peak
gemini-3-flash-previewUS/CanadaQ4 '258.3%69 wkbelow half its peak
claude-4.5-opusUS/CanadaQ4 '258.2%73 wkbelow half its peak
step-3.5-flash (free)ChinaQ1 '267.9%54 wkbelow half its peak
deepseek-chat-v3-0324 (free)ChinaQ1 '257.9%156 wkbelow half its peak
gemini-2.5-flash-preview-04-17US/CanadaQ2 '257.7%43 wkbelow half its peak
mimo-v2-flash (free)ChinaQ4 '257.7%51 wkbelow half its peak
kimi-k2.6ChinaQ2 '267.6%13 wkbelow half its peak
owl-alphaUnknownQ2 '267.4%81 wkbelow half its peak
claude-4.7-opusUS/CanadaQ2 '267.3%67 wkbelow half its peak
gemini-2.0-flash-lite-001US/CanadaQ1 '257.2%161 wkbelow half its peak
gemini-2.5-proUS/CanadaQ2 '257.1%29 wkbelow half its peak
deepseek-v3.2ChinaQ4 '256.8%913 wkbelow half its peak
claude-4.6-sonnetUS/CanadaQ1 '266.8%89 wkbelow half its peak
llama-3.3-70b-instructUS/CanadaQ1 '256.6%112 wkbelow half its peak
deepseek-chat-v3-0324ChinaQ1 '256.2%1410 wkbelow half its peak
grok-4.1-fastUS/CanadaQ4 '256.2%01 wkbelow half its peak
gemini-2.5-pro-exp-03-25US/CanadaQ2 '256.2%12 wkbelow half its peak
qwen3.6-plus-preview (free)ChinaQ1 '266.1%01 wkbelow half its peak
glm-5.2ChinaQ2 '266.1%37 wkbelow half its peak
claude-4.6-opusUS/CanadaQ1 '266.0%102 wkbelow half its peak
minimax-m2 (free)ChinaQ4 '255.8%21 wkbelow half its peak
deepseek-v4-proChinaQ2 '265.8%143 wkbelow half its peak
glm-5ChinaQ1 '265.8%12 wkbelow half its peak
minimax-m2.7ChinaQ1 '265.7%17 wkbelow half its peak
gpt-oss-120bUS/CanadaQ3 '255.6%182 wkbelow half its peak
horizon-betaUnknownQ3 '255.5%11 wkbelow half its peak
deepseek-r1 (free)ChinaQ1 '255.4%73 wkbelow half its peak
optimus-alphaUnknownQ2 '255.4%01 wkbelow half its peak
claude-3-7-sonnet (thinking)US/CanadaQ1 '255.3%37 wkbelow half its peak
nemotron-3-ultra-550b-a55b (free)US/CanadaQ2 '265.1%11above half its peak
glm-5-turboChinaQ1 '265.1%02 wkbelow half its peak
qwen3-coder-480b-a35b-07-25ChinaQ3 '254.7%33 wkbelow half its peak
sonoma-sky-alphaUnknownQ3 '254.6%21 wkbelow half its peak
qwen3-coder-480b-a35b-07-25 (free)ChinaQ3 '254.6%11 wkbelow half its peak
minimax-m2.1ChinaQ4 '254.5%62 wkbelow half its peak
hunter-alphaUnknownQ1 '264.5%11 wkbelow half its peak
claude-4.8-opusUS/CanadaQ2 '264.5%54 wkbelow half its peak
gemini-2.5-flash-lite-preview-06-17US/CanadaQ2 '254.5%11 wkbelow half its peak
minimax-m3 (free)ChinaQ3 '264.3%1above half its peak
deepseek-chat-v3.1 (free)ChinaQ3 '254.2%51 wkbelow half its peak
deepseek-chat-v3ChinaQ1 '254.1%04 wkbelow half its peak
mistral-nemoEuropeQ1 '254.1%51 wkbelow half its peak
deepseek-chat-v3.1ChinaQ3 '254.0%11 wkbelow half its peak
grok-4-fastUS/CanadaQ3 '254.0%211 wkbelow half its peak
gemini-2.5-flash-liteUS/CanadaQ3 '254.0%2412 wkbelow half its peak
trinity-large-preview (free)US/CanadaQ1 '263.9%43 wkbelow half its peak
minimax-m2ChinaQ4 '253.9%25 wkbelow half its peak
mimo-v2.5-proChinaQ2 '263.9%51 wkbelow half its peak
llama-3.2-1b-instructUS/CanadaQ1 '253.8%02 wkbelow half its peak
claude-opus-5US/CanadaQ3 '263.5%31 wkbelow half its peak
gpt-4.1-miniUS/CanadaQ2 '253.5%203 wkbelow half its peak
qwen3-30b-a3b-04-28ChinaQ2 '253.5%111 wkbelow half its peak
gemini-3.7-flashUS/CanadaQ3 '263.5%2above half its peak
deepseek-r1-0528 (free)ChinaQ2 '253.5%122 wkbelow half its peak
gemini-3.6-flashUS/CanadaQ3 '263.4%22 wkbelow half its peak
step-3.7-flashChinaQ2 '263.3%52 wkbelow half its peak
claude-3-7-sonnet (beta)US/CanadaQ1 '253.3%05 wkbelow half its peak
step-3.5-flashChinaQ1 '263.3%65 wkbelow half its peak
nemotron-3-super-120b-a12b (free)US/CanadaQ1 '263.1%48 wkbelow half its peak
claude-3-5-haikuUS/CanadaQ1 '253.1%101 wkbelow half its peak
gemini-2.0-pro-exp-02-05 (free)US/CanadaQ1 '253.1%43 wkbelow half its peak
gemini-3-pro-previewUS/CanadaQ4 '253.1%84 wkbelow half its peak
kimi-k2ChinaQ3 '253.0%52 wkbelow half its peak
gemini-2.5-pro-preview-06-05US/CanadaQ2 '253.0%21 wkbelow half its peak
gpt-oss-20bUS/CanadaQ3 '252.9%92 wkbelow half its peak
elephant-alphaUnknownQ2 '262.8%01 wkbelow half its peak
laguna-s-2.1 (free)US/CanadaQ3 '262.7%23 wkbelow half its peak
gpt-5US/CanadaQ3 '252.6%74 wkbelow half its peak
qwen3-coder-30b-a3b-instructChinaQ3 '252.6%11 wkbelow half its peak
glm-4.7ChinaQ4 '252.6%25 wkbelow half its peak
glm-5.3ChinaQ3 '262.6%2above half its peak
sherlock-think-alphaUnknownQ4 '252.6%11 wkbelow half its peak
gpt-4.1US/CanadaQ2 '252.6%010 wkbelow half its peak
polaris-alphaUnknownQ4 '252.5%11 wkbelow half its peak
kimi-k3ChinaQ3 '262.5%2above half its peak
gpt-5.4US/CanadaQ1 '262.5%63 wkbelow half its peak
mythomax-l2-13bOtherQ1 '252.5%21 wkbelow half its peak
gpt-5-nanoUS/CanadaQ3 '252.5%272 wkbelow half its peak
ling-3.0-flash (free)ChinaQ3 '262.5%11 wkbelow half its peak
devstral-2512 (free)EuropeQ4 '252.4%25 wkbelow half its peak
gpt-5.5US/CanadaQ2 '262.4%102 wkbelow half its peak
llama-3.1-70b-instructUS/CanadaQ1 '252.4%23 wkbelow half its peak
gpt-5.2US/CanadaQ4 '252.4%14 wkbelow half its peak
mimo-v2-omniChinaQ1 '262.4%12 wkbelow half its peak
gemini-2.5-flash-preview-04-17 (thinking)US/CanadaQ2 '252.4%41 wkbelow half its peak
glm-5.1ChinaQ2 '262.4%15 wkbelow half its peak
gemini-3.1-pro-previewUS/CanadaQ1 '262.3%74 wkbelow half its peak
horizon-alphaUnknownQ3 '252.2%01 wkbelow half its peak
glm-4.6ChinaQ3 '252.2%111 wkbelow half its peak
grok-4-07-09US/CanadaQ3 '252.1%62 wkbelow half its peak
llama-4-maverick-17b-128e-instructUS/CanadaQ1 '252.1%171 wkbelow half its peak
ling-2.6-1t (free)ChinaQ2 '262.0%11 wkbelow half its peak
kat-coder-pro-v1 (free)ChinaQ4 '252.0%62 wkbelow half its peak

The median of the 126 models peaks in week 1 and falls to half of that peak in week 8. At two quarters it sits at 0% of peak, and only 3 of the 84 models observed to that age remain above half their own peak. Fast rise, early peak and a long decline: that is the shape that repeats, and it is why what lasts is the method for evaluating and switching, not the choice of model.

11

What changed

The last period against four weeks earlier: who entered and left the top 10, who debuted and who gained and lost the most share. Responds to the window, the grouping and the filters.

Top-10 entries and exits

Week of Aug 31, 2026 vs. week of Aug 3, 2026

Entered the top 104
Left the top 104

Next to each name, the rank and share in the last period. A debut is a model that had no volume at all up to the comparison period.

Between the weeks of Aug 3, 2026 and Aug 31, 2026, 4 models entered the top 10 and 4 left; 23 debuted in the ranking. Biggest gain: hy4-preview, +12.7 pp, from 0.0% to 12.7%. Biggest drop: hy3, −8.2 pp, from 11.7% to 3.5%. A historical snapshot barely changes from one day to the next; turnover at the top does.

Biggest share changes

Share change in percentage points over the last 4 weeks, through Aug 31, 2026

Gained shareLost share
hy4-preview+12.7 pp
glm-5.3+2.6 pp
glm-5.2−3.0 pp
mimo-v2.5−5.8 pp
hy3−8.2 pp
See the numbers
ModelShare, week of Aug 3, 2026Share, week of Aug 31, 2026Change
hy4-preview0.00%12.71%+12.7 pp
glm-5.3-flash0.00%10.73%+10.7 pp
gpt-5.6-luna6.42%11.21%+4.8 pp
minimax-m3 (free)0.00%4.34%+4.3 pp
glm-5.30.00%2.62%+2.6 pp
gemini-3.7-flash0.00%1.82%+1.8 pp
deepseek-v4-pro3.78%1.09%−2.7 pp
glm-5.25.05%2.03%−3.0 pp
gemini-3.6-flash3.37%0.06%−3.3 pp
deepseek-v4-flash8.52%4.49%−4.0 pp
mimo-v2.57.81%2.04%−5.8 pp
hy311.66%3.46%−8.2 pp

5 of the 6 biggest gains come from models that started from near-zero share in the week of Aug 3, 2026: the gain is a debut, not a turnaround. In relative terms, gpt-5.6-luna multiplied its share by 1.7, from 6.42% to 11.21%. Percentage points favor models that are already large: for small models, the relative change tells more of the story.

12

Compare two models

Two models side by side. Adoption and trajectory respond to the window and the filter; peak and weeks in the top 10 cover the full history, and tasks come from the source's latest snapshot.

Compare

Pick two models to see adoption, price, context, lifecycle and use case side by side.

tencent/hy4-preview-20260827
openai/gpt-5.6-luna-20260709
Comparison of Hy4 preview and GPT-5.6 Luna. The better value in each row is marked.
IndicatorHy4 previewTencentGPT-5.6 LunaOpenAI
Shareweek of Aug 31, 202612.7% (better)11.2%
Tokens for the weekweek of Aug 31, 202615T (better)13T
Price per 1Mdeclared blend$1.25input $0.834 · output $2.50$0.45 (better)input $0.2 · output $1.20
Context window1M1M
Intelligence Index37.5
LaunchAug 28, 2026Jul 9, 2026
Weeks in the top 10full history26 (better)
Peak sharefull history12.7% (better)week of Aug 31, 202611.2%week of Aug 31, 2026
Median cost per sessionHermes Agent, 2 to 9 turns$0.028$0.008 (better)
What it is used for7-day snapshot from Sep 10, 2026; % is the model's share of the task
  • Workflow execution22.8%
  • File read and write26.7%
  • Multi-step planning10.2%
  • Code generation26.8%
  • Workflow execution8.8%
  • Debugging20.4%
Origin and weightsChina · Open weightsUS/Canada · Proprietary
Hy4 previewGPT-5.6 Luna
0%5%10%15%Jun 29Jul 13Jul 27Aug 10Aug 24

Weekly share within the filtered view, starting Jun 29, 2026, one period before either model first appears. A period with no recorded volume is left blank, not set to zero: outside the source's daily top 50, the model's volume is unknown.

See the numbers
WeekHy4 previewGPT-5.6 Luna
Aug 31, 202612.7%11.2%
Aug 24, 20262.7%6.9%
Aug 17, 20265.3%
Aug 10, 20267.0%
Aug 3, 20266.4%
Jul 27, 20263.4%
Jul 20, 20260.59%
Jul 13, 20260.44%
Jul 6, 20260.09%
Jun 29, 2026

Hy4 preview had 1.1 times the share of GPT-5.6 Luna in the week of Aug 31, 2026. Per 1M tokens at the declared blend, Hy4 preview costs 2.8 times as much as GPT-5.6 Luna. Across the full history, GPT-5.6 Luna spent 6 weeks in the top 10, versus 2 for Hy4 preview.

13

Season map

Each lab placed by the traction it has and the capability it declares, as a percentile against the rest. It is a map of relative position, like any market quadrant, with the weights open for adjustment.

Traction vs. capability, by lab

Relative position among the 19 labs with volume in the latest period. Vertical: traction, from token share, spend share and share growth. Horizontal: declared capability, from context window, multimodality, reasoning, catalog breadth and release cadence. The trail shows where each lab stood 12 weeks earlier.

US/CanadaChinaEuropeKorea
02550751000255075100LeadersChallengersProspectsNiche playersdeclared capability percentile →traction percentile →DeepSeekTencentZ.ai (GLM)OpenAIMiniMaxGoogleAnthropicNVIDIAXiaomiMoonshot AIPoolsideUpstageMetaInclusionAIQwenStepFunxAIThinking MachinesMistral AI

Bubble area is proportional to the lab's tokens in the latest period. Dashed trail: position 12 weeks earlier, for the 8 largest only.

Axis weights

The choice of weights is editorial. Move them to test whether anyone's position depends on it. Each axis is the weighted average of the percentiles; the number in parentheses is the effective share of the axis.

Vertical axis, traction
40(40%)
30(30%)
30(30%)
Horizontal axis, declared capability
25(25%)
20(20%)
20(20%)
20(20%)
15(15%)

Who sits where

Leaderstraction and capability above the medianDeepSeek, Z.ai (GLM), OpenAI, MiniMax, Google, Anthropic and Moonshot AI
Challengersheavy usage, capability below the medianTencent and NVIDIA
Prospectshigh capability, little tractionXiaomi and Meta
Niche playersboth below the medianPoolside, Upstage, InclusionAI, Qwen, StepFun, xAI, Thinking Machines and Mistral AI

The challengers quadrant is what the thesis predicts: traction built on price and availability, not on frontier technical features.

Biggest traction gain over the last 12 weeks: Z.ai (GLM), +37 percentile points. Biggest drop: Anthropic, −37.

methodBoth axes are percentiles among the labs present, not absolute values: moving up here can mean the others got worse.

See the numbers
LabOriginTractionCapabilityToken shareSpend shareGrowthModelsContextIndex
TencentChina91st45th17.2%16.5%+7.26 pp21M
Z.ai (GLM)China88th66th16.4%11.0%+14.52 pp31.31M44.9
OpenAIUS/Canada86th67th15.6%16.1%+9.62 pp81M52.8
DeepSeekChina80th71st19.1%3.9%+1.05 pp61.31M36.3
Moonshot AIChina59th58th1.9%8.0%+0.29 pp21M43.8
AnthropicUS/Canada56th68th4.8%35.7%−11.25 pp71M53.4
GoogleUS/Canada55th77th6.0%6.1%−3.95 pp101M41.2
NVIDIAUS/Canada52nd41st4.6%0.0%+1.83 pp3256k23.4
MiniMaxChina51st59th6.6%0.6%−4.57 pp31M29.6
UpstageKorea47th28th1.1%0.2%+1.12 pp1512k
MetaUS/Canada45th66th0.9%0.2%+0.93 pp31M
XiaomiChina39th56th2.4%0.5%−7.58 pp21M26.4
xAIUS/Canada38th40th0.3%0.8%+0.27 pp1488k44.4
InclusionAIChina33rd33rd0.8%0.0%+0.64 pp2256k
PoolsideUS/Canada33rd36th1.2%0.0%−0.33 pp11M
QwenChina32nd48th0.5%0.4%−1.08 pp41M40.3
StepFunChina25th37th0.4%0.2%−1.44 pp1256k
Thinking MachinesUS/Canada20th49th0.2%0.0%+0.17 pp11M25.5
Mistral AIEurope18th8th0.1%0.0%−0.35 pp1128k
14

Season signals

What is out of pattern in this window, detected automatically, and where the numbers go if the current rate holds. The second part is not a forecast, and the whole page exists to show that these rates change.

Findings

4 broken patterns in this window, detected by rule, not by curation.

Unusual acceleration+1.1pp

Solar Pro 4 rose 1.1 points in 4 weeks, 7.4 standard deviations above its own typical swing. It reached 1.1% of volume.

Gained share while pricier2.3× the median

Hy4 preview costs $1.25 per 1M, 2.3 times the market median, and still gained 12.7 points of share in 4 weeks. It runs against the dominant force in this dataset, which is price.

Outlasting its season19 weeks

MiMo-V2.5 launched 19 weeks ago and is still in the top 10, 2.4 times the median age of the leaders. The thesis says that is rare, which is exactly why it is worth looking at what it does differently.

The market is concentrating againHHI +11/week

The HHI had been falling across the window and reversed: over the last 26 weeks it rises 11 points per week. A reversal in concentration often comes before a dominant model arrives, or before a dominant one leaves.

If the current pace holds

Straight line fitted to the last 26 weeks, extended 13 more weeks.

IndicatorNowPace per weekIn 13 weeks
Chinese labs share61.3%+0.7pp rising70.6%
Open-weights share69.2%+1.4pp rising87.5%
Anthropic share4.5%0.4pp fallinghits 0.0% in 13 weeks
Google share5.7%0.4pp fallinghits 0.0% in 15 weeks
Traffic on free endpoints11.3%0.1pp falling9.4%
Top-5 concentration49.8%+0.7pp rising59.2%
Effective market price$1.01$0.02 falling$0.74

In red, the series whose line hits the floor or the ceiling before one and a half times the horizon. The end value is not published there: it is not a forecast that they will hit zero or saturate, it is proof that the current rate cannot hold that long.

This is not a forecast. It is the rate observed over the last 26 weeks, extended in a straight line. This page's thesis is precisely that these rates change, so the number is for sizing orders of magnitude and prompting the right question, not for planning.

Use-case signals new

The task that grew most between today's snapshot and the one 7 days earlier, and the model that gained most within it. It does not respond to the window or the filters.

Waiting for the archive

The archive holds 2 daily task snapshots, the latest from Sep 10, 2026. Each snapshot covers 7 days, so the first non-overlapping comparison needs 8 snapshots: 6 to go. If no day fails, it comes out with the snapshot of Sep 16, 2026.

Once it exists, this card shows the task that gained the most token share between the two snapshots, in percentage points, the model that gained most within it and, for contrast, the task that lost the most.

Today's snapshot is already stored: 29 classified tasks, led by Workflow execution, with 24.6% of classified tokens.

The source only publishes the current week and keeps no history. Every day not archived is a day that does not come back.

15

How to read each indicator

Each indicator answers three questions: what it is, how it's calculated and what it does not prove. The third matters most in a decision, and it is the one almost no dashboard publishes.

Source. A public daily ranking of token traffic on a language model router, under a CC BY 4.0 license, with full attribution in the footer. Quality indices come from Artificial Analysis through the same source, without recalculation.

Two time scales. History uses only full weeks, Monday to Sunday: 85 weeks, from Jan 6, 2025 to the week of Aug 31, 2026, with 403 models. The Now block uses daily data through the latest published day, Sep 10, 2026, in 7- and 30-day windows.

Spend is an estimate. The source sums input and output tokens without separating them, so spend uses a blend of 75% input and 25% output at list price, with a published floor and ceiling. A free endpoint costs zero.

Endpoint variants. In Now and on model pages, variants such as :free are added to the base model. In History they stay separate, because billing is one of the filter dimensions.

26 indicators
The shape of a season
what it is

A model's average adoption curve, from launch to decay.

how it's calculated

Each model is aligned on its own debut week and normalized to its own peak, so the horizontal axis becomes model age and the vertical axis becomes percent of peak. The thick line is the median and the band spans the first to the third quartile.

what it doesn't prove

It does not predict any single model's trajectory. It describes the observed distribution, and the spread within the band is wide.

Used in:
The source's aggregate line
what it is

Volume that fell outside each day's top 50, rolled up by the source into a single line.

how it's calculated

It comes straight from the API. It counts toward the total and the denominator of every percentage, but carries no metadata: we don't know which models it contains.

what it doesn't prove

That is why it drops out of any filtered view. With a filter active, percentages are calculated within the filtered view, and the filter bar shows how much volume was left out.

Used in:
Billing
what it is

Separates three things that are often confused: a paid model, a free endpoint of a paid model, and a model with no price.

how it's calculated

It comes from the catalog's pricing field. The :free suffix in the slug identifies a free endpoint of the same model, not a different model.

what it doesn't prove

Free endpoints usually carry rate limits and slower queues. Volume there does not signal willingness to pay.

Used in:
Median cost per session
what it is

What a typical session with a model costs, in dollars, inside an agent tool (the harness, such as a terminal coding assistant), broken out by the session's turn range.

how it's calculated

For each combination of harness, model and turn range, the source publishes the median cost per session over a rolling window. The site aggregates each harness and range by taking the median across models, without weighting by each model's volume, and keeps the minimum and maximum among them.

what it doesn't prove

A gap between harnesses does not prove that one tool saves money: each serves different tasks, codebases and users. The median across models weighs a rarely used model the same as a heavily used one. Cost is the price the router charges, with no contract or subscription discounts.

Used in:
Cost per evaluated task
what it is

What fraction of a standardized evaluation's tasks the model got right, and what each task cost on average, in dollars.

how it's calculated

It comes straight from the evaluations the source runs through its own router, at the price it charges: GPQA Diamond (graduate-level science questions) and τ-bench Verified in the airline scenario (the model serves a customer using tools). The efficiency frontier is calculated here: sorted by cost, a model makes the frontier if it beats the accuracy of every cheaper model.

what it doesn't prove

It is a snapshot taken on the evaluation date, with no confidence interval, on a fixed set of tasks. Cost depends on how many reasoning tokens the model spends and which provider served it, and scoring higher on a public test does not guarantee better results on your workload.

Used in:
Price tier and context window
what it is

Groupings of the blended price per million tokens and of the maximum context window.

how it's calculated

The blended price uses the declared mix of prompt and completion. The tiers are fixed cutoffs, published in the code, not rolling quantiles, so that what each tier means does not change from one week to the next.

what it doesn't prove

List price is not price paid: volume discounts, caching and contracts do not show up. An advertised maximum context also does not guarantee quality across the whole window.

Used in:
Use case
what it is

A breakdown of the volume the source can classify by use case: four top-level categories (coding, agents, data and general use) and the tasks within them, measured in tokens and in requests.

how it's calculated

The source classifies a sample of traffic over a rolling multi-day window and publishes, for each task, its share of tokens, its share of requests and the models with the most tokens in it. The other bucket is left out of the denominator. The site archives one snapshot a day, because the source keeps no history. The tokens ÷ requests ratio divides the two shares: above 1, each request for that task carries more tokens than the average across classified traffic.

what it doesn't prove

It is a sample, built on the source's own taxonomy and classifier, with no published uncertainty. Rolling windows overlap, so the difference between two consecutive snapshots is almost entirely noise. Share within a task shows which model is used most for it, not which one handles it best.

Used in:
Estimated spend
what it is

What the week's traffic would cost at list price.

how it's calculated

Tokens multiplied by each model's price. Because the source sums prompt and completion without splitting them, and completion costs several times more, we publish three numbers: a floor (all prompt), a ceiling (all completion) and the estimate at the declared mix. Free endpoints cost zero.

what it doesn't prove

It is not anyone's revenue. It ignores discounts, caching, batch pricing, contracts and the fact that part of the traffic runs on promotional credits. Use it to compare positioning across labs, not to size anyone's sales.

Used in:
Intelligence Index
what it is

A composite measure of model capability, produced by Artificial Analysis.

how it's calculated

It comes as-is through the OpenRouter API, along with the coding and agentic indexes. We do not recompute any of it.

what it doesn't prove

Coverage is partial, around a third of models with volume, and it is not random: lesser-known models get tested less. A missing index is not a sign of low quality.

Used in:
Daily 7- and 30-day windows
what it is

The sum of daily volume over the last 7 or 30 days the source has published, compared with the immediately preceding window of the same length. It is the measure behind the Now panel.

how it's calculated

The window ends on the last day the source published, never on today's date. Share is the model's volume divided by the window total, with the other line in the denominator. Endpoint variants, such as :free, are summed into the base model.

what it doesn't prove

Seven days is the most volatile view on the page. A strong week may be a launch spike or test traffic, and it only becomes a trend when it repeats across non-overlapping windows.

Used in:
Weights license
what it is

Whether the model's weights are publicly available for download.

how it's calculated

Evidence before heuristics: the presence of hugging_face_id in the catalog decides. Where the field is missing, classification falls back to a name-pattern heuristic, and the data records which of the two answered.

what it doesn't prove

Open weights does not mean a permissive license. Several models here restrict commercial use, and the distinction is not in the source's data.

Used in:
Leader by criterion
what it is

The top-ranked model under a stated criterion: highest score on an evaluation, highest observed usage, largest share gain or lowest price among models that clear a minimum quality bar.

how it's calculated

Each highlight states the criterion, source, date and how many models were compared. When two models tie on the published value, both appear as a group of leaders. The value pick uses a mix of 75% input tokens and 25% output tokens.

what it doesn't prove

Leading on evaluations is not leading on usage, and neither tells you which model fits your workload. The evaluation source publishes no uncertainty, so a small gap between first and second is not a proven advantage.

Used in:
Declared modality and reasoning
what it is

Two model attributes taken from the catalog: whether it accepts input beyond text (image, audio, file) and whether it declares reasoning support.

how it's calculated

Modality comes from the catalog's architecture field: input with more than one type counts as multimodal, text alone counts as text only. Reasoning is true when the catalog exposes the reasoning parameter for the model. A model with no catalog entry becomes "Unknown", never "no".

what it doesn't prove

It describes the model, not the request: traffic on a multimodal model may be text only, and the catalog does not say whether reasoning is mandatory or optional, or how many reasoning tokens were generated.

Used in:
Lab origin
what it is

The headquarters country of the organization that trained the model.

how it's calculated

Assigned by vendor, based on the slug prefix. Models under stealth/* and openrouter/* are anonymous tests and are classed as unknown.

what it doesn't prove

It does not say where inference runs or where data travels. A Chinese model served by a US provider still counts as Chinese here.

Used in:
Effective market price
what it is

What one million tokens actually consumed costs, on average.

how it's calculated

The week's estimated spend divided by the week's volume.

what it doesn't prove

A drop in effective price has two possible causes and the number does not separate them: models got cheaper, or demand shifted to cheaper models. In this dataset the second dominates.

Used in:
Price by provider
what it is

The same model served by different providers, each with its own price, maximum context, quantization and data retention policy.

how it's calculated

It comes from the endpoint list in the source's catalog, in dollars per million input and output tokens. Zero data retention flags an endpoint that declares it stores neither prompt nor response. Quantization is the declared numerical precision of the weights being served.

what it doesn't prove

List price excludes caching, discounts and rate limits. Lower quantization can change response quality, and not every provider declares its own. Latency and availability do not appear here.

Used in:
Money-to-volume ratio
what it is

How much of the money a lab captures relative to how many tokens it processes.

how it's calculated

Share of estimated spend divided by share of tokens. A ratio of 1 means the lab charges the market's average price; above 1, it charges more; below 1, less.

what it doesn't prove

It does not measure margin or efficiency. A lab with a high ratio may simply serve more expensive use cases rather than be more profitable.

Used in:
Top-10 turnover and age
what it is

How many of the ten most-used models are new compared with four weeks earlier, and how long ago the current ones debuted in the ranking.

how it's calculated

Turnover compares the current set with the one from four weeks ago. Median age counts weeks since each model's first appearance; the first twelve weeks of the series are censored, because we don't know when the models already present in January 2025 actually debuted.

what it doesn't prove

High turnover does not mean the models that dropped out got worse. Most of the time they were replaced by the same lab's next generation.

Used in:
Provider headquarters and zero data retention
what it is

Where the company serving inference is based, meaning who receives the request and returns the response, and whether it offers an endpoint with zero data retention (ZDR).

how it's calculated

Headquarters is the country each provider declares in the source's public provider list, counted once per provider. A provider with no declared headquarters is listed as not reported and is not assigned to any country. Zero data retention comes from the list of endpoints flagged as ZDR on the same date: we count endpoints, distinct models and distinct providers on that list.

what it doesn't prove

Headquarters is not where the data center sits or where data travels, and it says nothing about who trained the model. Zero data retention is the provider's declared policy for that endpoint, not an independent audit.

Used in:
Token share
what it is

The share of weekly volume captured by each model, lab or category.

how it's calculated

The entity's tokens divided by the week's total, with the other line included in the denominator. With a filter active, the denominator becomes the total of the filtered view.

what it doesn't prove

It is not market share by revenue or by users. A free model with heavy traffic has a high share and zero revenue. Always compare it with share of spend.

Used in:
Top-5 share and HHI
what it is

Two concentration measures: how much the five largest models capture, and the Herfindahl-Hirschman index across all models.

how it's calculated

The top 5 excludes the other line from the ranking but keeps it in the denominator: the question is how much of the market the five largest capture. HHI sums the squared percentage shares, normalized over named models only, because the aggregate line has no defined share. Below 1500 counts as unconcentrated on the conventional antitrust scale.

what it doesn't prove

Low concentration does not mean switching models is cheap. It measures available substitutes, not switching cost.

Used in:
Signals and extrapolation
what it is

Two different things in one section. Findings are measured facts that cut against the window's dominant pattern. Extrapolation is the recent rate extended in a straight line.

how it's calculated

Each finding follows a fixed rule published in the code: acceleration above two standard deviations of the model's own fluctuation, share gain for a model priced above the median, a stay in the top 10 longer than twice the median age, a sign reversal in the HHI trend, and a gap of more than 4 points between the China and open-weights curves. The extrapolation is a least-squares line fitted to the last third of the window (between 4 and 26 periods), extended by a quarter of the observed span (at most 13 weeks or 3 months). When the line hits the floor or the ceiling within one and a half times that horizon, the page shows when it breaks through rather than where it would end up.

what it doesn't prove

Nothing here is a forecast. The line assumes the rate holds, and this page's thesis is that it does not: no model has held the lead for more than two quarters in this dataset. The number is there to gauge order of magnitude and prompt the right question, never to plan around. Nor does a finding explain cause: it says something broke from the pattern, and figuring out why is still up to you.

Used in:
Time to peak and half-life
what it is

How many weeks a model takes from launch to its highest share, and how many more it takes afterward to lose half of it.

how it's calculated

Calculated per model, over the full series. Only models that reached at least 2% of weekly volume are included, so that noise from tiny models does not dominate the median. A model that has not yet fallen to half its peak shows no half-life rather than counting as zero.

what it doesn't prove

A short half-life is not a sign of a bad model. It is a sign of a market with a fast release cadence, and the effect shows up in good and bad models alike.

Used in:
Tokens per week
what it is

Total volume processed via OpenRouter in the week, summing prompt and completion across every model in the daily top 50 plus the aggregate line for the rest.

how it's calculated

The sum of daily totals for the ISO week, Monday through Sunday. Only full weeks are included. When grouped by month, the value shown is the month's weekly average, so that a five-week month does not look 25% larger than a four-week one.

what it doesn't prove

It is not global AI consumption. It is one router's traffic, with the bias described in the footer. Nor is it comparable between 2025 and 2026 without normalizing: the denominator grew 231-fold.

Used in:
Four-week share change
what it is

The difference, in percentage points, between a model's current share and its share four weeks earlier.

how it's calculated

Simple subtraction of percentages. When grouped by month, the comparison is with the previous month, the equivalent period.

what it doesn't prove

A percentage point is not a percent: going from 1% to 2% is +1 pp and also a doubling. For small models, the relative change tells more of the story.

Used in:
Volume by app
what it is

Tokens each app that identifies itself to the source processed in a day, along with the number of requests it made.

how it's calculated

It comes straight from the source, for the last published day, in the overall ranking and in each app category. Tokens per request divides one by the other. The trending list follows the source's own recent-growth criterion, and the volume shown next to it is that day's.

what it doesn't prove

Only apps that send identification are included; direct calls and anonymous traffic are left out. An app can appear in more than one category. A single day is a short window, exposed to launches, promotions and the app's own instability.

Used in: