Skip to content
Model Season every model has a season

gpt-oss-120b

openai/gpt-oss-120b
OpenAIUS/CanadaOpen weights

The 33rd most-used model on the router over the last 7 days, with 0.38% of tokens. Weekly share peaked at 5.6% in the week of Dec 8, 2025, with 18T tokens processed since Aug 4, 2025.

Launch
Aug 5, 2025
Context
128k tokens
Modality
text → text
Reasoning
yes
License
Open weights
Input
$0.037 / 1M
Output
$0.17 / 1M
Share, last 7 days
0.38%
33rd of 61 · Sep 4, 2026 to Sep 10, 2026
Peak weekly share
5.6%
week of Dec 8, 2025
Cumulative tokens
18T
since the week of Aug 4, 2025
Weeks in the top 10
5
of 57 weeks with volume

Weekly token share

Share of the router's weekly volume, since the first week with volume

0%2%4%6%8%Sep '25NovJan '26MarMayJulpeak 5.6% on Dec 8, 2025
See the numbers
WeekShare
Aug 31, 20260.400%
Aug 24, 20260.418%
Aug 17, 20260.553%
Aug 10, 20260.588%
Aug 3, 20260.606%
Jul 27, 20260.956%
Jul 20, 20260.810%
Jul 13, 20260.842%
Jul 6, 20260.987%
Jun 29, 20261.006%
Jun 22, 20261.175%
Jun 15, 20261.240%
Jun 8, 20261.414%
Jun 1, 20261.680%
May 25, 20261.828%
May 18, 20261.933%
May 11, 20262.062%
May 4, 20262.052%
Apr 27, 20262.091%
Apr 20, 20262.186%
Apr 13, 20262.323%
Apr 6, 20261.972%
Mar 30, 20261.609%
Mar 23, 20262.033%
Mar 16, 20262.016%
Mar 9, 20262.453%
Mar 2, 20263.359%
Feb 23, 20262.466%
Feb 16, 20262.235%
Feb 9, 20262.458%
Feb 2, 20262.643%
Jan 26, 20263.690%
Jan 19, 20263.492%
Jan 12, 20262.995%
Jan 5, 20261.889%
Dec 29, 20251.972%
Dec 22, 20251.950%
Dec 15, 20253.771%
Dec 8, 20255.596%
Dec 1, 20252.291%
Nov 24, 20251.024%
Nov 17, 20251.019%
Nov 10, 20250.893%
Nov 3, 20251.213%
Oct 27, 20251.292%
Oct 20, 20251.697%
Oct 13, 20251.665%
Oct 6, 20251.400%
Sep 29, 20251.300%
Sep 22, 20250.954%
Sep 15, 20251.008%
Sep 8, 20251.887%
Sep 1, 20250.768%
Aug 25, 20251.460%
Aug 18, 20251.466%
Aug 11, 20251.620%
Aug 4, 20251.398%

Debuted in the week of Aug 4, 2025 and peaked at 5.6% 18 weeks later. The latest full week came in at 0.40%, 7% of peak.

What this model is used for new

Tasks where the model ranks among the leaders, and its share of each

Ranks among the leaders in 8 of the 29 classified tasks. Its largest share is in summarization, 2.7% of the task's tokens; in the heaviest task where it appears, classification (5.4% of classified volume), it holds 2.3%.

Snapshot of the rolling 7-day window through Sep 10, 2026. Share: the model's tokens in the task over the task's total tokens. Weight: the task's share of classified volume.

Where to run it

Endpoints in the source's catalog, from cheapest to most expensive on input

ProviderInputOutputContextQuantizationZero retention
AkashML$0.03$0.17128kbf16yes
CoreWeave$0.03$0.17128kfp4yes
DekaLLM$0.03$0.18128kbf16no
DeepInfra$0.037$0.17128kbf16yes
Novita$0.05$0.25128kfp4yes
DigitalOcean$0.055$0.385125knot reportedyes
Mancer 2$0.055$0.50128kfp8yes
Google$0.09$0.36128knot reportedyes
BaseTen$0.10$0.50125kfp4yes
BaseTen$0.10$0.50125kfp4yes
Parasail$0.10$0.75128kfp4yes
SambaNova$0.14$0.95128knot reportedyes
Amazon Bedrock$0.15$0.60128knot reportedyes
Nebius$0.15$0.60128kfp4yes
Amazon Bedrock$0.15$0.60128knot reportedyes
DeepInfra$0.15$0.60128kbf16yes
SiliconFlow$0.15$0.60128kfp8yes
Phala$0.15$0.60128knot reportedyes
Together$0.15$0.60128knot reportedyes
Groq$0.15$0.60128knot reportedyes
Mara$0.15$0.75128knot reportedyes
DeepInfra$0.20$0.95128kfp8yes
Cerebras$0.35$0.75128kfp16yes

23 endpoints across 19 providers, with input from $0.03 to $0.35 per 1M tokens. 22 offer zero data retention.

Prices in US dollars per million tokens, from the latest archived endpoint catalog.

Benchmarks

Artificial Analysis indexes and benchmarks run by the source, with cost per task

Intelligence
12.3
63rd of 82
Coding
30.4
80th of 111
Agents
6.2
61st of 86

GPQA Diamond 75.1%

$0.007 per task, 4353 tasks. 85th of 131 evaluated. Median of evaluated models: 80.2%, $0.024 per task.

τ-bench verified, airline 64.4%

$0.013 per task, 1448 tasks. 78th of 122 evaluated. Median of evaluated models: 70.7%, $0.056 per task.

Artificial Analysis indexes, exposed by the catalog; benchmarks run by the source and archived on Sep 11, 2026. The gray mark is the median of evaluated models.

Against the week's largest models

The five largest proprietary and five largest open-weights models over the last 7 days

ModelLab7-day shareBlended priceContextIntelligenceWeeks in top 10
Proprietary
GPT-5.6 LunaOpenAIOpenAI11.5%$0.451M37.56
Gemini 3.8 FlashGoogleGoogle2.1%$1.501M41.20
Muse Spark 1.3 ContributorMetaMeta1.5%$0.1251M0
Claude Opus 5AnthropicAnthropic1.4%$10.001M50.71
GPT-5.6 SolOpenAIOpenAI1.4%$4.001M47.10
Open weights
Hy4 previewTencentTencent15.5%$1.251M1
DeepSeek V4 Flash 0731DeepSeekDeepSeek10.0%$0.0941.31M34.55
GLM 5.3 FlashZ.ai (GLM)Z.ai (GLM)10.0%$0.2371.31M41.92
MiMo-V2.5XiaomiXiaomi4.2%$0.1751M22.315
DeepSeek V4 Flash 0423DeepSeekDeepSeek3.9%$0.1111M24.819
gpt-oss-120bthis pageOpenAIOpenAI0.38%$0.07128k12.35

Share over the last 7 days (Sep 4, 2026 to Sep 10, 2026). Blended price in US dollars per million tokens, weighted 75% input and 25% output. A dash means the number is not published, not zero.

What this page doesn't say. Share here is a slice of one router's traffic, where the deciding factor is usually price per token. It is not market, revenue or user share, and it does not tell you which model fits your workload.