Released on Aug 5, 2025

GPT-OSS API Benchmarks, Pricing & Provider Data

Compare GPT-OSS with another model

Choose a model to open its comparison page.

Share on X
LLM

GPT-OSS API pricing covers 454 API providers, from $0.0004/M to $495.00/M. GPT-OSS free API options are available from 26 providers. The page also shows measured API speed and first-token latency.

GPT-OSS is an open-weight language model family designed for self-hosted inference, research, and cost-efficient alternatives to proprietary GPT-class models.

Speed
374char/s
3.23 s
Cost
$0.0004/ 1M · 8:1 in:out
$0.0000658 in · $0.0003 out

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
131.1Ktokens
157.3 pages of text
OUTPUT
131.1Ktokens
8K128K1M4M
131.1K

Features

Technical Details

Input
Output
Total parameters
120B
Released
Aug 2025
Knowledge cutoff
2024-06-30
Tokenizer
GPT
Architecture
text->text
Moderated
Yes
Supported parameters
include_reasoningmax_tokensmin_preasoningseedstoptemperaturetool_choicetoolstop_atop_ktop_p

Pricing Comparison

Compare GPT-OSS API pricing across 428 providers. Prices range from $0.0004/M to $495.00/M. 6345ywz API offers the lowest rate at $0.0004/M. 26 providers offer free API credits or a free tier.

ProviderHealthModel VariantGroupInput ($/M)Output ($/M)Speed (t/s)First tokenAudit
L1
99%
gpt-oss-120b
default
$0.0014/request
-
1637.3 t/s
0.91 s
L1
99%
gpt-oss-120b-medium
default
$0.055/M
$0.110/M
255.8 t/s
1.45 s
L1
100%
gpt-oss-120b
default
$4.38/M
Audio in$0.219/M
$13.14/M
1319.0 t/s
0.61 s
L1
99%
gpt-oss-120b
model
$0.073/request
-
1201.8 t/s
0.78 s
L1
99%
gpt-oss-20b
91vip
$0.438/M
Cache read$0.0013/MCache write$0.026/MCache write 1h$0.042/M
$0.048/M
216.3 t/s
1.31 s
L1
98%
FAST/gpt-oss-120b
翻译
$0.205/M
$0.068/M
750.0 t/s
0.92 s
L1
98%
gpt-oss-120b
翻译
$0.068/M
$0.288/M
126.7 t/s
0.79 s
L1
98%
openai/gpt-oss-20b
0倍倍率分组
$0.0004/M
Cache read$0.0002/M
$0.0022/M
L1
98%
openai/gpt-oss-120b
0倍倍率分组
$0.0021/M
Cache read$0.0002/M
$0.0082/M
L1
100%
openai/gpt-oss-20b
default
$0.010/request
-
234.4 t/s
1.48 s
L1
100%
openai/gpt-oss-120b
default
$0.010/request
-
L1
100%
openai/gpt-oss-120b:free
default
$0.010/request
-
L1
100%
openai/gpt-oss-20b:free
default
$0.010/request
-
L1
12%
openai/gpt-oss-120b
default
$359.64/M
Cache read$359.64/M
$1798.20/M
124.7 t/s
2.30 s
L1
12%
openai/gpt-oss-20b
default
$289.71/M
Cache read$289.71/M
$1398.60/M
L1
100%
gpt-oss-120b
default
$0.075/M
$0.301/M
L1
100%
gpt-oss-20b
测试
$0.137/M
$0.548/M
L1
100%
gpt-oss-120b
default
$0.018/request
-
L1
100%
gpt-oss-20b
default
$0.027/M
$0.110/M
L1
100%
gpt-oss-120b
default
$0.151/M
$0.603/M
Showing 20 model IDs of 155.

Alternatives & Similar Models

Frequently Asked Questions

What benchmark data does GPT-OSS include?
LMSpeed shows GPT-OSS benchmark context, API price, output speed, first-token latency, and provider data across 454 providers when those signals are available.
What is the GPT-OSS API price?
GPT-OSS has pricing from 454 providers, ranging from $0.0004/M to $495.00/M. 6345ywz API has the lowest listed price.
What does the GPT-OSS API pricing table include?
The GPT-OSS API pricing table compares 454 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
Which provider has the cheapest GPT-OSS API pricing?
6345ywz API currently has the lowest listed GPT-OSS price at $0.0004/M across 454 providers.
Can I compare GPT-OSS API price and speed together?
Yes. LMSpeed shows GPT-OSS API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
Is GPT-OSS API free?
Yes, GPT-OSS free API options are available through 26 providerundefined other undefined} on LMSpeed, including 兔子API, 兔子API, 猫羽雫API, 猫羽霖API, Moyanjdc API. These providers offer free API credits or a free tier with no per-token charges.
Where can I get GPT-OSS free API access?
LMSpeed currently lists 26 free API providerundefined other undefined} for GPT-OSS: 兔子API, 兔子API, 猫羽雫API, 猫羽霖API, Moyanjdc API. Check each provider row before using it because free tier limits can change.

Also known as

@cf/openai/gpt-oss-120b@cf/openai/gpt-oss-20bCerebras/gpt-oss-120bFAST/gpt-oss-120bGPT-OSS-120B

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation