GPT-OSS API Benchmarks, Pricing & Provider Data
Compare GPT-OSS with another model
Choose a model to open its comparison page.
GPT-OSS API pricing covers undefined API provider} other undefined API providers}}, from $0.000025/request to $75.00/M. GPT-OSS free API options are available from undefined provider} other undefined providers}}. The page also shows measured API speed and first-token latency.
GPT-OSS is an open-weight language model family designed for self-hosted inference, research, and cost-efficient alternatives to proprietary GPT-class models.
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Pricing Comparison
Compare GPT-OSS API pricing across 289 providers. Prices range from $0.000025/request to $75.00/M. CM-API 公益站 offers the lowest rate at $0.000025/request. 24 providers offer free API credits or a free tier.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 100% | gpt-oss-120b | default | $0.137/request | - | 1637.3 t/s | 0.91 s | — | |
L1 100% | gpt-oss-120b-medium | default | $5.48/M | $10.96/M | 255.8 t/s | 1.45 s | — | |
L1 0% | cerebras/gpt-oss-120b | 高并发渠道 | $0.280/M | $0.600/M | 1344.7 t/s | 0.38 s | — | |
L1 0% | gpt-oss-20b | nvidia | $0.00006/M Cache read$0.00000018/MCache write$0.0000036/MCache write 1h$0.00000576/M | $0.0000066/M | — | — | — | |
L1 0% | gpt-oss-120b | nvidia | $0.000078/M Cache read$0.000000304/MCache write$0.00000608/MCache write 1h$0.00000973/M | $0.0000148/M | — | — | — | |
L1 100% | gpt-oss-120b | default | $4.38/M Audio in$0.219/M | $13.14/M | 1319.0 t/s | 0.61 s | — | |
L1 99% | FAST/gpt-oss-120b | 翻译 | $0.205/M | $0.068/M | 750.0 t/s | 0.92 s | ||
L1 99% | gpt-oss-120b | 翻译 | $0.068/M | $0.288/M | 126.7 t/s | 0.79 s | — | |
L1 99% | openai/gpt-oss-20b | 0倍倍率分组 | $0.0004/M Cache read$0.0002/M | $0.0022/M | — | — | — | |
L1 99% | openai/gpt-oss-120b | 临时渠道 | $0.0021/M Cache read$0.0002/M | $0.0082/M | — | — | — | |
L1 61% | openai/gpt-oss-20b | default | $0.010/request | - | 234.4 t/s | 1.48 s | — | |
L1 61% | openai/gpt-oss-120b:free | default | $0.010/request | - | — | — | — | |
L1 61% | openai/gpt-oss-120b | default | $0.010/request | - | — | — | — | |
L1 61% | openai/gpt-oss-20b:free | default | $0.010/request | - | — | — | — | |
L1 99% | gpt-oss-20b | default | $0.073/request | - | 216.3 t/s | 1.31 s | ||
L1 0% | openai/gpt-oss-120b | default | $2.97/M Cache read$2.97/M | $16.83/M | 124.7 t/s | 2.30 s | ||
L1 0% | openai/gpt-oss-20b | default | $2.97/M Cache read$2.97/M | $13.86/M | — | — | — | |
Jasper Free | L1 100% L2 100% | gpt-oss-20b | Unlimited | Free | Free | 100.0 t/s | 0.30 s | — |
DeadlySignal API Free | L1 99% L2 100% | gpt-oss-20b | default | Free | Free | 15.2 t/s | 11.66 s | — |
L1 99% L2 100% | openai/gpt-oss-20b | nvidia-free | Free | Free | — | — | — |
Alternatives & Similar Models
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
GPT-5.4
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
DeepSeek V3.2
deepseek-v3-2
DeepSeek V3.2 is an upgraded V3-series MoE model with stronger reasoning, coding, and math performance, widely available through OpenAI-compatible API relays.
Kimi K2.5
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.
DeepSeek V4 Pro
deepseek-v4-pro
DeepSeek V4 Pro is the professional-tier DeepSeek V4 model, targeting frontier reasoning, coding, and agent workflows with maximum capability.
MiniMax M2.5
minimax-m2-5
MiniMax M2.5 is MiniMax's flagship text model for coding and agents, with SOTA-level programming and agentic performance, improved token efficiency, and fast high-TPS API deployment.
Frequently Asked Questions
- What benchmark data does GPT-OSS include?
- LMSpeed shows GPT-OSS benchmark context, API price, output speed, first-token latency, and provider data across 313 providers when those signals are available.
- What is the GPT-OSS API price?
- GPT-OSS has pricing from undefined provider} other undefined providers}}, ranging from $0.000025/request to $75.00/M. CM-API 公益站 has the lowest listed price.
- What does the GPT-OSS API pricing table include?
- The GPT-OSS API pricing table compares 313 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest GPT-OSS API pricing?
- CM-API 公益站 currently has the lowest listed GPT-OSS price at $0.000025/request across undefined provider} other undefined providers}}.
- Can I compare GPT-OSS API price and speed together?
- Yes. LMSpeed shows GPT-OSS API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
- Is GPT-OSS API free?
- Yes, GPT-OSS free API options are available through 24 providerundefined other undefined} on LMSpeed, including 兔子API, 猫羽霖API, DeadlySignal API, 猫羽霖API, Zero API. These providers offer free API credits or a free tier with no per-token charges.
- Where can I get GPT-OSS free API access?
- LMSpeed currently lists 24 free API providerundefined other undefined} for GPT-OSS: 兔子API, 猫羽霖API, DeadlySignal API, 猫羽霖API, Zero API. Check each provider row before using it because free tier limits can change.
