OpenAI
·Released on Feb 12, 2025

O3 Mini API Benchmarks, Pricing & Provider Data

Compare O3 Mini with another model

Choose a model to open its comparison page.

Share on X
LLM

O3 Mini benchmark, API pricing, and provider data cover 1126 API providers, with prices starting at $0.0062/request. O3 Mini free API options are available from 11 providers. The page also shows measured API speed and first-token latency.

OpenAI O3 Mini is a compact reasoning model in the O series, designed for efficient reasoning and problem-solving tasks.

Quality
#67of 100
51.0
LMSpeed score
Speed
64char/s
6.90 s
Cost
#115of 181
$0.0062/ 1M · 8:1 in:out
$0.0010 in · $0.0052 out

Category Performance

Observed-capability estimates with uncertainty reported separately from independent benchmark families.

Coverage
6 / 8
Methodology
V3.0
Category PerformanceObserved-capability estimates with uncertainty reported separately from independent benchmark families.Agents39.7Coding40.5Reasoning49.4Knowledge47.8Math58.5Multilingual-Multimodal-Instruction following55.6
#1Math58.5Estimated2/4 Measured dimensions
80% interval: 46.770.4
foundational math59.2
math_500 · math_500
competition math57.9
aime · aime
proof frontierPrior only
applied tool mathPrior only
#2Instruction following55.6Provisional1/4 Measured dimensions
80% interval: 39.371.9
constraint following55.6
ifeval · ifeval
structured outputPrior only
novel instruction generalizationPrior only
long multiturn instructionPrior only
#3Reasoning49.4Estimated2/4 Measured dimensions
80% interval: 38.360.6
abstract logicPrior only
scientific causal48.1
gpqa · gpqa / hle · hle
multistep constraints50.8
mmlu_pro · mmlu_pro
evidence verificationPrior only
#4Knowledge47.8Provisional1/4 Measured dimensions
80% interval: 30.365.2
broad knowledge47.8
mmlu · mmlu
professional knowledgePrior only
factualityPrior only
retrieval open bookPrior only
#5Coding40.5Estimated2/4 Measured dimensions
80% interval: 29.351.7
code generation53.1
livecodebench · livecodebench / scicode · scicode
repository engineeringPrior only
debugging testing27.9
swe_verified · swe_verified
tooling qualityPrior only
#6Agents39.7Provisional1/4 Measured dimensions
80% interval: 23.755.8
planningPrior only
tool use39.7
tau · tau2_bench
environment executionPrior only
recovery reliabilityPrior only
No data:MultilingualMultimodal

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
200Ktokens
240 pages of text
OUTPUT
100Ktokens
8K128K1M4M
200K

Features

Technical Details

Input
Output
Released
Feb 2025
Knowledge cutoff
2023-10-31
Documentation
Tokenizer
GPT
Architecture
text+file->text
Moderated
Yes
Supported parameters
include_reasoningmax_tokensreasoningresponse_formatseedstructured_outputstool_choicetools

Rankings

Excels at

Falls behind in

Detailed scores

Updated: Sep 1, 2026

Overall

undefined metric} other undefined metrics}}

Overall score51.0#67 / 100
Overall score
LMSpeed rank#67
Score51.0
UpdatedAug 20, 2026
Confidence1

Pricing

undefined metric} other undefined metrics}}

Input price$1.10/M#115 / 181Output price$4.40/M#116 / 181
Input price
LMSpeed rank#115
Score$1.10/M
UpdatedSep 1, 2026
Confidence4
Output price
LMSpeed rank#116
Score$4.40/M
UpdatedSep 1, 2026
Confidence4

Agents

V3.0

undefined metric} other undefined metrics}} · Provisional

Score39.780% interval23.7–55.81/4 Measured dimensionsΤ²-bench results28.7#72 / 82
Dimensions and evidence
Planning & decompositionPrior only
Tool use39.7
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
tau · tau2_bench · z -1.48 · q 1.00
Environment & long-horizon executionPrior only
Recovery & completion reliabilityPrior only
Τ²-bench resultsV3 evidence
LMSpeed rank#72
Score28.7
UpdatedAug 20, 2026
Confidence1

Coding

V3.0

undefined metric} other undefined metrics}} · Estimated

Score40.580% interval29.3–51.72/4 Measured dimensionsLiveCodeBench71.7%#27 / 115SciCode39.9%#86 / 206
Dimensions and evidence
Code generation53.1
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
livecodebench · livecodebench · z 0.65 · q 1.00
scicode · scicode · z 0.20 · q 1.00
Repository engineeringPrior only
Debugging & testing27.9
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
swe_verified · swe_verified · z -3.00 · q 1.00
Tool-assisted development & qualityPrior only
LiveCodeBenchV3 evidence
LMSpeed rank#27
Score71.7%
UpdatedSep 1, 2026
Confidence4
SciCodeV3 evidence
LMSpeed rank#86
Score39.9%
UpdatedSep 1, 2026
Confidence4
Coding score
LMSpeed rank#80
Score21.3
UpdatedAug 20, 2026
Confidence1
SWE-bench VerifiedV3 evidence
LMSpeed rank#46
Score49.3
UpdatedAug 20, 2026
Confidence1
AA-SciCodeV3 evidence
LMSpeed rank#70
Score39.9
UpdatedAug 20, 2026
Confidence1

Reasoning

V3.0

undefined metric} other undefined metrics}} · Estimated

Score49.480% interval38.3–60.62/4 Measured dimensionsMMLU-Pro79.1%#70 / 129GPQA74.8%#117 / 213
Dimensions and evidence
Abstract logicPrior only
Scientific & causal reasoning48.1
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
gpqa · gpqa · z -0.12 · q 1.00
hle · hle · z -0.41 · q 1.00
Multi-step constraints50.8
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
mmlu_pro · mmlu_pro · z -0.10 · q 1.00
Evidence integration & verificationPrior only
MMLU-ProV3 evidence
LMSpeed rank#70
Score79.1%
UpdatedSep 1, 2026
Confidence4
GPQAV3 evidence
LMSpeed rank#117
Score74.8%
UpdatedSep 1, 2026
Confidence4
HLEV3 evidence
LMSpeed rank#129
Score7.9%
UpdatedSep 1, 2026
Confidence4

Knowledge

V3.0

undefined metric} other undefined metrics}} · Provisional

Score47.880% interval30.3–65.21/4 Measured dimensionsKnowledge score65.6#38 / 69MMLU86.9#4 / 5
Dimensions and evidence
Broad knowledge47.8
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
mmlu · mmlu · z -0.22 · q 0.21
Professional knowledgePrior only
FactualityPrior only
Retrieval & open-book usePrior only
Knowledge score
LMSpeed rank#38
Score65.6
UpdatedAug 20, 2026
Confidence1
MMLUV3 evidence
LMSpeed rank#4
Score86.9
UpdatedAug 20, 2026
Confidence1
Artificial Analysis Intelligence Index
LMSpeed rank#87
Score19.2
UpdatedAug 20, 2026
Confidence1
AA-GPQA Diamond
LMSpeed rank#82
Score74.8
UpdatedAug 20, 2026
Confidence1
AA-HLE
LMSpeed rank#83
Score7.9
UpdatedAug 20, 2026
Confidence1

Math

V3.0

undefined metric} other undefined metrics}} · Estimated

Score58.580% interval46.7–70.42/4 Measured dimensionsAIME77.0%#14 / 68MATH-50097.3%#11 / 73
Dimensions and evidence
Foundational math59.2
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
math_500 · math_500 · z 1.26 · q 1.00
Competition math57.9
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
aime · aime · z 1.03 · q 1.00
Advanced proofsPrior only
Applied & tool-assisted mathPrior only
AIMEV3 evidence
LMSpeed rank#14
Score77.0%
UpdatedSep 1, 2026
Confidence4
MATH-500V3 evidence
LMSpeed rank#11
Score97.3%
UpdatedSep 1, 2026
Confidence4
AIME 2024
LMSpeed rank#1
Score87.3
UpdatedAug 20, 2026
Confidence1

Multilingual

V3.0

undefined metric} other undefined metrics}} · No data

No data80% interval30.8–69.20/4 Measured dimensions
Dimensions and evidence
Cross-language understandingPrior only
Multilingual generationPrior only
Reasoning transferPrior only
Low-resource robustnessPrior only

Multimodal

V3.0

undefined metric} other undefined metrics}} · No data

No data80% interval30.8–69.20/4 Measured dimensions
Dimensions and evidence
Perception & OCRPrior only
Document & spatial understandingPrior only
Visual reasoningPrior only
Video & grounded actionPrior only

Instruction following

V3.0

undefined metric} other undefined metrics}} · Provisional

Score55.680% interval39.3–71.91/4 Measured dimensionsInstruction Following score91.0#6 / 25IFEval93.9#5 / 16
Dimensions and evidence
Constraint following55.6
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
ifeval · ifeval · z 0.54 · q 1.00
Structured outputPrior only
Novel-instruction generalizationPrior only
Multi-turn & long instructionsPrior only
Instruction Following score
LMSpeed rank#6
Score91.0
UpdatedAug 20, 2026
Confidence1
IFEvalV3 evidence
LMSpeed rank#5
Score93.9
UpdatedAug 20, 2026
Confidence1

OpenRouter endpoints

1 endpoints

Third-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.

Provider endpointInputOutput1d uptime30m latency30m throughputContext / output
OpenAI
openai
$1.10/M$4.40/M100%undefined tokens / undefined tokens

Pricing Comparison

Compare O3 Mini API pricing across 1115 providers. Prices range from $0.0062/request to $300.00/M. KKSJ-AI offers the lowest rate at $0.0062/request. 11 providers offer free API credits or a free tier.

ProviderHealthModel VariantGroupInput ($/M)Output ($/M)Speed (t/s)First tokenAudit
L1
100%
o3-mini
default
-86%$0.151/M
Cache read$0.075/M
-86%$0.603/M
116.7 t/s
4.81 s
L1
100%
o3-mini-2025-01-31
default
-86%$0.151/M
Cache read$0.075/M
-86%$0.603/M
L1
100%
o3-mini
default
-93%$0.075/M
Cache read$0.038/M
-93%$0.301/M
L1
100%
o3-mini-2025-01-31
default
-93%$0.075/M
Cache read$0.038/M
-93%$0.301/M
L1
100%
o3-mini-all
default
$0.080/request
-
L1
100%
o3-mini-high-all
default
$0.160/request
-
L1
100%
o3-mini-2025-01-31
default
$1.10/M
$4.40/M
L1
100%
o3-mini
default
$1.10/M
$4.40/M
L1
100%
o3-mini-low
default
-86%$0.151/M
-86%$0.603/M
L1
100%
o3-mini-medium
default
-86%$0.151/M
-86%$0.603/M
L1
100%
o3-mini
default
-86%$0.151/M
-86%$0.603/M
L1
100%
o3-mini-2025-01-31
default
-86%$0.151/M
-86%$0.603/M
L1
100%
o3-mini-high
default
-86%$0.151/M
-86%$0.603/M
L1
100%
o3-mini-2025-01-31-high
az_stable_high_22
-86%$0.151/M
-86%$0.603/M
L1
100%
o3-mini-2025-01-31-low
az_stable_high_22
-86%$0.151/M
-86%$0.603/M
L1
100%
o3-mini-2025-01-31-medium
az_stable_high_22
-86%$0.151/M
-86%$0.603/M
L1
100%
o3-mini-all
default
-86%$0.151/M
-86%$0.603/M
L1
100%
o3-mini-high-all
default
-86%$0.151/M
-86%$0.603/M
L1
100%
o3-mini-2025-01-31-high
default
-23%$0.844/M
-23%$3.38/M
L1
100%
o3-mini-2025-01-31
default
-23%$0.844/M
-23%$3.38/M
Showing 20 model IDs of 246.

Alternatives & Similar Models

Frequently Asked Questions

What benchmark data does O3 Mini include?
LMSpeed shows O3 Mini benchmark context, API price, output speed, first-token latency, and provider data across 1126 providers when those signals are available.
What is the O3 Mini API price?
O3 Mini has pricing from 1126 providers, ranging from $0.0062/request to $300.00/M. KKSJ-AI has the lowest listed price.
What does the O3 Mini API pricing table include?
The O3 Mini API pricing table compares 1126 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
Which provider has the cheapest O3 Mini API pricing?
KKSJ-AI currently has the lowest listed O3 Mini price at $0.0062/request across 1126 providers.
Can I compare O3 Mini API price and speed together?
Yes. LMSpeed shows O3 Mini API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
Is O3 Mini API free?
Yes, O3 Mini free API options are available through 11 providerundefined other undefined} on LMSpeed, including 兔子API, 兔子API, 兔子API, 兔子API, 兔子API. These providers offer free API credits or a free tier with no per-token charges.
Where can I get O3 Mini free API access?
LMSpeed currently lists 11 free API providerundefined other undefined} for O3 Mini: 兔子API, 兔子API, 兔子API, 兔子API, 兔子API. Check each provider row before using it because free tier limits can change.

Also known as

o3-minio3-mini-2025-01-31o3-mini-2025-01-31-higho3-mini-2025-01-31-lowo3-mini-2025-01-31-medium

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation