OpenAI
·Released on Mar 17, 2026

GPT-5.4 Mini API Benchmarks, Pricing & Provider Data

Compare GPT-5.4 Mini with another model

Choose a model to open its comparison page.

Share on X
LLM

GPT-5.4 Mini benchmark, API pricing, and provider data cover 959 API providers, with prices starting at $0.0030/request. GPT-5.4 Mini free API options are available from 5 providers. The page also shows measured API speed and first-token latency.

OpenAI GPT-5.4 Mini is a compact language model in the GPT-5 series, optimized for quick responses and high throughput.

Quality
#55of 100
55.0
LMSpeed score
Speed
134char/s
3.54 s
Cost
#104of 180
$0.0030/ 1M · 8:1 in:out
$0.0005 in · $0.0025 out

Category Performance

Observed-capability estimates with uncertainty reported separately from independent benchmark families.

Coverage
7 / 8
Methodology
V3.0
Category PerformanceObserved-capability estimates with uncertainty reported separately from independent benchmark families.Agents51Coding57.5Reasoning55.6Knowledge46.1Math49.7Multilingual-Multimodal47.1Instruction following53.6
#1Coding57.5Estimated2/4 Measured dimensions
80% interval: 45.669.3
code generation58.9
scicode · scicode
repository engineering56
vibecode · vibe_code_bench
debugging testingPrior only
tooling qualityPrior only
#2Reasoning55.6Estimated2/4 Measured dimensions
80% interval: 44.866.4
abstract logicPrior only
scientific causal57.1
critpt · critpt / gpqa · gpqa / hle · hle / hle · hle_no_tools
multistep constraintsPrior only
evidence verification54.2
lcr · lcr
#3Instruction following53.6Provisional1/4 Measured dimensions
80% interval: 37.669.6
constraint followingPrior only
structured outputPrior only
novel instruction generalization53.6
ifbench · aa_if_bench
long multiturn instructionPrior only
#4Agents51RatedGlobal rank #323/4 Measured dimensions
80% interval: 42.959.2
planningPrior only
tool use48.4
mcp_atlas · mcp_atlas / tau · tau2_bench / toolathlon · toolathlon
environment execution50.9
gdpval_aa · benchlm_agentic_gdpval_aa / osworld · os_world_verified / terminalbench · benchlm_agentic_terminal_bench2
recovery reliability53.8
apex_agents · apex_agents_aa
#5Math49.7Provisional1/4 Measured dimensions
80% interval: 33.665.7
foundational mathPrior only
competition mathPrior only
proof frontier49.7
frontiermath · frontier_math_v2_tier4 / frontiermath · frontier_math_v2_tiers13
applied tool mathPrior only
#6Multimodal47.1Provisional1/4 Measured dimensions
80% interval: 31.063.3
perception ocrPrior only
document spatialPrior only
visual reasoning47.1
mmmu_pro · mmmu_pro
video actionPrior only
#7Knowledge46.1Provisional1/4 Measured dimensions
80% interval: 32.160.1
broad knowledge46.1
aa_omniscience · aa_omniscience_index / benchlm_category_knowledge · benchlm_category_knowledge
professional knowledgePrior only
factualityPrior only
retrieval open bookPrior only
No data:Multilingual

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
400Ktokens
480 pages of text
OUTPUT
128Ktokens
8K128K1M4M
400K

Features

Technical Details

Input
Output
Released
Mar 2026
Knowledge cutoff
2025-08-31
Documentation
Tokenizer
GPT
Architecture
text+image+file->text
Moderated
Yes
Supported parameters
include_reasoningmax_completion_tokensmax_tokensreasoningreasoning_effortresponse_formatseedstructured_outputstool_choicetools

Rankings

Excels at

It's decent at

Falls behind in

Detailed scores

Updated: Aug 31, 2026

Overall

undefined metric} other undefined metrics}}

Overall score55.0#55 / 100
Overall score
LMSpeed rank#55
Score55.0
UpdatedAug 31, 2026
Confidence3

Pricing

undefined metric} other undefined metrics}}

Input price$0.750/M#104 / 180Output price$4.50/M#121 / 180
Input price
LMSpeed rank#104
Score$0.750/M
UpdatedAug 31, 2026
Confidence4
Output price
LMSpeed rank#121
Score$4.50/M
UpdatedAug 31, 2026
Confidence4

Agents

V3.0

undefined metric} other undefined metrics}} · Rated

Score51#3280% interval42.9–59.23/4 Measured dimensionsAgentic score55.6#41 / 69Terminal-Bench 2.060.0#31 / 53
Dimensions and evidence
Planning & decompositionPrior only
Tool use48.4
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
mcp_atlas · mcp_atlas · z -1.22 · q 1.00
tau · tau2_bench · z 0.45 · q 1.00
toolathlon · toolathlon · z -0.52 · q 1.00
Environment & long-horizon execution50.9
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
gdpval_aa · benchlm_agentic_gdpval_aa · z 0.00 · q 1.00
osworld · os_world_verified · z -0.09 · q 1.00
terminalbench · benchlm_agentic_terminal_bench2 · z -0.24 · q 1.00
Recovery & completion reliability53.8
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
apex_agents · apex_agents_aa · z 0.21 · q 1.00
Agentic score
LMSpeed rank#41
Score55.6
UpdatedAug 31, 2026
Confidence3
Terminal-Bench 2.0V3 evidence
LMSpeed rank#31
Score60.0
UpdatedAug 31, 2026
Confidence3
OSWorld-VerifiedV3 evidence
LMSpeed rank#16
Score72.1
UpdatedAug 31, 2026
Confidence3
MCP AtlasV3 evidence
LMSpeed rank#24
Score57.7
UpdatedAug 31, 2026
Confidence3
ToolathlonV3 evidence
LMSpeed rank#18
Score42.9
UpdatedAug 31, 2026
Confidence3
Τ²-bench resultsV3 evidence
LMSpeed rank#29
Score93.4
UpdatedAug 31, 2026
Confidence3
AA Agentic Index
LMSpeed rank#26
Score31.5
UpdatedAug 31, 2026
Confidence3
APEX-Agents-AAV3 evidence
LMSpeed rank#11
Score28.2
UpdatedAug 31, 2026
Confidence3
GDPval-AA
LMSpeed rank#34
Score33.4
UpdatedAug 31, 2026
Confidence3
GDPval-AAV3 evidence
LMSpeed rank#33
Score1169.0
UpdatedAug 31, 2026
Confidence3

Coding

V3.0

undefined metric} other undefined metrics}} · Estimated

Score57.580% interval45.6–69.32/4 Measured dimensionsSciCode44.2%#51 / 205Coding score52.7#44 / 81
Dimensions and evidence
Code generation58.9
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
scicode · scicode · z 1.06 · q 1.00
Repository engineering56
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
vibecode · vibe_code_bench · z 0.79 · q 1.00
Debugging & testingPrior only
Tool-assisted development & qualityPrior only
SciCodeV3 evidence
LMSpeed rank#51
Score44.2%
UpdatedAug 31, 2026
Confidence4
Coding score
LMSpeed rank#44
Score52.7
UpdatedAug 31, 2026
Confidence3
Vibe Code BenchV3 evidence
LMSpeed rank#10
Score48.0
UpdatedAug 31, 2026
Confidence3
AA Coding Index
LMSpeed rank#31
Score56.1
UpdatedAug 31, 2026
Confidence3
AA-SciCodeV3 evidence
LMSpeed rank#29
Score49.9
UpdatedAug 31, 2026
Confidence3
FrontierCode 1.1 Main
LMSpeed rank#8
Score27.0
UpdatedAug 31, 2026
Confidence3
EEBench
LMSpeed rank#16
Score18.3
UpdatedAug 20, 2026
Confidence3

Reasoning

V3.0

undefined metric} other undefined metrics}} · Estimated

Score55.680% interval44.8–66.42/4 Measured dimensionsGPQA82.3%#82 / 212HLE18.6%#85 / 209
Dimensions and evidence
Abstract logicPrior only
Scientific & causal reasoning57.1
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
critpt · critpt · z 0.55 · q 1.00
gpqa · gpqa · z 0.69 · q 1.00
hle · hle · z 0.58 · q 1.00
hle · hle_no_tools · z -1.54 · q 1.00
Multi-step constraintsPrior only
Evidence integration & verification54.2
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
lcr · lcr · z 0.20 · q 1.00
GPQAV3 evidence
LMSpeed rank#82
Score82.3%
UpdatedAug 31, 2026
Confidence4
HLEV3 evidence
LMSpeed rank#85
Score18.6%
UpdatedAug 31, 2026
Confidence4
AA-LCRV3 evidence
LMSpeed rank#39
Score73.0
UpdatedAug 31, 2026
Confidence3
CritPtV3 evidence
LMSpeed rank#29
Score10.0
UpdatedAug 31, 2026
Confidence3

Knowledge

V3.0

undefined metric} other undefined metrics}} · Provisional

Score46.180% interval32.1–60.11/4 Measured dimensionsKnowledge score57.3#52 / 69HLE w/o tools28.2#20 / 21
Dimensions and evidence
Broad knowledge46.1
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
aa_omniscience · aa_omniscience_index · z -0.50 · q 1.00
benchlm_category_knowledge · benchlm_category_knowledge · z -0.61 · q 1.00
Professional knowledgePrior only
FactualityPrior only
Retrieval & open-book usePrior only
Knowledge scoreV3 evidence
LMSpeed rank#52
Score57.3
UpdatedAug 31, 2026
Confidence3
HLE w/o tools
LMSpeed rank#20
Score28.2
UpdatedAug 31, 2026
Confidence3
Artificial Analysis Intelligence IndexV3 evidence
LMSpeed rank#40
Score40.9
UpdatedAug 31, 2026
Confidence3
AA-GPQA Diamond
LMSpeed rank#42
Score87.5
UpdatedAug 31, 2026
Confidence3
AA-HLE
LMSpeed rank#50
Score28.1
UpdatedAug 31, 2026
Confidence3
AA-Omniscience Accuracy
LMSpeed rank#40
Score37.5
UpdatedAug 31, 2026
Confidence3
AA-Omniscience Hallucination Rate
LMSpeed rank#18
Score90.2
UpdatedAug 31, 2026
Confidence3

Math

V3.0

undefined metric} other undefined metrics}} · Provisional

Score49.780% interval33.6–65.71/4 Measured dimensionsMath score44.6#38 / 62FrontierMath v2 (Tiers 1-3)28.3#20 / 47
Dimensions and evidence
Foundational mathPrior only
Competition mathPrior only
Advanced proofs49.7
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
frontiermath · frontier_math_v2_tier4 · z -0.65 · q 1.00
frontiermath · frontier_math_v2_tiers13 · z 0.13 · q 1.00
Applied & tool-assisted mathPrior only
Math score
LMSpeed rank#38
Score44.6
UpdatedAug 31, 2026
Confidence3
FrontierMath v2 (Tiers 1-3)V3 evidence
LMSpeed rank#20
Score28.3
UpdatedAug 31, 2026
Confidence3
FrontierMath v2 (Tier 4)V3 evidence
LMSpeed rank#35
Score2.1
UpdatedAug 31, 2026
Confidence3

Multilingual

V3.0

undefined metric} other undefined metrics}} · No data

No data80% interval30.8–69.20/4 Measured dimensions
Dimensions and evidence
Cross-language understandingPrior only
Multilingual generationPrior only
Reasoning transferPrior only
Low-resource robustnessPrior only

Multimodal

V3.0

undefined metric} other undefined metrics}} · Provisional

Score47.180% interval31.0–63.31/4 Measured dimensionsMultimodal Grounded score51.3#32 / 44MMMU-Pro76.6#22 / 31
Dimensions and evidence
Perception & OCRPrior only
Document & spatial understandingPrior only
Visual reasoning47.1
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
mmmu_pro · mmmu_pro · z -0.45 · q 1.00
Video & grounded actionPrior only
Multimodal Grounded score
LMSpeed rank#32
Score51.3
UpdatedAug 31, 2026
Confidence3
MMMU-ProV3 evidence
LMSpeed rank#22
Score76.6
UpdatedAug 31, 2026
Confidence3
MMMU-Pro w/ Python
LMSpeed rank#8
Score78.0
UpdatedAug 31, 2026
Confidence3
AA-MMMU-Pro
LMSpeed rank#40
Score73.3
UpdatedAug 31, 2026
Confidence3

Instruction following

V3.0

undefined metric} other undefined metrics}} · Provisional

Score53.680% interval37.6–69.61/4 Measured dimensionsAA-IFBench73.3#24 / 84
Dimensions and evidence
Constraint followingPrior only
Structured outputPrior only
Novel-instruction generalization53.6
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
ifbench · aa_if_bench · z 0.21 · q 1.00
Multi-turn & long instructionsPrior only
AA-IFBenchV3 evidence
LMSpeed rank#24
Score73.3
UpdatedAug 31, 2026
Confidence3

OpenRouter endpoints

5 endpoints

Third-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.

Provider endpointInputOutput1d uptime30m latency30m throughputContext / output
OpenAI
openai/flex
$0.375/M$2.25/M100%undefined tokens / undefined tokens
Azure
azure
$0.750/M$4.50/M100.0%undefined tokens / undefined tokens
OpenAI
openai
$0.750/M$4.50/M99.8%undefined tokens / undefined tokens
OpenAI
openai/fast
$1.50/M$9/M99.4%undefined tokens / undefined tokens
Azure
azure/us
$0.825/M$4.95/Mundefined tokens / undefined tokens

Pricing Comparison

Compare GPT-5.4 Mini API pricing across 954 providers. Prices range from $0.0030/request to $7492.50/M. Crond offers the lowest rate at $0.0030/request. 5 providers offer free API credits or a free tier.

ProviderHealthModel VariantGroupInput ($/M)Output ($/M)Speed (t/s)First tokenAudit
L1
0%
gpt-5.4-mini
default
$0.750/M
Cache read$0.075/M
$4.50/M
245.0 t/s
2.62 s
L1
100%
gpt-5.4-mini
default
-93%$0.051/M
Cache read$0.0051/M
-93%$0.308/M
200.7 t/s
1.44 s
L1
100%
gpt-5.4-mini-2026-03-17
default
-93%$0.051/M
-93%$0.308/M
L1
99%
gpt-5.4-mini
default
$0.750/M
Cache read$0.075/M
$4.50/M
139.4 t/s
2.84 s
L1
100%
gpt-5.4-mini
default
$0.750/M
Cache read$0.075/M
$4.50/M
97.9 t/s
4.04 s
L1
100%
gpt-5.4-mini
GPT-Entry
-97%$0.022/M
Cache read$0.0022/M
-97%$0.135/M
94.8 t/s
2.34 s
L1
100%
gpt-5.4-mini
codex
-93%$0.054/M
Cache read$0.0054/M
-93%$0.321/M
L1
100%
gpt-5.4-mini
GPT · Best cash equivalent (conditional promotion)
-95%$0.037/M
Cache read$0.0037/M
-95%$0.225/M
L1
100%
gpt-5.4-mini
default
$0.750/M
$4.50/M
L1
100%
gpt-5.4-mini-xhigh
default
$0.750/M
$4.50/M
L1
100%
gpt-5.4-mini-medium
default
$0.750/M
$4.50/M
L1
100%
gpt-5.4-mini-low
default
$0.750/M
$4.50/M
L1
100%
gpt-5.4-mini-high
default
$0.750/M
$4.50/M
L1
100%
gpt-5.4-mini-2026-03-17
default
$0.750/M
$4.50/M
L1
99%
gpt-5.4-mini
Codex-Plus
-99%$0.0044/M
Cache read$0.0004/MCache write$0.0055/MCache write 1h$0.0088/M
-99%$0.026/M
L1
100%
gpt-5.4-mini
default
-86%$0.103/M
-86%$0.616/M
L1
100%
gpt-5.4-mini-2026-03-17
default
-86%$0.103/M
-86%$0.616/M
L1
100%
gpt-5.4-mini-high
default
-86%$0.103/M
-86%$0.616/M
L1
100%
gpt-5.4-mini-medium
default
-86%$0.103/M
-86%$0.616/M
L1
100%
gpt-5.4-mini
default
-86%$0.103/M
Cache read$0.010/M
-86%$0.616/M
Showing 20 model IDs of 170.

Alternatives & Similar Models

Frequently Asked Questions

What benchmark data does GPT-5.4 Mini include?
LMSpeed shows GPT-5.4 Mini benchmark context, API price, output speed, first-token latency, and provider data across 959 providers when those signals are available.
What is the GPT-5.4 Mini API price?
GPT-5.4 Mini has pricing from 959 providers, ranging from $0.0030/request to $7492.50/M. Crond has the lowest listed price.
What does the GPT-5.4 Mini API pricing table include?
The GPT-5.4 Mini API pricing table compares 959 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
Which provider has the cheapest GPT-5.4 Mini API pricing?
Crond currently has the lowest listed GPT-5.4 Mini price at $0.0030/request across 959 providers.
Can I compare GPT-5.4 Mini API price and speed together?
Yes. LMSpeed shows GPT-5.4 Mini API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
Is GPT-5.4 Mini API free?
Yes, GPT-5.4 Mini free API options are available through 5 providerundefined other undefined} on LMSpeed, including 兔子API, 兔子API, Moyanjdc API, 兔子API, 兔子API. These providers offer free API credits or a free tier with no per-token charges.
Where can I get GPT-5.4 Mini free API access?
LMSpeed currently lists 5 free API providerundefined other undefined} for GPT-5.4 Mini: 兔子API, 兔子API, Moyanjdc API, 兔子API, 兔子API. Check each provider row before using it because free tier limits can change.

Also known as

11/gpt-5.4-mini15/gpt-5.4-mini64/gpt-5.4-miniCPA/gpt-5.4-miniGPT-5.4 Mini

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation