OpenAI
·Released on Jul 9, 2026

GPT-5.6 Luna API Benchmarks, Pricing & Provider Data

Compare GPT-5.6 Luna with another model

Choose a model to open its comparison page.

Share on X
LLM

GPT-5.6 Luna benchmark, API pricing, and provider data cover 674 API providers, with prices starting at $0.0016/M. GPT-5.6 Luna free API options are available from 5 providers. The page also shows measured API speed and first-token latency.

GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providin...

Quality
#19of 100
70.0
LMSpeed score
Speed
#26of 74
654char/s
6.86 s
Cost
#33of 180
$0.0016/ 1M · 8:1 in:out
$0.0003 in · $0.0013 out

Category Performance

Observed-capability estimates with uncertainty reported separately from independent benchmark families.

Coverage
6 / 8
Methodology
V3.0
Category PerformanceObserved-capability estimates with uncertainty reported separately from independent benchmark families.Agents58.3Coding56.7Reasoning54.5Knowledge56.5Math62.4Multilingual-Multimodal50.3Instruction following-
#1Math62.4Provisional1/4 Measured dimensions
80% interval: 46.378.5
foundational mathPrior only
competition mathPrior only
proof frontier62.4
frontiermath · frontier_math / frontiermath · frontier_math_v2_tier4 / frontiermath · frontier_math_v2_tiers13
applied tool mathPrior only
#2Agents58.3RatedGlobal rank #123/4 Measured dimensions
80% interval: 50.965.7
planningPrior only
tool use54.1
tau · aa_tau3_banking / toolathlon · toolathlon
environment execution62.3
browsecomp · browse_comp / gdpval_aa · benchlm_agentic_gdpval_aa / itbench · aa_itbench / osworld · os_world2 / terminalbench · benchlm_agentic_terminal_bench2
recovery reliability58.5
apex_agents · apex_agents_aa / cyber_gym · cyber_gym / exploit_gym · exploit_gym / harvey_lab · aa_harvey_lab
#3Coding56.7Estimated3/4 Measured dimensions
80% interval: 47.566.0
code generation56.4
scicode · scicode
repository engineering57.8
swe_pro · swe_pro
debugging testingPrior only
tooling quality55.9
terminalbench · aa_terminal_bench21 / terminalbench · benchlm_coding_terminal_bench2
#4Knowledge56.5Provisional1/4 Measured dimensions
80% interval: 39.873.1
broad knowledgePrior only
professional knowledge56.5
healthbench · health_bench_hard
factualityPrior only
retrieval open bookPrior only
#5Reasoning54.5RatedGlobal rank #183/4 Measured dimensions
80% interval: 45.763.2
abstract logic47.5
arc_agi · arc_agi2
scientific causal58.8
critpt · critpt / gpqa · gpqa / hle · hle
multistep constraintsPrior only
evidence verification57.1
lcr · lcr
#6Multimodal50.3Provisional1/4 Measured dimensions
80% interval: 34.266.4
perception ocrPrior only
document spatialPrior only
visual reasoning50.3
mmmu_pro · mmmu_pro
video actionPrior only
No data:MultilingualInstruction following

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
1.1Mtokens
1.3K pages of text
OUTPUT
128Ktokens
8K128K1M4M
1.1M

Features

Technical Details

Input
Output
Released
Jul 2026
Knowledge cutoff
2026-02-16
Documentation
Tokenizer
GPT
Architecture
text+image+file->text
Moderated
Yes
Supported parameters
include_reasoningmax_completion_tokensmax_tokensreasoningreasoning_effortresponse_formatseedstructured_outputstool_choicetools

Rankings

Excels at

Falls behind in

Detailed scores

Updated: Aug 24, 2026

Overall

undefined metric} other undefined metrics}}

Overall score70.0#19 / 100DeepSWE67.2#5 / 16
Overall score
LMSpeed rank#19
Score70.0
UpdatedAug 24, 2026
Confidence3
DeepSWE
LMSpeed rank#5
Score67.2
UpdatedAug 24, 2026
Confidence3
ExploitBench
LMSpeed rank#3
Score33.2
UpdatedAug 24, 2026
Confidence3

Speed & latency

undefined metric} other undefined metrics}}

Output speed132.4 tok/s#26 / 74Time to first token86.96 s#73 / 74
Output speed
LMSpeed rank#26
Score132.4 tok/s
UpdatedAug 24, 2026
Confidence4
Time to first token
LMSpeed rank#73
Score86.96 s
UpdatedAug 24, 2026
Confidence4

Pricing

undefined metric} other undefined metrics}}

Input price$0.200/M#33 / 180Output price$1.20/M#49 / 180
Input price
LMSpeed rank#33
Score$0.200/M
UpdatedAug 24, 2026
Confidence4
Output price
LMSpeed rank#49
Score$1.20/M
UpdatedAug 24, 2026
Confidence4

Agents

V3.0

undefined metric} other undefined metrics}} · Rated

Score58.3#1280% interval50.9–65.73/4 Measured dimensionsAgentic score85.0#8 / 69Terminal-Bench 2.084.7#4 / 53
Dimensions and evidence
Planning & decompositionPrior only
Tool use54.1
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
tau · aa_tau3_banking · z -0.43 · q 1.00
toolathlon · toolathlon · z 0.55 · q 1.00
Environment & long-horizon execution62.3
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
browsecomp · browse_comp · z 0.27 · q 1.00
gdpval_aa · benchlm_agentic_gdpval_aa · z 1.21 · q 1.00
itbench · aa_itbench · z -1.26 · q 1.00
osworld · os_world2 · z 0.89 · q 1.00
terminalbench · benchlm_agentic_terminal_bench2 · z 1.68 · q 1.00
Recovery & completion reliability58.5
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
apex_agents · apex_agents_aa · z 0.65 · q 1.00
cyber_gym · cyber_gym · z 0.12 · q 1.00
exploit_gym · exploit_gym · z -0.05 · q 0.50
harvey_lab · aa_harvey_lab · z -0.35 · q 1.00
Agentic score
LMSpeed rank#8
Score85.0
UpdatedAug 24, 2026
Confidence3
Terminal-Bench 2.0V3 evidence
LMSpeed rank#4
Score84.7
UpdatedAug 24, 2026
Confidence3
BrowseCompV3 evidence
LMSpeed rank#13
Score83.3
UpdatedAug 24, 2026
Confidence3
OSWorld 2.0V3 evidence
LMSpeed rank#5
Score45.6
UpdatedAug 24, 2026
Confidence3
CyberGymV3 evidence
LMSpeed rank#7
Score77.9
UpdatedAug 24, 2026
Confidence3
ExploitGymV3 evidence
LMSpeed rank#4
Score12.4
UpdatedAug 24, 2026
Confidence3
ToolathlonV3 evidence
LMSpeed rank#7
Score53.4
UpdatedAug 24, 2026
Confidence3
AA Agentic Index
LMSpeed rank#15
Score46.9
UpdatedAug 24, 2026
Confidence3
GDPval-AA
LMSpeed rank#11
Score53.9
UpdatedAug 24, 2026
Confidence3
GDPval-AAV3 evidence
LMSpeed rank#11
Score1582.0
UpdatedAug 24, 2026
Confidence3
AA Harvey LABV3 evidence
LMSpeed rank#9
Score87.9
UpdatedAug 24, 2026
Confidence3
AA ITBenchV3 evidence
LMSpeed rank#7
Score40.3
UpdatedAug 24, 2026
Confidence3
AA Tau3 BankingV3 evidence
LMSpeed rank#11
Score31.1
UpdatedAug 24, 2026
Confidence3
AA AutomationBench
LMSpeed rank#11
Score42.2
UpdatedAug 24, 2026
Confidence3
APEX-Agents-AAV3 evidence
LMSpeed rank#5
Score35.8
UpdatedAug 24, 2026
Confidence3
Terminal-Bench 3.0
LMSpeed rank#10
Score14.3
UpdatedAug 24, 2026
Confidence3

Coding

V3.0

undefined metric} other undefined metrics}} · Estimated

Score56.780% interval47.5–66.03/4 Measured dimensionsSciCode52.5%#21 / 204Coding score57.5#37 / 81
Dimensions and evidence
Code generation56.4
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
scicode · scicode · z 0.70 · q 1.00
Repository engineering57.8
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
swe_pro · swe_pro · z 1.10 · q 1.00
Debugging & testingPrior only
Tool-assisted development & quality55.9
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
terminalbench · aa_terminal_bench21 · z 0.06 · q 1.00
terminalbench · benchlm_coding_terminal_bench2 · z 1.29 · q 1.00
SciCodeV3 evidence
LMSpeed rank#21
Score52.5%
UpdatedAug 24, 2026
Confidence4
Coding score
LMSpeed rank#37
Score57.5
UpdatedAug 24, 2026
Confidence3
SWE-bench ProV3 evidence
LMSpeed rank#10
Score62.7
UpdatedAug 24, 2026
Confidence3
Terminal-Bench 2.0V3 evidence
LMSpeed rank#3
Score84.7
UpdatedAug 24, 2026
Confidence3
FrontierCode 1.1 Extended
LMSpeed rank#5
Score55.1
UpdatedAug 24, 2026
Confidence3
CursorBench v3.2
LMSpeed rank#9
Score61.1
UpdatedAug 24, 2026
Confidence3
AA Coding Index
LMSpeed rank#15
Score71.5
UpdatedAug 24, 2026
Confidence3
AA-SciCodeV3 evidence
LMSpeed rank#24
Score52.5
UpdatedAug 24, 2026
Confidence3
AA Terminal-Bench 2.1V3 evidence
LMSpeed rank#8
Score80.9
UpdatedAug 24, 2026
Confidence3

Reasoning

V3.0

undefined metric} other undefined metrics}} · Rated

Score54.5#1880% interval45.7–63.23/4 Measured dimensionsGPQA91.1%#21 / 211HLE39.5%#27 / 208
Dimensions and evidence
Abstract logic47.5
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
arc_agi · arc_agi2 · z -0.51 · q 1.00
Scientific & causal reasoning58.8
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
critpt · critpt · z 0.90 · q 1.00
gpqa · gpqa · z 0.57 · q 1.00
hle · hle · z 0.51 · q 1.00
Multi-step constraintsPrior only
Evidence integration & verification57.1
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
lcr · lcr · z 0.59 · q 1.00
GPQAV3 evidence
LMSpeed rank#21
Score91.1%
UpdatedAug 24, 2026
Confidence4
HLEV3 evidence
LMSpeed rank#27
Score39.5%
UpdatedAug 24, 2026
Confidence4
ARC-AGI-3
LMSpeed rank#9
Score0.2
UpdatedAug 24, 2026
Confidence3
AA-LCRV3 evidence
LMSpeed rank#13
Score78.3
UpdatedAug 24, 2026
Confidence3
CritPtV3 evidence
LMSpeed rank#12
Score20.6
UpdatedAug 24, 2026
Confidence3
Reasoning score
LMSpeed rank#23
Score57.8
UpdatedAug 24, 2026
Confidence3
ARC-AGI-2V3 evidence
LMSpeed rank#11
Score59.5
UpdatedAug 24, 2026
Confidence3

Knowledge

V3.0

undefined metric} other undefined metrics}} · Provisional

Score56.580% interval39.8–73.11/4 Measured dimensionsKnowledge score81.5#14 / 69GPQA-D92.3#12 / 31
Dimensions and evidence
Broad knowledgePrior only
Professional knowledge56.5
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
healthbench · health_bench_hard · z 0.00 · q 0.88
FactualityPrior only
Retrieval & open-book usePrior only
Knowledge score
LMSpeed rank#14
Score81.5
UpdatedAug 24, 2026
Confidence3
GPQA-D
LMSpeed rank#12
Score92.3
UpdatedAug 24, 2026
Confidence3
HealthBench Professional
LMSpeed rank#5
Score55.7
UpdatedAug 24, 2026
Confidence3
HealthBench HardV3 evidence
LMSpeed rank#4
Score32.0
UpdatedAug 24, 2026
Confidence3
Artificial Analysis Intelligence Index
LMSpeed rank#21
Score51.2
UpdatedAug 24, 2026
Confidence3
AA-GPQA Diamond
LMSpeed rank#21
Score91.1
UpdatedAug 24, 2026
Confidence3
AA-HLE
LMSpeed rank#25
Score39.5
UpdatedAug 24, 2026
Confidence3
AA-Omniscience Accuracy
LMSpeed rank#25
Score42.7
UpdatedAug 24, 2026
Confidence3
AA-Omniscience Hallucination Rate
LMSpeed rank#10
Score92.6
UpdatedAug 24, 2026
Confidence3

Math

V3.0

undefined metric} other undefined metrics}} · Provisional

Score62.480% interval46.3–78.51/4 Measured dimensionsMath score96.9#1 / 62FrontierMath (legacy)78.6#3 / 7
Dimensions and evidence
Foundational mathPrior only
Competition mathPrior only
Advanced proofs62.4
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
frontiermath · frontier_math · z 1.10 · q 0.58
frontiermath · frontier_math_v2_tier4 · z 1.73 · q 1.00
frontiermath · frontier_math_v2_tiers13 · z 1.42 · q 1.00
Applied & tool-assisted mathPrior only
Math score
LMSpeed rank#1
Score96.9
UpdatedAug 24, 2026
Confidence3
FrontierMath (legacy)V3 evidence
LMSpeed rank#3
Score78.6
UpdatedAug 24, 2026
Confidence3
FrontierMath v2 (Tiers 1-3)V3 evidence
LMSpeed rank#3
Score78.6
UpdatedAug 24, 2026
Confidence3
FrontierMath v2 (Tier 4)V3 evidence
LMSpeed rank#3
Score58.5
UpdatedAug 24, 2026
Confidence3

Multilingual

V3.0

undefined metric} other undefined metrics}} · No data

No data80% interval30.8–69.20/4 Measured dimensions
Dimensions and evidence
Cross-language understandingPrior only
Multilingual generationPrior only
Reasoning transferPrior only
Low-resource robustnessPrior only

Multimodal

V3.0

undefined metric} other undefined metrics}} · Provisional

Score50.380% interval34.2–66.41/4 Measured dimensionsMultimodal Grounded score66.0#24 / 44MMMU-Pro78.4#17 / 31
Dimensions and evidence
Perception & OCRPrior only
Document & spatial understandingPrior only
Visual reasoning50.3
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
mmmu_pro · mmmu_pro · z -0.02 · q 1.00
Video & grounded actionPrior only
Multimodal Grounded score
LMSpeed rank#24
Score66.0
UpdatedAug 24, 2026
Confidence3
MMMU-ProV3 evidence
LMSpeed rank#17
Score78.4
UpdatedAug 24, 2026
Confidence3
MMMU-Pro w/ Python
LMSpeed rank#7
Score79.5
UpdatedAug 24, 2026
Confidence3
AA-MMMU-Pro
LMSpeed rank#17
Score78.6
UpdatedAug 24, 2026
Confidence3

Instruction following

V3.0

undefined metric} other undefined metrics}} · No data

No data80% interval30.8–69.20/4 Measured dimensions
Dimensions and evidence
Constraint followingPrior only
Structured outputPrior only
Novel-instruction generalizationPrior only
Multi-turn & long instructionsPrior only

OpenRouter endpoints

7 endpoints

Third-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.

Provider endpointInputOutput1d uptime30m latency30m throughputContext / output
Amazon Bedrock
amazon-bedrock/us-east-1
$0.220/M$1.32/M100%undefined tokens / undefined tokens
Azure
azure/eu
$0.220/M$1.32/M100.0%undefined tokens / undefined tokens
OpenAI
openai/flex
$0.100/M$0.600/M98.9%undefined tokens / undefined tokens
OpenAI
openai
$0.200/M$1.20/M98.9%undefined tokens / undefined tokens
OpenAI
openai/priority
$0.400/M$2.40/M98.9%undefined tokens / undefined tokens
Azure
azure
$0.200/M$1.20/M94.6%undefined tokens / undefined tokens
Azure
azure/us
$0.220/M$1.32/M89.7%undefined tokens / undefined tokens

Pricing Comparison

Compare GPT-5.6 Luna API pricing across 669 providers. Prices range from $0.0016/M to $6993.00/M. 熊猫 API offers the lowest rate at $0.0016/M. 5 providers offer free API credits or a free tier.

ProviderHealthModel VariantGroupInput ($/M)Output ($/M)Speed (t/s)First tokenAudit
L1
100%
gpt-5.6-luna
default
$0.029/request
-
150.9 t/s+14%
3.07 s-96%
L1
100%
gpt-5.6-luna
Codex
$0.250/M
Cache read$0.025/MCache write$0.313/MCache write 1h$0.500/M
$1.50/M
40.0 t/s
13.90 s-84%
L1
100%
gpt-5.6-luna
codex
-64%$0.071/M
Cache read$0.0071/M
-64%$0.429/M
L1
100%
gpt-5.6-luna
default
-66%$0.068/M
Cache read$0.0068/M
-66%$0.411/M
L1
99%
gpt-5.6-luna
default
$1.00/M
$6.00/M
L1
99%
gpt-5.6-luna
Codex-Plus
-97%$0.0059/M
Cache read$0.0006/M
-96%$0.047/M
L1
100%
gpt-5.6-luna
GPT-Entry
-85%$0.030/M
Cache read$0.0030/M
-85%$0.180/M
L1
100%
gpt-5.6-luna
default
-32%$0.137/M
-32%$0.822/M
L1
100%
gpt-5.6-luna-medium
CodeX_05_token
-66%$0.068/M
-66%$0.411/M
L1
100%
gpt-5.6-luna-high
CodeX_05_token
-66%$0.068/M
-66%$0.411/M
L1
100%
gpt-5.6-luna-high-openai-compact
CodeX_05_token
-66%$0.068/M
-66%$0.411/M
L1
100%
gpt-5.6-luna-low
CodeX_05_token
-66%$0.068/M
-66%$0.411/M
L1
100%
gpt-5.6-luna-low-openai-compact
CodeX_05_token
-66%$0.068/M
-66%$0.411/M
L1
100%
gpt-5.6-luna-medium-openai-compact
CodeX_05_token
-66%$0.068/M
-66%$0.411/M
L1
100%
gpt-5.6-luna-openai-compact
CodeX_05_token
-66%$0.068/M
-66%$0.411/M
L1
100%
gpt-5.6-luna-xhigh
CodeX_05_token
-66%$0.068/M
-66%$0.411/M
L1
100%
gpt-5.6-luna-xhigh-openai-compact
CodeX_05_token
-66%$0.068/M
-66%$0.411/M
L1
100%
gpt-5.6-luna-2026-07-09
default
-32%$0.137/M
-32%$0.822/M
L1
100%
gpt-5.6-luna
discount-codex
-75%$0.050/M
Cache read$0.0050/MCache write$0.063/MCache write 1h$0.100/M
-75%$0.300/M
L1
100%
gpt-5.6-luna
gpt-精品-个人版
-48%$0.103/M
-31%$0.827/M
Showing 20 model IDs of 173.

Alternatives & Similar Models

Frequently Asked Questions

What benchmark data does GPT-5.6 Luna include?
LMSpeed shows GPT-5.6 Luna benchmark context, API price, output speed, first-token latency, and provider data across 674 providers when those signals are available.
What is the GPT-5.6 Luna API price?
GPT-5.6 Luna has pricing from 674 providers, ranging from $0.0016/M to $6993.00/M. 熊猫 API has the lowest listed price.
What does the GPT-5.6 Luna API pricing table include?
The GPT-5.6 Luna API pricing table compares 674 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
Which provider has the cheapest GPT-5.6 Luna API pricing?
熊猫 API currently has the lowest listed GPT-5.6 Luna price at $0.0016/M across 674 providers.
Can I compare GPT-5.6 Luna API price and speed together?
Yes. LMSpeed shows GPT-5.6 Luna API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
Is GPT-5.6 Luna API free?
Yes, GPT-5.6 Luna free API options are available through 5 providerundefined other undefined} on LMSpeed, including 兔子API, Moyanjdc API, APIMart, 兔子API, 兔子API. These providers offer free API credits or a free tier with no per-token charges.
Where can I get GPT-5.6 Luna free API access?
LMSpeed currently lists 5 free API providerundefined other undefined} for GPT-5.6 Luna: 兔子API, Moyanjdc API, APIMart, 兔子API, 兔子API. Check each provider row before using it because free tier limits can change.

Also known as

[lq]q|XXQ/gpt-5.6-lunagpt-5-6-lunagpt-5.6-lunagpt-5.6-luna-2026-07-09gpt-5.6-luna-high

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation