Anthropic
·Released on Feb 4, 2026

Claude Opus 4.6 API Benchmarks, Pricing & Provider Data

Compare Claude Opus 4.6 with another model

Choose a model to open its comparison page.

Share on X
LLM

Claude Opus 4.6 benchmark, API pricing, and provider data cover undefined API provider} other undefined API providers}}, with prices starting at $0.015/M. Claude Opus 4.6 free API options are available from undefined provider} other undefined providers}}. The page also shows measured API speed and first-token latency.

Anthropic Claude Opus 4.6 is the most capable Claude Opus tier, optimized for complex analysis, long-horizon coding, and high-stakes enterprise reasoning workloads.

Quality
#42of 112
62.0
LMSpeed score
Speed
45char/s
5.30 s
Cost
#168of 186
$0.015/ 1M · 8:1 in:out
$0.0024 in · $0.013 out

Category Performance

Observed-capability estimates with uncertainty reported separately from independent benchmark families.

Coverage
8 / 8
Methodology
V3.0
Category PerformanceObserved-capability estimates with uncertainty reported separately from independent benchmark families.Agents54.3Coding56.6Reasoning53.4Knowledge53.3Math55.7Multilingual48.4Multimodal47.8Instruction following44.8
#1Coding56.6RatedGlobal rank #103/4 Measured dimensions
80% interval: 48.664.7
code generation49.2
livecodebench · live_code_bench_pro / scicode · aa_sci_code
repository engineering54.2
react_native_bench · react_native_evals / swe_pro · swe_pro / vibecode · vibe_code_bench
debugging testing66.5
swe_rebench · swe_rebench / swe_verified · swe_verified
tooling qualityPrior only
#2Math55.7Provisional1/4 Measured dimensions
80% interval: 39.771.8
foundational mathPrior only
competition mathPrior only
proof frontier55.7
frontiermath · frontier_math_v2_tier4 / frontiermath · frontier_math_v2_tiers13
applied tool mathPrior only
#3Agents54.3RatedGlobal rank #254/4 Measured dimensions
80% interval: 48.560.0
planning56.6
deep_planning · gert_labs
tool use50.4
tau · tau2_bench
environment execution51.6
browsecomp · browse_comp / deep_search_qa · deep_search_qa / osworld · os_world_verified / terminalbench · benchlm_agentic_terminal_bench2
recovery reliability58.5
apex_agents · apex_agents_aa / claw_eval · claw_eval / cyber_gym · cyber_gym / jobbench · job_bench / researchclaw · research_claw_bench
#4Reasoning53.4RatedGlobal rank #293/4 Measured dimensions
80% interval: 44.862.1
abstract logicPrior only
scientific causal57.3
critpt · critpt / gpqa · gpqa / hle · hle / hle · hle_no_tools
multistep constraints53.3
mmlu_pro · mmlu_pro
evidence verification49.7
lcr · lcr
#5Knowledge53.3Provisional1/4 Measured dimensions
80% interval: 39.766.9
broad knowledgePrior only
professional knowledge53.3
healthbench · health_bench_hard / medxpert · med_xpert_qa_text / supergpqa · super_gpqa
factualityPrior only
retrieval open bookPrior only
#6Multilingual48.4Provisional1/4 Measured dimensions
80% interval: 31.365.4
cross language understanding48.4
global_mmlu · aa_global_mmlu_lite
multilingual generationPrior only
reasoning transferPrior only
low resource robustnessPrior only
#7Multimodal47.8Estimated2/4 Measured dimensions
80% interval: 36.658.9
perception ocrPrior only
document spatialPrior only
visual reasoning41.2
erqa · erqa / medxpert · med_xpert_qa_mm / mmmu_pro · mmmu_pro
video action54.3
screenspot · screen_spot_pro
#8Instruction following44.8Provisional1/4 Measured dimensions
80% interval: 28.860.8
constraint followingPrior only
structured outputPrior only
novel instruction generalization44.8
ifbench · aa_if_bench
long multiturn instructionPrior only

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
1Mtokens
1.2K pages of text
OUTPUT
128Ktokens
8K128K1M4M
1M

Features

Technical Details

Input
Output
Released
Feb 2026
Documentation
Tokenizer
Claude
Architecture
text+image+file->text
Moderated
Yes
Supported parameters
include_reasoningmax_completion_tokensmax_tokensreasoningreasoning_effortresponse_formatstopstructured_outputstemperaturetool_choicetoolstop_ktop_pverbosity

Rankings

Excels at

It's decent at

Falls behind in

Detailed scores

Updated: Sep 13, 2026

Overall

undefined metric} other undefined metrics}}

Overall score62.0#42 / 112
Overall score
LMSpeed rank#42
Score62.0
UpdatedSep 13, 2026
Confidence4

Pricing

undefined metric} other undefined metrics}}

Input price$5.00/M#168 / 186Output price$25.00/M#169 / 186
Input price
LMSpeed rank#168
Score$5.00/M
UpdatedSep 13, 2026
Confidence4
Output price
LMSpeed rank#169
Score$25.00/M
UpdatedSep 13, 2026
Confidence4

Agents

V3.0

undefined metric} other undefined metrics}} · Rated

Score54.3#2580% interval48.5–60.04/4 Measured dimensionsAgentic score63.5#37 / 77Terminal-Bench 2.065.4#24 / 53
Dimensions and evidence
Planning & decomposition56.6
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
deep_planning · gert_labs · z 0.89 · q 1.00
Tool use50.4
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
tau · tau2_bench · z -0.05 · q 1.00
Environment & long-horizon execution51.6
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
browsecomp · browse_comp · z 0.21 · q 1.00
deep_search_qa · deep_search_qa · z -0.52 · q 1.00
osworld · os_world_verified · z -0.04 · q 1.00
terminalbench · benchlm_agentic_terminal_bench2 · z 0.10 · q 1.00
Recovery & completion reliability58.5
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
apex_agents · apex_agents_aa · z 0.50 · q 1.00
claw_eval · claw_eval · z 1.87 · q 1.00
cyber_gym · cyber_gym · z -0.99 · q 1.00
jobbench · job_bench · z 0.16 · q 1.00
researchclaw · research_claw_bench · z 0.57 · q 1.00
Agentic score
LMSpeed rank#37
Score63.5
UpdatedSep 13, 2026
Confidence4
Terminal-Bench 2.0V3 evidence
LMSpeed rank#24
Score65.4
UpdatedSep 13, 2026
Confidence4
BrowseCompV3 evidence
LMSpeed rank#11
Score83.7
UpdatedSep 13, 2026
Confidence4
OSWorld-VerifiedV3 evidence
LMSpeed rank#15
Score72.7
UpdatedSep 13, 2026
Confidence4
Τ²-bench resultsV3 evidence
LMSpeed rank#46
Score84.8
UpdatedSep 13, 2026
Confidence4
Claw-EvalV3 evidence
LMSpeed rank#3
Score70.4
UpdatedSep 13, 2026
Confidence4
DeepSearchQAV3 evidence
LMSpeed rank#11
Score73.7
UpdatedSep 13, 2026
Confidence4
CyberGymV3 evidence
LMSpeed rank#11
Score66.6
UpdatedSep 13, 2026
Confidence4
Gert LabsV3 evidence
LMSpeed rank#10
Score61.9
UpdatedSep 13, 2026
Confidence4
ResearchClawBenchV3 evidence
LMSpeed rank#4
Score19.9
UpdatedSep 13, 2026
Confidence4
JobBenchV3 evidence
LMSpeed rank#9
Score36.7
UpdatedSep 13, 2026
Confidence4
ApprenticeBench
LMSpeed rank#16
Score5.0
UpdatedSep 13, 2026
Confidence4
APEX-Agents-AAV3 evidence
LMSpeed rank#8
Score33.0
UpdatedAug 20, 2026
Confidence1

Coding

V3.0

undefined metric} other undefined metrics}} · Rated

Score56.6#1080% interval48.6–64.73/4 Measured dimensionsCoding score50.3#54 / 87SWE-bench Verified80.8#8 / 49
Dimensions and evidence
Code generation49.2
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
livecodebench · live_code_bench_pro · z -0.90 · q 0.50
scicode · aa_sci_code · z 0.24 · q 1.00
Repository engineering54.2
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
react_native_bench · react_native_evals · z 0.80 · q 1.00
swe_pro · swe_pro · z -0.64 · q 1.00
vibecode · vibe_code_bench · z 1.10 · q 1.00
Debugging & testing66.5
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
swe_rebench · swe_rebench · z 2.55 · q 1.00
swe_verified · swe_verified · z 0.85 · q 1.00
Tool-assisted development & qualityPrior only
Coding score
LMSpeed rank#54
Score50.3
UpdatedSep 13, 2026
Confidence4
SWE-bench VerifiedV3 evidence
LMSpeed rank#8
Score80.8
UpdatedSep 13, 2026
Confidence4
SWE-bench Verified*
LMSpeed rank#1
Score75.6
UpdatedSep 13, 2026
Confidence4
LiveCodeBench ProV3 evidence
LMSpeed rank#4
Score70.7
UpdatedSep 13, 2026
Confidence4
SWE-bench ProV3 evidence
LMSpeed rank#41
Score53.4
UpdatedSep 13, 2026
Confidence4
SWE-RebenchV3 evidence
LMSpeed rank#1
Score65.3
UpdatedSep 13, 2026
Confidence4
React Native EvalsV3 evidence
LMSpeed rank#3
Score84.1
UpdatedSep 13, 2026
Confidence4
Vibe Code BenchV3 evidence
LMSpeed rank#5
Score57.6
UpdatedSep 13, 2026
Confidence4
FrontierCode 1.1 Main
LMSpeed rank#10
Score26.9
UpdatedSep 13, 2026
Confidence4
AA-SciCodeV3 evidence
LMSpeed rank#54
Score45.7
UpdatedSep 8, 2026
Confidence4

Reasoning

V3.0

undefined metric} other undefined metrics}} · Rated

Score53.4#2980% interval44.8–62.13/4 Measured dimensionsMMLU-Pro82.0%#50 / 129GPQA84.0%#75 / 218
Dimensions and evidence
Abstract logicPrior only
Scientific & causal reasoning57.3
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
critpt · critpt · z -0.04 · q 1.00
gpqa · gpqa · z 0.84 · q 1.00
hle · hle · z 0.86 · q 1.00
hle · hle_no_tools · z -0.30 · q 1.00
Multi-step constraints53.3
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
mmlu_pro · mmlu_pro · z 0.23 · q 1.00
Evidence integration & verification49.7
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
lcr · lcr · z -0.39 · q 1.00
MMLU-ProV3 evidence
LMSpeed rank#50
Score82.0%
UpdatedSep 13, 2026
Confidence4
GPQAV3 evidence
LMSpeed rank#75
Score84.0%
UpdatedSep 13, 2026
Confidence4
HLEV3 evidence
LMSpeed rank#89
Score19.1%
UpdatedSep 13, 2026
Confidence4
AA-LCRV3 evidence
LMSpeed rank#71
Score67.0
UpdatedSep 13, 2026
Confidence4
CritPtV3 evidence
LMSpeed rank#57
Score2.8
UpdatedSep 13, 2026
Confidence4
Reasoning score
LMSpeed rank#42
Score67.1
UpdatedSep 13, 2026
Confidence4

Knowledge

V3.0

undefined metric} other undefined metrics}} · Provisional

Score53.380% interval39.7–66.91/4 Measured dimensionsKnowledge score79.8#15 / 83GPQA-D89.2#19 / 32
Dimensions and evidence
Broad knowledgePrior only
Professional knowledge53.3
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
healthbench · health_bench_hard · z -1.99 · q 1.00
medxpert · med_xpert_qa_text · z -0.45 · q 0.50
supergpqa · super_gpqa · z 3.00 · q 1.00
FactualityPrior only
Retrieval & open-book usePrior only
Knowledge score
LMSpeed rank#15
Score79.8
UpdatedSep 13, 2026
Confidence4
GPQA-D
LMSpeed rank#19
Score89.2
UpdatedSep 13, 2026
Confidence4
SuperGPQAV3 evidence
LMSpeed rank#1
Score95.0
UpdatedSep 13, 2026
Confidence4
MMLU-Pro (Arcee)
LMSpeed rank#1
Score89.1
UpdatedSep 13, 2026
Confidence4
HLE w/o tools
LMSpeed rank#14
Score40.0
UpdatedSep 13, 2026
Confidence4
HealthBench HardV3 evidence
LMSpeed rank#8
Score14.8
UpdatedSep 13, 2026
Confidence4
MedXpertQA (Text)V3 evidence
LMSpeed rank#3
Score52.1
UpdatedSep 13, 2026
Confidence4
Artificial Analysis Intelligence Index
LMSpeed rank#57
Score26.4
UpdatedSep 13, 2026
Confidence4
AA-GPQA Diamond
LMSpeed rank#67
Score84.0
UpdatedSep 13, 2026
Confidence4
AA-HLE
LMSpeed rank#74
Score19.1
UpdatedSep 13, 2026
Confidence4
AA-Omniscience Index
LMSpeed rank#36
Score2.4
UpdatedSep 13, 2026
Confidence4
AA-Omniscience Accuracy
LMSpeed rank#23
Score45.8
UpdatedSep 13, 2026
Confidence4
AA-Omniscience Hallucination Rate
LMSpeed rank#43
Score80.1
UpdatedSep 13, 2026
Confidence4

Math

V3.0

undefined metric} other undefined metrics}} · Provisional

Score55.780% interval39.7–71.81/4 Measured dimensionsMath score58.6#29 / 63AIME25 (Arcee)99.8#1 / 5
Dimensions and evidence
Foundational mathPrior only
Competition mathPrior only
Advanced proofs55.7
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
frontiermath · frontier_math_v2_tier4 · z 0.79 · q 1.00
frontiermath · frontier_math_v2_tiers13 · z 0.45 · q 1.00
Applied & tool-assisted mathPrior only
Math score
LMSpeed rank#29
Score58.6
UpdatedSep 13, 2026
Confidence4
AIME25 (Arcee)
LMSpeed rank#1
Score99.8
UpdatedSep 13, 2026
Confidence4
FrontierMath v2 (Tiers 1-3)V3 evidence
LMSpeed rank#10
Score40.7
UpdatedSep 13, 2026
Confidence4
FrontierMath v2 (Tier 4)V3 evidence
LMSpeed rank#11
Score22.9
UpdatedSep 13, 2026
Confidence4

Multilingual

V3.0

undefined metric} other undefined metrics}} · Provisional

Score48.480% interval31.3–65.41/4 Measured dimensionsAA Global-MMLU-Lite92.2#3 / 5
Dimensions and evidence
Cross-language understanding48.4
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
global_mmlu · aa_global_mmlu_lite · z 0.00 · q 0.63
Multilingual generationPrior only
Reasoning transferPrior only
Low-resource robustnessPrior only
AA Global-MMLU-LiteV3 evidence
LMSpeed rank#3
Score92.2
UpdatedAug 20, 2026
Confidence1

Multimodal

V3.0

undefined metric} other undefined metrics}} · Estimated

Score47.880% interval36.6–58.92/4 Measured dimensionsMultimodal Grounded score59.5#40 / 56MMMU-Pro77.3#21 / 31
Dimensions and evidence
Perception & OCRPrior only
Document & spatial understandingPrior only
Visual reasoning41.2
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
erqa · erqa · z -1.94 · q 1.00
medxpert · med_xpert_qa_mm · z -0.93 · q 0.75
mmmu_pro · mmmu_pro · z -0.29 · q 1.00
Video & grounded action54.3
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
screenspot · screen_spot_pro · z 0.23 · q 1.00
Multimodal Grounded score
LMSpeed rank#40
Score59.5
UpdatedSep 13, 2026
Confidence4
MMMU-ProV3 evidence
LMSpeed rank#21
Score77.3
UpdatedSep 13, 2026
Confidence4
ERQAV3 evidence
LMSpeed rank#8
Score51.6
UpdatedSep 13, 2026
Confidence4
ScreenSpot ProV3 evidence
LMSpeed rank#6
Score83.1
UpdatedSep 13, 2026
Confidence4
MedXpertQA (MM)V3 evidence
LMSpeed rank#6
Score64.8
UpdatedSep 13, 2026
Confidence4
AA-MMMU-Pro
LMSpeed rank#45
Score72.5
UpdatedSep 13, 2026
Confidence4
Design Arena Website
LMSpeed rank#14
Score1304.0
UpdatedSep 13, 2026
Confidence4

Instruction following

V3.0

undefined metric} other undefined metrics}} · Provisional

Score44.880% interval28.8–60.81/4 Measured dimensionsAA-IFBench44.6#59 / 84Instruction Following score52.6#49 / 52
Dimensions and evidence
Constraint followingPrior only
Structured outputPrior only
Novel-instruction generalization44.8
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
ifbench · aa_if_bench · z -0.98 · q 1.00
Multi-turn & long instructionsPrior only
AA-IFBenchV3 evidence
LMSpeed rank#59
Score44.6
UpdatedSep 13, 2026
Confidence4
Instruction Following score
LMSpeed rank#49
Score52.6
UpdatedSep 13, 2026
Confidence4

OpenRouter endpoints

6 endpoints

Third-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.

Provider endpointInputOutput1d uptime30m latency30m throughputContext / output
Azure
azure/global
$5/M$25/M100%undefined tokens / undefined tokens
Google
google-vertex/europe
$5.50/M$27.50/M100%undefined tokens / undefined tokens
Claude Platform on AWS
claude-on-aws
$5/M$25/M100.0%undefined tokens / undefined tokens
Anthropic
anthropic
$5/M$25/M99.9%undefined tokens / undefined tokens
Amazon Bedrock
amazon-bedrock
$5/M$25/M99.8%undefined tokens / undefined tokens
Google
google-vertex/global
$5/M$25/M99.8%undefined tokens / undefined tokens

Pricing Comparison

Compare Claude Opus 4.6 API pricing across 923 providers. Prices range from $0.015/M to $16900.00/M. 熊猫 API offers the lowest rate at $0.015/M. 15 providers offer free API credits or a free tier.

ProviderHealthModel VariantGroupInput ($/M)Output ($/M)Speed (t/s)First tokenAudit
L1
100%
claude-opus-4-6
default
$6.00/M
Cache read$0.600/MCache write$7.50/MCache write 1h$12.00/M
$30.00/M
112.9 t/s
2.19 s
L1
97%
claude-opus-4-6
活动分组1
-94%$0.286/M
-94%$1.43/M
83.7 t/s
3.92 s
L1
97%
claude-opus-4-6-thinking
活动分组1
-94%$0.286/M
-94%$1.43/M
L1
100%
claude-opus-4-6
Kiro-Claude-1
-97%$0.131/M
Cache read$0.013/M
-97%$0.654/M
45.9 t/s
2.56 s
L1
0%
claude-opus-4-6
kiro低缓
-95%$0.258/M
Cache read$0.026/MCache write$0.322/MCache write 1h$0.515/M
-95%$1.29/M
45.1 t/s
2.28 s
L1
0%
anthropic/claude-opus-4.6
or
-59%$2.06/M
Cache read$0.206/MCache write$2.58/MCache write 1h$4.12/M
-59%$10.30/M
L1
0%
claude-opus-4-6-fast
特价纯血-bq渠道
$6.18/M
Cache read$0.618/MCache write$7.72/MCache write 1h$12.36/M
$30.90/M
L1
100%
claude-opus-4-6
Claude-Entry
-94%$0.300/M
Cache read$0.030/MCache write$0.375/MCache write 1h$0.600/M
-94%$1.50/M
44.3 t/s
2.80 s
L1
100%
claude-opus-4-6
claude带缓存
$10.95/M
Cache read$1.09/MCache write$13.69/MCache write 1h$21.90/M
$54.75/M
43.6 t/s
3.35 s
L1
100%
claude-opus-4-6
default
$5.35/M
Cache read$0.535/MCache write$6.69/MCache write 1h$10.71/M
$26.77/M
42.5 t/s
1.44 s
L1
100%
claude-opus-4-6
vip
-90%$0.500/M
Cache read$0.100/MCache write$1.00/MCache write 1h$1.60/M
-90%$2.50/M
40.5 t/s
16.14 s
L1
100%
claude-opus-4-6-nothinking
claude
-90%$0.500/M
Cache read$0.100/MCache write$1.00/MCache write 1h$1.60/M
-90%$2.50/M
L1
100%
claude-opus-4-6
default
$10.27/M
$51.37/M
40.3 t/s
3.74 s
L1
100%
[按次特价]claude-opus-4-6
default
$0.685/request
-
L1
100%
L2
0%
claude-opus-4-6
Claude_AG
-98%$0.096/M
Cache read$0.0096/M
-98%$0.479/M
38.4 t/s
2.61 s
L1
99%
claude-opus-4-6
default
$10.00/M
Cache read$2.63/MCache write$1.57/MCache write 1h$2.52/M
$50.00/M
37.2 t/s
3.09 s
L1
0%
claude-opus-4-6
default
Free
Free
33.0 t/s
14.80 s
L1
100%
claude-opus-4-6
default
$5.00/M
Cache read$0.500/MCache write$6.25/MCache write 1h$10.00/M
$25.00/M
27.0 t/s
15.40 s
L1
100%
claude-opus-4-6-20260101
default
$5.00/M
Cache read$0.500/MCache write$6.25/MCache write 1h$10.00/M
$25.00/M
L1
100%
claude-opus-4-6-low
claude_kiro
-65%$1.75/M
-65%$8.75/M
Showing 20 model IDs of 249.

Alternatives & Similar Models

Frequently Asked Questions

What benchmark data does Claude Opus 4.6 include?
LMSpeed shows Claude Opus 4.6 benchmark context, API price, output speed, first-token latency, and provider data across 938 providers when those signals are available.
What is the Claude Opus 4.6 API price?
Claude Opus 4.6 has pricing from undefined provider} other undefined providers}}, ranging from $0.015/M to $16900.00/M. 熊猫 API has the lowest listed price.
What does the Claude Opus 4.6 API pricing table include?
The Claude Opus 4.6 API pricing table compares 938 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
Which provider has the cheapest Claude Opus 4.6 API pricing?
熊猫 API currently has the lowest listed Claude Opus 4.6 price at $0.015/M across undefined provider} other undefined providers}}.
Can I compare Claude Opus 4.6 API price and speed together?
Yes. LMSpeed shows Claude Opus 4.6 API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
Is Claude Opus 4.6 API free?
Yes, Claude Opus 4.6 free API options are available through 15 providerundefined other undefined} on LMSpeed, including 我不是AI神, 兔子API, 兔子API, 兔子API, 兔子API. These providers offer free API credits or a free tier with no per-token charges.
Where can I get Claude Opus 4.6 free API access?
LMSpeed currently lists 15 free API providerundefined other undefined} for Claude Opus 4.6: 我不是AI神, 兔子API, 兔子API, 兔子API, 兔子API. Check each provider row before using it because free tier limits can change.

Also known as

26479061/claude-opus-4-63h15pm/claude-opus-4.642API/claude-opus-4-6AWS/claude-opus-4-6Claude Opus 4.6

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation