SpaceXAI
·Released on Apr 30, 2026

Grok 4.3 API Benchmarks, Pricing & Provider Data

Compare Grok 4.3 with another model

Choose a model to open its comparison page.

Share on X
LLM

Grok 4.3 benchmark, API pricing, and provider data cover undefined API provider} other undefined API providers}}, with prices starting at $0.010/request. Grok 4.3 free API options are available from undefined provider} other undefined providers}}. The page also shows measured API speed and first-token latency.

Grok 4.3 is a reasoning model from xAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...

Quality
#61of 112
58.0
LMSpeed score
Speed
#38of 80
49char/s
6.49 s
Cost
#123of 186
$0.010/ 1M · 8:1 in:out
$0.0016 in · $0.0084 out

Category Performance

Observed-capability estimates with uncertainty reported separately from independent benchmark families.

Coverage
6 / 8
Methodology
V3.0
Category PerformanceObserved-capability estimates with uncertainty reported separately from independent benchmark families.Agents49Coding55Reasoning53.3Knowledge51.1Math-Multilingual-Multimodal49.8Instruction following57.1
#1Instruction following57.1Provisional1/4 Measured dimensions
80% interval: 41.173.1
constraint followingPrior only
structured outputPrior only
novel instruction generalization57.1
ifbench · aa_if_bench
long multiturn instructionPrior only
#2Coding55Provisional1/4 Measured dimensions
80% interval: 39.071.0
code generation55
scicode · scicode
repository engineeringPrior only
debugging testingPrior only
tooling qualityPrior only
#3Reasoning53.3Estimated2/4 Measured dimensions
80% interval: 42.564.1
abstract logicPrior only
scientific causal58
critpt · critpt / gpqa · gpqa / hle · hle
multistep constraintsPrior only
evidence verification48.6
lcr · lcr
#4Knowledge51.1Provisional1/4 Measured dimensions
80% interval: 37.165.1
broad knowledge51.1
aa_omniscience · aa_omniscience_index / benchlm_category_knowledge · benchlm_category_knowledge
professional knowledgePrior only
factualityPrior only
retrieval open bookPrior only
#5Multimodal49.8Provisional1/4 Measured dimensions
80% interval: 33.765.9
perception ocrPrior only
document spatialPrior only
visual reasoning49.8
mmmu_pro · mmmu_pro
video actionPrior only
#6Agents49RatedGlobal rank #404/4 Measured dimensions
80% interval: 42.455.6
planning47.4
deep_planning · gert_labs
tool use58.8
tau · tau2_bench
environment execution49.5
gdpval_aa · benchlm_agentic_gdpval_aa
recovery reliability40.3
apex_agents · apex_agents_aa / researchclaw · research_claw_bench
No data:MathMultilingual

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
1Mtokens
1.2K pages of text
OUTPUT
900Ktokens
8K128K1M4M
1M

Features

Technical Details

Input
Output
Released
Apr 2026
Documentation
Tokenizer
Grok
Architecture
text+image+file->text
Moderated
No
Supported parameters
include_reasoninglogprobsmax_tokensreasoningreasoning_effortresponse_formatseedstructured_outputstemperaturetool_choicetoolstop_logprobstop_p

Rankings

Excels at

It's decent at

Falls behind in

Detailed scores

Updated: Sep 13, 2026

Overall

undefined metric} other undefined metrics}}

Overall score58.0#61 / 112
Overall score
LMSpeed rank#61
Score58.0
UpdatedAug 20, 2026
Confidence1

Pricing

undefined metric} other undefined metrics}}

Input price$1.25/M#123 / 186Output price$2.50/M#89 / 186
Input price
LMSpeed rank#123
Score$1.25/M
UpdatedSep 13, 2026
Confidence4
Output price
LMSpeed rank#89
Score$2.50/M
UpdatedSep 13, 2026
Confidence4

Agents

V3.0

undefined metric} other undefined metrics}} · Rated

Score49#4080% interval42.4–55.64/4 Measured dimensionsΤ²-bench results97.7#9 / 82GDPval-AA29.2#42 / 66
Dimensions and evidence
Planning & decomposition47.4
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
deep_planning · gert_labs · z -0.34 · q 1.00
Tool use58.8
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
tau · tau2_bench · z 1.08 · q 1.00
Environment & long-horizon execution49.5
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
gdpval_aa · benchlm_agentic_gdpval_aa · z -0.19 · q 1.00
Recovery & completion reliability40.3
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
apex_agents · apex_agents_aa · z -0.60 · q 1.00
researchclaw · research_claw_bench · z -2.01 · q 1.00
Τ²-bench resultsV3 evidence
LMSpeed rank#9
Score97.7
UpdatedAug 20, 2026
Confidence1
GDPval-AA
LMSpeed rank#42
Score29.2
UpdatedAug 20, 2026
Confidence1
AA Agentic Index
LMSpeed rank#39
Score24.2
UpdatedAug 20, 2026
Confidence1
APEX-Agents-AAV3 evidence
LMSpeed rank#15
Score17.0
UpdatedAug 20, 2026
Confidence1
GDPval-AAV3 evidence
LMSpeed rank#43
Score1088.0
UpdatedAug 20, 2026
Confidence1
Gert LabsV3 evidence
LMSpeed rank#30
Score43.9
UpdatedAug 20, 2026
Confidence1
ResearchClawBenchV3 evidence
LMSpeed rank#18
Score12.4
UpdatedAug 20, 2026
Confidence1
Agentic score
LMSpeed rank#56
Score50.4
UpdatedAug 20, 2026
Confidence1

Coding

V3.0

undefined metric} other undefined metrics}} · Provisional

Score5580% interval39.0–71.01/4 Measured dimensionsSciCode39.4%#63 / 89Coding score57.4#32 / 87
Dimensions and evidence
Code generation55
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
scicode · scicode · z 0.49 · q 1.00
Repository engineeringPrior only
Debugging & testingPrior only
Tool-assisted development & qualityPrior only
SciCodeV3 evidence
LMSpeed rank#63
Score39.4%
UpdatedSep 13, 2026
Confidence4
Coding score
LMSpeed rank#32
Score57.4
UpdatedAug 20, 2026
Confidence1
SciCodeV3 evidence
LMSpeed rank#7
Score47.3
UpdatedAug 20, 2026
Confidence1
AA Coding Index
LMSpeed rank#55
Score42.3
UpdatedAug 20, 2026
Confidence1
AA-SciCodeV3 evidence
LMSpeed rank#44
Score47.3
UpdatedAug 20, 2026
Confidence1
EEBench
LMSpeed rank#19
Score14.3
UpdatedAug 20, 2026
Confidence1
SpaceXAI MTS Eval
LMSpeed rank#6
Score40.6
UpdatedAug 20, 2026
Confidence1

Reasoning

V3.0

undefined metric} other undefined metrics}} · Estimated

Score53.380% interval42.5–64.12/4 Measured dimensionsGPQA90.1%#33 / 218HLE37.2%#37 / 216
Dimensions and evidence
Abstract logicPrior only
Scientific & causal reasoning58
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
critpt · critpt · z 0.39 · q 1.00
gpqa · gpqa · z 0.78 · q 1.00
hle · hle · z 0.58 · q 1.00
Multi-step constraintsPrior only
Evidence integration & verification48.6
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
lcr · lcr · z -0.54 · q 1.00
GPQAV3 evidence
LMSpeed rank#33
Score90.1%
UpdatedSep 13, 2026
Confidence4
HLEV3 evidence
LMSpeed rank#37
Score37.2%
UpdatedSep 13, 2026
Confidence4
AA-LCRV3 evidence
LMSpeed rank#77
Score64.3
UpdatedAug 20, 2026
Confidence1
CritPtV3 evidence
LMSpeed rank#40
Score8.0
UpdatedAug 20, 2026
Confidence1

Knowledge

V3.0

undefined metric} other undefined metrics}} · Provisional

Score51.180% interval37.1–65.11/4 Measured dimensionsKnowledge score55.4#65 / 83Artificial Analysis Intelligence Index37.6#33 / 117
Dimensions and evidence
Broad knowledge51.1
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
aa_omniscience · aa_omniscience_index · z 0.72 · q 1.00
benchlm_category_knowledge · benchlm_category_knowledge · z -0.85 · q 1.00
Professional knowledgePrior only
FactualityPrior only
Retrieval & open-book usePrior only
Knowledge scoreV3 evidence
LMSpeed rank#65
Score55.4
UpdatedAug 20, 2026
Confidence1
Artificial Analysis Intelligence IndexV3 evidence
LMSpeed rank#33
Score37.6
UpdatedAug 20, 2026
Confidence1
AA-Omniscience Accuracy
LMSpeed rank#46
Score34.6
UpdatedAug 20, 2026
Confidence1
AA-Omniscience Hallucination Rate
LMSpeed rank#103
Score25.0
UpdatedAug 20, 2026
Confidence1
AA-GPQA Diamond
LMSpeed rank#33
Score90.1
UpdatedAug 20, 2026
Confidence1
AA-HLE
LMSpeed rank#35
Score37.2
UpdatedAug 20, 2026
Confidence1
AA-Omniscience IndexV3 evidence
LMSpeed rank#20
Score18.0
UpdatedAug 20, 2026
Confidence1

Math

V3.0

undefined metric} other undefined metrics}} · No data

No data80% interval30.8–69.20/4 Measured dimensions
Dimensions and evidence
Foundational mathPrior only
Competition mathPrior only
Advanced proofsPrior only
Applied & tool-assisted mathPrior only

Multilingual

V3.0

undefined metric} other undefined metrics}} · No data

No data80% interval30.8–69.20/4 Measured dimensions
Dimensions and evidence
Cross-language understandingPrior only
Multilingual generationPrior only
Reasoning transferPrior only
Low-resource robustnessPrior only

Multimodal

V3.0

undefined metric} other undefined metrics}} · Provisional

Score49.880% interval33.7–65.91/4 Measured dimensionsMultimodal Grounded score64.9#34 / 56MMMU-Pro78.1#18 / 31
Dimensions and evidence
Perception & OCRPrior only
Document & spatial understandingPrior only
Visual reasoning49.8
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
mmmu_pro · mmmu_pro · z -0.10 · q 1.00
Video & grounded actionPrior only
Multimodal Grounded score
LMSpeed rank#34
Score64.9
UpdatedAug 20, 2026
Confidence1
MMMU-ProV3 evidence
LMSpeed rank#18
Score78.1
UpdatedAug 20, 2026
Confidence1
Design Arena Website
LMSpeed rank#48
Score1204.0
UpdatedAug 20, 2026
Confidence1
AA-MMMU-Pro
LMSpeed rank#24
Score78.1
UpdatedAug 20, 2026
Confidence1

Instruction following

V3.0

undefined metric} other undefined metrics}} · Provisional

Score57.180% interval41.1–73.11/4 Measured dimensionsInstruction Following score94.5#1 / 52IFBench81.3#3 / 14
Dimensions and evidence
Constraint followingPrior only
Structured outputPrior only
Novel-instruction generalization57.1
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
ifbench · aa_if_bench · z 0.66 · q 1.00
Multi-turn & long instructionsPrior only
Instruction Following score
LMSpeed rank#1
Score94.5
UpdatedAug 20, 2026
Confidence1
IFBenchV3 evidence
LMSpeed rank#3
Score81.3
UpdatedAug 20, 2026
Confidence1
AA-IFBenchV3 evidence
LMSpeed rank#2
Score81.3
UpdatedAug 20, 2026
Confidence1

OpenRouter endpoints

4 endpoints

Third-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.

Provider endpointInputOutput1d uptime30m latency30m throughputContext / output
xAI
xai
$1.25/M$2.50/M99.7%undefined tokens / undefined tokens
xAI
xai/zdr/priority
$2.50/M$5/M99.6%undefined tokens / undefined tokens
xAI
xai/zdr
$1.25/M$2.50/M99.5%undefined tokens / undefined tokens
xAI
xai/priority
$2.50/M$5/M98.7%undefined tokens / undefined tokens

Pricing Comparison

Compare Grok 4.3 API pricing across 153 providers. Prices range from $0.010/request to $1071.43/M. Astrdark offers the lowest rate at $0.010/request. 5 providers offer free API credits or a free tier.

ProviderHealthModel VariantGroupInput ($/M)Output ($/M)Speed (t/s)First tokenAudit
L1
100%
grok-4.3
gf
$2.50/M
$5.00/M
L1
100%
grok-4.3
default
-93%$0.086/M
-93%$0.171/M
L1
100%
grok-4.3
default
-23%$0.959/M
-23%$1.92/M
L1
100%
grok-4.3
default
$4.28/M
Cache read$0.685/M
$8.56/M
L1
100%
grok-4.3
free
-99%$0.017/M
Cache read$0.0027/M
-99%$0.034/M
L1
100%
grok-4.3
mix
-97%$0.034/M
Cache read$0.0055/M
-97%$0.068/M
L1
100%
grok-4.3
Chat
-93%$0.086/M
Cache read$0.014/M
-93%$0.171/M
L1
100%
grok-4.3-fast
Chat
-93%$0.086/M
-93%$0.171/M
L1
100%
grok-4.3
grok
-66%$0.428/M
-66%$0.856/M
L1
100%
grok-4.3
default
-57%$0.533/M
-73%$0.666/M
L1
100%
grok-4.3
default
$1.25/M
Cache read$0.200/M
$2.50/M
L1
100%
grok-4.3
gork-官转
$1.26/M
$2.53/M
L1
100%
[lq]q|HP/grok-4.3
default
$10.00/request
-
L1
100%
[lq]q|XQ/grok-4.3
default
$15.00/request
-
L1
99%
L2
100%
grok-4.3
default
Free
Free
L1
100%
L2
100%
grok-4.3
Free
-99%$0.013/M
Cache read$0.0020/M
-99%$0.025/M
L1
100%
grok-4.3
Grok
-80%$0.250/M
Cache read$0.040/M
-80%$0.500/M
L1
100%
grok-4.3
xai
-71%$0.360/M
Cache read$0.058/M
-71%$0.719/M
L1
100%
grok-4.3
default
-55%$0.563/M
-44%$1.41/M
L1
100%
L2
100%
grok-4.3
default
-50%$0.625/M
Cache read$0.100/M
-50%$1.25/M
Showing 20 model IDs of 91.

Alternatives & Similar Models

Frequently Asked Questions

What benchmark data does Grok 4.3 include?
LMSpeed shows Grok 4.3 benchmark context, API price, output speed, first-token latency, and provider data across 158 providers when those signals are available.
What is the Grok 4.3 API price?
Grok 4.3 has pricing from undefined provider} other undefined providers}}, ranging from $0.010/request to $1071.43/M. Astrdark has the lowest listed price.
What does the Grok 4.3 API pricing table include?
The Grok 4.3 API pricing table compares 158 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
Which provider has the cheapest Grok 4.3 API pricing?
Astrdark currently has the lowest listed Grok 4.3 price at $0.010/request across undefined provider} other undefined providers}}.
Can I compare Grok 4.3 API price and speed together?
Yes. LMSpeed shows Grok 4.3 API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
Is Grok 4.3 API free?
Yes, Grok 4.3 free API options are available through 5 providerundefined other undefined} on LMSpeed, including 兔子API, DeadlySignal API, 兔子API, Moyanjdc API, 兔子API. These providers offer free API credits or a free tier with no per-token charges.
Where can I get Grok 4.3 free API access?
LMSpeed currently lists 5 free API providerundefined other undefined} for Grok 4.3: 兔子API, DeadlySignal API, 兔子API, Moyanjdc API, 兔子API. Check each provider row before using it because free tier limits can change.

Also known as

[lq]q|HP/grok-4.3[lq]q|XQ/grok-4.3[次]grok-4.3grok-4-3grok-4.3

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation