Google
·Released on Feb 19, 2026

Gemini 3.1 Pro API Benchmarks, Pricing & Provider Data

Compare Gemini 3.1 Pro with another model

Choose a model to open its comparison page.

Share on X
LLM

Gemini 3.1 Pro benchmark, API pricing, and provider data cover 659 API providers, with prices starting at $0.0041/request. Gemini 3.1 Pro free API options are available from 8 providers. The page also shows measured API speed and first-token latency.

Google Gemini 3.1 Pro is a Gemini 3 series model with advanced multimodal reasoning, long-context support, and strong performance on coding and analytical tasks.

Quality
#22of 100
67.0
LMSpeed score
Speed
149char/s
16.52 s
Cost
$0.0041/ 1M · 8:1 in:out
$0.0007 in · $0.0035 out

Category Performance

Observed-capability estimates with uncertainty reported separately from independent benchmark families.

Coverage
8 / 8
Methodology
V3.0
Category PerformanceObserved-capability estimates with uncertainty reported separately from independent benchmark families.Agents50.6Coding58.8Reasoning58.3Knowledge57.8Math55.3Multilingual59.7Multimodal55.4Instruction following55.1
#1Multilingual59.7Provisional1/4 Measured dimensions
80% interval: 42.577.0
cross language understanding59.7
global_mmlu · aa_global_mmlu_lite
multilingual generationPrior only
reasoning transferPrior only
low resource robustnessPrior only
#2Coding58.8Estimated2/4 Measured dimensions
80% interval: 48.169.4
code generation65.6
livecodebench · live_code_bench_pro / scicode · aa_sci_code
repository engineering51.9
react_native_bench · react_native_evals / vibecode · vibe_code_bench
debugging testingPrior only
tooling qualityPrior only
#3Reasoning58.3RatedGlobal rank #93/4 Measured dimensions
80% interval: 49.567.0
abstract logic53.1
arc_agi · arc_agi2
scientific causal64.1
critpt · critpt / gpqa · aa_gpqa_diamond / hle · aa_hle / hle · hle_no_tools
multistep constraintsPrior only
evidence verification57.5
lcr · lcr
#4Knowledge57.8Provisional1/4 Measured dimensions
80% interval: 42.773.0
broad knowledgePrior only
professional knowledge57.8
healthbench · health_bench_hard / medxpert · med_xpert_qa_text
factualityPrior only
retrieval open bookPrior only
#5Multimodal55.4RatedGlobal rank #24/4 Measured dimensions
80% interval: 48.862.1
perception ocr53
simplevqa · simple_vqa
document spatial48
charxiv · charxiv
visual reasoning64.1
erqa · erqa / medxpert · med_xpert_qa_mm / mmmu_pro · mmmu_pro
video action56.7
screenspot · screen_spot_pro
#6Math55.3Provisional1/4 Measured dimensions
80% interval: 39.271.4
foundational mathPrior only
competition mathPrior only
proof frontier55.3
frontiermath · frontier_math_v2_tier4 / frontiermath · frontier_math_v2_tiers13
applied tool mathPrior only
#7Instruction following55.1Provisional1/4 Measured dimensions
80% interval: 39.171.1
constraint followingPrior only
structured outputPrior only
novel instruction generalization55.1
ifbench · aa_if_bench
long multiturn instructionPrior only
#8Agents50.6RatedGlobal rank #344/4 Measured dimensions
80% interval: 44.456.7
planning54
deep_planning · gert_labs
tool use56
tau · tau2_bench
environment execution45.9
deep_search_qa · deep_search_qa / gdpval_aa · benchlm_agentic_gdpval_aa
recovery reliability46.4
apex_agents · apex_agents_aa / claw_eval · claw_eval / researchclaw · research_claw_bench

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
1.0Mtokens
1.3K pages of text
OUTPUT
65.5Ktokens
8K128K1M4M
1.0M

Features

Technical Details

Input
Output
Released
Feb 2026
Documentation
Tokenizer
Gemini
Architecture
text+image+file+audio+video->text
Moderated
No
Supported parameters
include_reasoningmax_tokensreasoningresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_p

Rankings

Excels at

It's decent at

Falls behind in

Detailed scores

Updated: Sep 1, 2026

Overall

undefined metric} other undefined metrics}}

Overall score67.0#22 / 100
Overall score
LMSpeed rank#22
Score67.0
UpdatedSep 1, 2026
Confidence3

Agents

V3.0

undefined metric} other undefined metrics}} · Rated

Score50.6#3480% interval44.4–56.74/4 Measured dimensionsAgentic score67.5#26 / 69Claw-Eval57.8#16 / 27
Dimensions and evidence
Planning & decomposition54
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
deep_planning · gert_labs · z 0.55 · q 1.00
Tool use56
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
tau · tau2_bench · z 0.69 · q 1.00
Environment & long-horizon execution45.9
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
deep_search_qa · deep_search_qa · z -0.58 · q 1.00
gdpval_aa · benchlm_agentic_gdpval_aa · z -0.59 · q 1.00
Recovery & completion reliability46.4
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
apex_agents · apex_agents_aa · z 0.44 · q 1.00
claw_eval · claw_eval · z -0.29 · q 1.00
researchclaw · research_claw_bench · z -1.64 · q 1.00
Agentic score
LMSpeed rank#26
Score67.5
UpdatedSep 1, 2026
Confidence3
Claw-EvalV3 evidence
LMSpeed rank#16
Score57.8
UpdatedSep 1, 2026
Confidence3
DeepSearchQAV3 evidence
LMSpeed rank#12
Score69.7
UpdatedSep 1, 2026
Confidence3
Τ²-bench resultsV3 evidence
LMSpeed rank#17
Score95.6
UpdatedSep 1, 2026
Confidence3
AA Agentic Index
LMSpeed rank#43
Score23.1
UpdatedSep 1, 2026
Confidence3
APEX-Agents-AAV3 evidence
LMSpeed rank#9
Score32.0
UpdatedSep 1, 2026
Confidence3
GDPval-AA
LMSpeed rank#49
Score23.2
UpdatedSep 1, 2026
Confidence3
GDPval-AAV3 evidence
LMSpeed rank#49
Score965.0
UpdatedSep 1, 2026
Confidence3
Gert LabsV3 evidence
LMSpeed rank#14
Score56.9
UpdatedSep 1, 2026
Confidence3
ResearchClawBenchV3 evidence
LMSpeed rank#17
Score13.3
UpdatedSep 1, 2026
Confidence3
AA AutomationBench
LMSpeed rank#14
Score37.5
UpdatedAug 20, 2026
Confidence3

Coding

V3.0

undefined metric} other undefined metrics}} · Estimated

Score58.880% interval48.1–69.42/4 Measured dimensionsCoding score59.4#23 / 81LiveCodeBench Pro82.9#2 / 4
Dimensions and evidence
Code generation65.6
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
livecodebench · live_code_bench_pro · z 0.54 · q 0.50
scicode · aa_sci_code · z 1.89 · q 1.00
Repository engineering51.9
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
react_native_bench · react_native_evals · z 0.10 · q 1.00
vibecode · vibe_code_bench · z 0.27 · q 1.00
Debugging & testingPrior only
Tool-assisted development & qualityPrior only
Coding score
LMSpeed rank#23
Score59.4
UpdatedSep 1, 2026
Confidence3
LiveCodeBench ProV3 evidence
LMSpeed rank#2
Score82.9
UpdatedSep 1, 2026
Confidence3
React Native EvalsV3 evidence
LMSpeed rank#6
Score78.9
UpdatedSep 1, 2026
Confidence3
Vibe Code BenchV3 evidence
LMSpeed rank#13
Score32.0
UpdatedSep 1, 2026
Confidence3
AA Coding Index
LMSpeed rank#21
Score68.8
UpdatedSep 1, 2026
Confidence3
AA-SciCodeV3 evidence
LMSpeed rank#2
Score58.9
UpdatedSep 1, 2026
Confidence3
EEBench
LMSpeed rank#5
Score46.3
UpdatedAug 20, 2026
Confidence3
CADGenBench Generation
LMSpeed rank#8
Score21.1
UpdatedAug 20, 2026
Confidence3

Reasoning

V3.0

undefined metric} other undefined metrics}} · Rated

Score58.3#980% interval49.5–67.03/4 Measured dimensionsReasoning score72.4#12 / 29ARC-AGI-277.1#6 / 17
Dimensions and evidence
Abstract logic53.1
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
arc_agi · arc_agi2 · z 0.24 · q 1.00
Scientific & causal reasoning64.1
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
critpt · critpt · z 0.83 · q 1.00
gpqa · aa_gpqa_diamond · z 1.50 · q 1.00
hle · aa_hle · z 1.13 · q 1.00
hle · hle_no_tools · z 0.43 · q 1.00
Multi-step constraintsPrior only
Evidence integration & verification57.5
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
lcr · lcr · z 0.65 · q 1.00
Reasoning score
LMSpeed rank#12
Score72.4
UpdatedSep 1, 2026
Confidence3
ARC-AGI-2V3 evidence
LMSpeed rank#6
Score77.1
UpdatedSep 1, 2026
Confidence3
AA-LCRV3 evidence
LMSpeed rank#10
Score79.0
UpdatedSep 1, 2026
Confidence3
CritPtV3 evidence
LMSpeed rank#15
Score17.7
UpdatedSep 1, 2026
Confidence3
ARC-AGI-3
LMSpeed rank#6
Score0.4
UpdatedSep 1, 2026
Confidence3

Knowledge

V3.0

undefined metric} other undefined metrics}} · Provisional

Score57.880% interval42.7–73.01/4 Measured dimensionsKnowledge score66.7#37 / 69GPQA-D94.3#2 / 31
Dimensions and evidence
Broad knowledgePrior only
Professional knowledge57.8
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
healthbench · health_bench_hard · z -1.24 · q 0.88
medxpert · med_xpert_qa_text · z 2.02 · q 0.50
FactualityPrior only
Retrieval & open-book usePrior only
Knowledge score
LMSpeed rank#37
Score66.7
UpdatedSep 1, 2026
Confidence3
GPQA-D
LMSpeed rank#2
Score94.3
UpdatedSep 1, 2026
Confidence3
HLE w/o tools
LMSpeed rank#5
Score45.4
UpdatedSep 1, 2026
Confidence3
HealthBench HardV3 evidence
LMSpeed rank#5
Score20.6
UpdatedSep 1, 2026
Confidence3
MedXpertQA (Text)V3 evidence
LMSpeed rank#1
Score71.5
UpdatedSep 1, 2026
Confidence3
Artificial Analysis Intelligence Index
LMSpeed rank#23
Score47.7
UpdatedSep 1, 2026
Confidence3
AA-GPQA Diamond
LMSpeed rank#3
Score94.1
UpdatedSep 1, 2026
Confidence3
AA-HLE
LMSpeed rank#6
Score47.0
UpdatedSep 1, 2026
Confidence3
AA-Omniscience Index
LMSpeed rank#3
Score31.9
UpdatedSep 1, 2026
Confidence3
AA-Omniscience Accuracy
LMSpeed rank#7
Score54.9
UpdatedSep 1, 2026
Confidence3
AA-Omniscience Hallucination Rate
LMSpeed rank#73
Score50.9
UpdatedSep 1, 2026
Confidence3

Math

V3.0

undefined metric} other undefined metrics}} · Provisional

Score55.380% interval39.2–71.41/4 Measured dimensionsMath score55.1#34 / 62FrontierMath v2 (Tiers 1-3)36.9#15 / 47
Dimensions and evidence
Foundational mathPrior only
Competition mathPrior only
Advanced proofs55.3
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
frontiermath · frontier_math_v2_tier4 · z 0.62 · q 1.00
frontiermath · frontier_math_v2_tiers13 · z 0.36 · q 1.00
Applied & tool-assisted mathPrior only
Math score
LMSpeed rank#34
Score55.1
UpdatedSep 1, 2026
Confidence3
FrontierMath v2 (Tiers 1-3)V3 evidence
LMSpeed rank#15
Score36.9
UpdatedSep 1, 2026
Confidence3
FrontierMath v2 (Tier 4)V3 evidence
LMSpeed rank#13
Score16.7
UpdatedSep 1, 2026
Confidence3

Multilingual

V3.0

undefined metric} other undefined metrics}} · Provisional

Score59.780% interval42.5–77.01/4 Measured dimensionsAA Global-MMLU-Lite93.2#1 / 4
Dimensions and evidence
Cross-language understanding59.7
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
global_mmlu · aa_global_mmlu_lite · z 0.76 · q 0.50
Multilingual generationPrior only
Reasoning transferPrior only
Low-resource robustnessPrior only
AA Global-MMLU-LiteV3 evidence
LMSpeed rank#1
Score93.2
UpdatedSep 1, 2026
Confidence3

Multimodal

V3.0

undefined metric} other undefined metrics}} · Rated

Score55.4#280% interval48.8–62.14/4 Measured dimensionsMultimodal Grounded score76.8#9 / 44MMMU-Pro83.9#2 / 31
Dimensions and evidence
Perception & OCR53
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
simplevqa · simple_vqa · z 0.43 · q 1.00
Document & spatial understanding48
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
charxiv · charxiv · z -0.14 · q 1.00
Visual reasoning64.1
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
erqa · erqa · z 0.61 · q 1.00
medxpert · med_xpert_qa_mm · z 0.86 · q 0.75
mmmu_pro · mmmu_pro · z 1.46 · q 1.00
Video & grounded action56.7
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
screenspot · screen_spot_pro · z 0.60 · q 1.00
Multimodal Grounded score
LMSpeed rank#9
Score76.8
UpdatedSep 1, 2026
Confidence3
MMMU-ProV3 evidence
LMSpeed rank#2
Score83.9
UpdatedSep 1, 2026
Confidence3
CharXivV3 evidence
LMSpeed rank#20
Score80.2
UpdatedSep 1, 2026
Confidence3
ERQAV3 evidence
LMSpeed rank#3
Score69.4
UpdatedSep 1, 2026
Confidence3
SimpleVQAV3 evidence
LMSpeed rank#4
Score72.4
UpdatedSep 1, 2026
Confidence3
ScreenSpot ProV3 evidence
LMSpeed rank#4
Score84.4
UpdatedSep 1, 2026
Confidence3
ZeroBench
LMSpeed rank#2
Score29.0
UpdatedSep 1, 2026
Confidence3
MedXpertQA (MM)V3 evidence
LMSpeed rank#1
Score81.3
UpdatedSep 1, 2026
Confidence3
AA-MMMU-Pro
LMSpeed rank#6
Score82.4
UpdatedSep 1, 2026
Confidence3
Design Arena Website
LMSpeed rank#28
Score1272.0
UpdatedSep 1, 2026
Confidence3

Instruction following

V3.0

undefined metric} other undefined metrics}} · Provisional

Score55.180% interval39.1–71.11/4 Measured dimensionsAA-IFBench77.1#8 / 84
Dimensions and evidence
Constraint followingPrior only
Structured outputPrior only
Novel-instruction generalization55.1
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
ifbench · aa_if_bench · z 0.41 · q 1.00
Multi-turn & long instructionsPrior only
AA-IFBenchV3 evidence
LMSpeed rank#8
Score77.1
UpdatedSep 1, 2026
Confidence3

Pricing Comparison

Compare Gemini 3.1 Pro API pricing across 651 providers. Prices range from $0.0041/request to $1600.00/M. 钠 API offers the lowest rate at $0.0041/request. 8 providers offer free API credits or a free tier.

ProviderHealthModel VariantGroupInput ($/M)Output ($/M)Speed (t/s)First tokenAudit
L1
99%
gemini-3.1-pro-preview
default
$2.00/M
$12.00/M
105.4 t/s
12.72 s
L1
99%
gemini-3.1-pro
Gemini
$0.600/M
Cache read$0.060/M
$3.60/M
97.9 t/s
12.32 s
L1
99%
gemini-3.1-pro-preview
default
$0.0041/request
-
95.5 t/s
13.02 s
L1
99%
gemini-3.1-pro-high
default
$0.0041/request
-
L1
99%
gemini-3.1-pro-low
default
$0.0041/request
-
L1
99%
gemini-3.1-pro
default
$20.00/M
Cache read$2.00/MCache write$10.00/MCache write 1h$16.00/M
$100.00/M
88.8 t/s
11.90 s
L1
99%
gemini-3.1-pro-preview
default
$2.80/M
$16.80/M
70.1 t/s
23.27 s
L1
99%
gemini-3.1-pro
default
$2.80/M
$16.80/M
L1
99%
gemini-3.1-pro-preview-customtools
default
$2.80/M
$16.80/M
L1
1%
gemini-3.1-pro-preview
Gemini特价
$0.036/request
-
68.9 t/s
19.70 s
L1
100%
gemini-3.1-pro-preview
特供-gemini45折
$0.137/M
Cache read$0.014/M
$0.822/M
L1
100%
gemini-3.1-pro-preview
gemini-officially
$2.00/M
$12.00/M
L1
100%
gemini-3.1-pro-preview-low
gemini_cli
$1.00/M
$6.00/M
L1
100%
gemini-3.1-pro-preview-high
gemini_cli
$1.00/M
$6.00/M
L1
100%
gemini-3.1-pro-preview
gemini_cli
$1.00/M
$6.00/M
L1
100%
gemini-3.1-pro-preview
default
$0.014/request
-
L1
100%
gemini-3.1-pro-preview-maxthinking
default
$0.274/M
$1.64/M
L1
100%
gemini-3.1-pro-high
default
$0.274/M
$1.64/M
L1
100%
gemini-3.1-pro
default
$0.274/M
$1.64/M
L1
100%
gemini-3.1-pro-preview-nothinking
default
$0.274/M
$1.64/M
Showing 20 model IDs of 237.

Alternatives & Similar Models

Frequently Asked Questions

What benchmark data does Gemini 3.1 Pro include?
LMSpeed shows Gemini 3.1 Pro benchmark context, API price, output speed, first-token latency, and provider data across 659 providers when those signals are available.
What is the Gemini 3.1 Pro API price?
Gemini 3.1 Pro has pricing from 659 providers, ranging from $0.0041/request to $1600.00/M. 钠 API has the lowest listed price.
What does the Gemini 3.1 Pro API pricing table include?
The Gemini 3.1 Pro API pricing table compares 659 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
Which provider has the cheapest Gemini 3.1 Pro API pricing?
钠 API currently has the lowest listed Gemini 3.1 Pro price at $0.0041/request across 659 providers.
Can I compare Gemini 3.1 Pro API price and speed together?
Yes. LMSpeed shows Gemini 3.1 Pro API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
Is Gemini 3.1 Pro API free?
Yes, Gemini 3.1 Pro free API options are available through 8 providerundefined other undefined} on LMSpeed, including 兔子API, 兔子API, 兔子API, 兔子API, 梦德 API. These providers offer free API credits or a free tier with no per-token charges.
Where can I get Gemini 3.1 Pro free API access?
LMSpeed currently lists 8 free API providerundefined other undefined} for Gemini 3.1 Pro: 兔子API, 兔子API, 兔子API, 兔子API, 梦德 API. Check each provider row before using it because free tier limits can change.

Also known as

Gemini-3-1-ProGemini-3.1-ProNotion/gemini-3.1-pro-preview[C][0.35/次]流式抗截断/gemini-3.1-pro-preview[L]gemini-3.1-pro-preview

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation