Gemini 2.5 Flash API Benchmarks, Pricing & Provider Data
Compare Gemini 2.5 Flash with another model
Choose a model to open its comparison page.
Gemini 2.5 Flash benchmark, API pricing, and provider data cover 573 API providers, with prices starting at $0.0010/request. Gemini 2.5 Flash free API options are available from 15 providers. The page also shows measured API speed and first-token latency.
Google Gemini 2.5 Flash is a fast multimodal model balancing speed and intelligence for chat, tool use, and large-context workloads at lower cost than Pro tiers.
Category Performance
Observed-capability estimates with uncertainty reported separately from independent benchmark families.
- Coverage
- 6 / 8
- Methodology
- V3.0
#1Math50.8Estimated3/4 Measured dimensions
#2Coding50.6Provisional1/4 Measured dimensions
#3Reasoning49RatedGlobal rank #433/4 Measured dimensions
#4Instruction following43.1Provisional1/4 Measured dimensions
#5Knowledge40.8Provisional1/4 Measured dimensions
#6Agents36.3Provisional1/4 Measured dimensions
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Rankings
Excels at
It's decent at
Falls behind in
Detailed scores
Updated: Sep 1, 2026Overall
undefined metric} other undefined metrics}}
Overall score45.0#82 / 100
Overall
undefined metric} other undefined metrics}}
Pricing
undefined metric} other undefined metrics}}
Input price$0.300/M#55 / 181Output price$2.50/M#87 / 181
Pricing
undefined metric} other undefined metrics}}
Agents
V3.0undefined metric} other undefined metrics}} · Provisional
Score36.380% interval20.3–52.41/4 Measured dimensionsΤ²-bench results14.9#82 / 82
Agents
V3.0undefined metric} other undefined metrics}} · Provisional
Coding
V3.0undefined metric} other undefined metrics}} · Provisional
Score50.680% interval36.7–64.51/4 Measured dimensionsLiveCodeBench69.5%#30 / 115SciCode39.4%#96 / 206
Coding
V3.0undefined metric} other undefined metrics}} · Provisional
Reasoning
V3.0undefined metric} other undefined metrics}} · Rated
Score49#4380% interval40.3–57.63/4 Measured dimensionsMMLU-Pro83.2%#42 / 129GPQA79.0%#95 / 213
Reasoning
V3.0undefined metric} other undefined metrics}} · Rated
Knowledge
V3.0undefined metric} other undefined metrics}} · Provisional
Score40.880% interval24.8–56.91/4 Measured dimensionsArtificial Analysis Intelligence Index14.2#97 / 111AA-GPQA Diamond68.3#87 / 108
Knowledge
V3.0undefined metric} other undefined metrics}} · Provisional
Math
V3.0undefined metric} other undefined metrics}} · Estimated
Score50.880% interval41.5–60.03/4 Measured dimensionsAIME82.3%#11 / 68MATH-50098.1%#7 / 73
Math
V3.0undefined metric} other undefined metrics}} · Estimated
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
Multimodal
V3.0undefined metric} other undefined metrics}} · No data
80% interval30.8–69.20/4 Measured dimensionsAA-MMMU-Pro65.5#50 / 65Design Arena Website1127.0#64 / 75
Multimodal
V3.0undefined metric} other undefined metrics}} · No data
Instruction following
V3.0undefined metric} other undefined metrics}} · Provisional
Score43.180% interval27.1–59.11/4 Measured dimensionsAA-IFBench39.0#72 / 84
Instruction following
V3.0undefined metric} other undefined metrics}} · Provisional
OpenRouter endpoints
7 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
Google AI Studio google-ai-studio/flex | $0.150/M | $1.25/M | 100% | — | — | undefined tokens / undefined tokens |
Google AI Studio google-ai-studio | $0.300/M | $2.50/M | 99.9% | — | — | undefined tokens / undefined tokens |
Google AI Studio google-ai-studio/priority | $0.540/M | $4.50/M | 99.6% | — | — | undefined tokens / undefined tokens |
Google google-vertex/global | $0.300/M | $2.50/M | 99.6% | — | — | undefined tokens / undefined tokens |
Google google-vertex/eu | $0.300/M | $2.50/M | 98.7% | — | — | undefined tokens / undefined tokens |
Google google-vertex/global/priority | $0.540/M | $4.50/M | 98.6% | — | — | undefined tokens / undefined tokens |
Google google-vertex | $0.300/M | $2.50/M | 79.3% | — | — | undefined tokens / undefined tokens |
Pricing Comparison
Compare Gemini 2.5 Flash API pricing across 558 providers. Prices range from $0.0010/request to $210.00/M. 酒馆无限制免费API offers the lowest rate at $0.0010/request. 15 providers offer free API credits or a free tier.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 79% L2 0% | gemini-2.5-flash | 默认分组 | $2.74/M Cache read$0.547/M | $10.95/M | 208.1 t/s | 5.69 s | — | |
L1 100% | gemini-2.5-flash | gemini-cli | -86%$0.041/M Cache read$0.0041/M | -86%$0.343/M | 200.1 t/s | 8.21 s | — | |
L1 0% | gemini-2.5-flash | default | $0.300/M Cache read$0.030/MImage$0.300/MAudio in$0.999/M | $2.50/M | 158.0 t/s | 9.77 s | — | |
L1 100% | gemini-2.5-flash | default | $0.010/request | - | 136.0 t/s | 11.32 s | — | |
L1 100% | anti/gemini-2.5-flash | default | $0.010/request | - | — | — | — | |
L1 100% | cli/gemini-2.5-flash | default | $0.010/request | - | — | — | — | |
L1 100% | gemini-2.5-flash | 特供-gemini45折 | -93%$0.021/M Cache read$0.0021/M | -93%$0.171/M | — | — | — | |
L1 100% | gemini-2.5-flash | gemini-officially | $0.300/M | $2.50/M | — | — | — | |
L1 100% | gemini-2.5-flash-nothinking | gemini_cli | -50%$0.150/M | -50%$1.25/M | — | — | — | |
L1 100% | gemini-2.5-flash | gemini_cli | -50%$0.150/M | -50%$1.25/M | — | — | — | |
L1 100% | gemini-2.5-flash-maxthinking | default | -86%$0.041/M | -86%$0.343/M | — | — | — | |
L1 100% | gemini-2.5-flash-nothinking | default | -86%$0.041/M | -86%$0.343/M | — | — | — | |
L1 100% | gemini-2.5-flash | default | -86%$0.041/M | -86%$0.343/M | — | — | — | |
L1 100% | gemini-2.5-flash-all | default | -86%$0.041/M | -86%$0.343/M | — | — | — | |
L1 100% | gemini-2.5-flash-preview-09-2025 | Gemini_vip | -86%$0.041/M | -86%$0.343/M | — | — | — | |
L1 100% | gemini-2.5-flash-preview-09-2025-maxthinking | Gemini_vip | -86%$0.041/M | -95%$0.123/M | — | — | — | |
L1 100% | gemini-2.5-flash-preview-09-2025-nothinking | Gemini_vip | -86%$0.041/M | -86%$0.343/M | — | — | — | |
L1 100% | gemini-2.5-flash-nothinking | default | -23%$0.230/M | -23%$1.92/M | — | — | — | |
L1 100% | gemini-2.5-flash | default | -23%$0.230/M | -23%$1.92/M | — | — | — | |
L1 100% | gemini-2.5-flash | gemini官 | $0.420/M Cache read$0.042/M | $3.50/M | — | — | — |
Alternatives & Similar Models
Gemini 2.5 Pro
gemini-2-5-pro
Google Gemini 2.5 Pro is Google advanced multimodal model with a 1M-token context window, strong STEM reasoning, and native support for images, audio, and video understanding.
GPT-5.4
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
Gemini 3 Flash
gemini-3-flash
Google Gemini 3 Flash is a next-generation fast multimodal model for responsive assistants, document understanding, and high-throughput API traffic.
Gemini 3.1 Pro
gemini-3-1-pro
Google Gemini 3.1 Pro is a Gemini 3 series model with advanced multimodal reasoning, long-context support, and strong performance on coding and analytical tasks.
Claude Sonnet 4.6
claude-sonnet-4-6
Anthropic Claude Sonnet 4.6 extends the Sonnet line with improved tool use, coding reliability, and long-context performance for everyday production workloads.
GPT-5.3 Codex
gpt-5-3-codex
OpenAI GPT-5.3 Codex is a code-specialized variant in the GPT-5 series, optimized for code generation, debugging, and software development tasks.
Frequently Asked Questions
- What benchmark data does Gemini 2.5 Flash include?
- LMSpeed shows Gemini 2.5 Flash benchmark context, API price, output speed, first-token latency, and provider data across 573 providers when those signals are available.
- What is the Gemini 2.5 Flash API price?
- Gemini 2.5 Flash has pricing from 573 providers, ranging from $0.0010/request to $210.00/M. 酒馆无限制免费API has the lowest listed price.
- What does the Gemini 2.5 Flash API pricing table include?
- The Gemini 2.5 Flash API pricing table compares 573 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest Gemini 2.5 Flash API pricing?
- 酒馆无限制免费API currently has the lowest listed Gemini 2.5 Flash price at $0.0010/request across 573 providers.
- Can I compare Gemini 2.5 Flash API price and speed together?
- Yes. LMSpeed shows Gemini 2.5 Flash API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
- Is Gemini 2.5 Flash API free?
- Yes, Gemini 2.5 Flash free API options are available through 15 providerundefined other undefined} on LMSpeed, including 兔子API, 兔子API, 兔子API, 兔子API, 兔子API. These providers offer free API credits or a free tier with no per-token charges.
- Where can I get Gemini 2.5 Flash free API access?
- LMSpeed currently lists 15 free API providerundefined other undefined} for Gemini 2.5 Flash: 兔子API, 兔子API, 兔子API, 兔子API, 兔子API. Check each provider row before using it because free tier limits can change.
