Gemini 3.8 Flash API Benchmarks, Pricing & Provider Data
Compare Gemini 3.8 Flash with another model
Choose a model to open its comparison page.
Gemini 3.8 Flash benchmark, API pricing, and provider data cover undefined API provider} other undefined API providers}}, with prices starting at $0.0021/request. Gemini 3.8 Flash free API options are available from undefined provider} other undefined providers}}.
Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.
Category Performance
Observed-capability estimates with uncertainty reported separately from independent benchmark families.
- Coverage
- 5 / 8
- Methodology
- V3.0
#1Knowledge61.3Provisional1/4 Measured dimensions
#2Agents59.9Estimated2/4 Measured dimensions
#3Reasoning59.7Estimated2/4 Measured dimensions
#4Coding58.8Estimated2/4 Measured dimensions
#5Multimodal54Provisional1/4 Measured dimensions
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Rankings
Excels at
It's decent at
Falls behind in
Detailed scores
Updated: Sep 13, 2026Overall
undefined metric} other undefined metrics}}
Overall score75.0#8 / 112DeepSWE73.8#3 / 21
Overall
undefined metric} other undefined metrics}}
Speed & latency
undefined metric} other undefined metrics}}
Output speed338.3 tok/s#4 / 80Time to first token14.92 s#69 / 80
Speed & latency
undefined metric} other undefined metrics}}
Pricing
undefined metric} other undefined metrics}}
Input price$0.750/M#107 / 186Output price$3.75/M#109 / 186
Pricing
undefined metric} other undefined metrics}}
Agents
V3.0undefined metric} other undefined metrics}} · Estimated
Score59.980% interval48.6–71.12/4 Measured dimensionsAgentic score74.4#26 / 77Finance Agent v261.4#1 / 4
Agents
V3.0undefined metric} other undefined metrics}} · Estimated
Coding
V3.0undefined metric} other undefined metrics}} · Estimated
Score58.880% interval46.9–70.72/4 Measured dimensionsSciCode56.6%#11 / 89Coding score75.2#4 / 87
Coding
V3.0undefined metric} other undefined metrics}} · Estimated
Reasoning
V3.0undefined metric} other undefined metrics}} · Estimated
Score59.780% interval49.0–70.52/4 Measured dimensionsGPQA95.3%#1 / 218HLE47.8%#10 / 216
Reasoning
V3.0undefined metric} other undefined metrics}} · Estimated
Knowledge
V3.0undefined metric} other undefined metrics}} · Provisional
Score61.380% interval47.3–75.31/4 Measured dimensionsKnowledge score83.0#9 / 83HLE-Verified54.9#1 / 6
Knowledge
V3.0undefined metric} other undefined metrics}} · Provisional
Math
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Math
V3.0undefined metric} other undefined metrics}} · No data
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
Multimodal
V3.0undefined metric} other undefined metrics}} · Provisional
Score5480% interval37.7–70.41/4 Measured dimensionsMultimodal Grounded score82.8#8 / 56CharXiv w/o tools86.2#2 / 10
Multimodal
V3.0undefined metric} other undefined metrics}} · Provisional
Instruction following
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Instruction following
V3.0undefined metric} other undefined metrics}} · No data
OpenRouter endpoints
6 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
Google google-vertex/global/priority | $1.35/M | $6.75/M | 99.9% | — | — | undefined tokens / undefined tokens |
Google AI Studio google-ai-studio/flex | $0.375/M | $1.88/M | 99.9% | — | — | undefined tokens / undefined tokens |
Google AI Studio google-ai-studio | $0.750/M | $3.75/M | 99.6% | — | — | undefined tokens / undefined tokens |
Google google-vertex/global | $0.750/M | $3.75/M | 98.8% | — | — | undefined tokens / undefined tokens |
Google AI Studio google-ai-studio/priority | $1.35/M | $6.75/M | 97.9% | — | — | undefined tokens / undefined tokens |
Google google-vertex/global/flex | $0.375/M | $1.88/M | 90.9% | — | — | undefined tokens / undefined tokens |
Pricing Comparison
Compare Gemini 3.8 Flash API pricing across 257 providers. Prices range from $0.0021/request to $1200.00/M. MyDamoxing offers the lowest rate at $0.0021/request. 4 providers offer free API credits or a free tier.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 100% | gemini-3.8-flash | gemini_cli | -60%$0.300/M | -60%$1.50/M | — | — | — | |
L1 100% | gemini-3.8-flash | gemini-slb | -57%$0.321/M Cache read$0.032/M | -57%$1.61/M | — | — | — | |
L1 100% | gemini-3.8-flash-nothinking | default | -86%$0.103/M | -86%$0.514/M | — | — | — | |
L1 100% | gemini-3.8-flash-maxthinking | default | -86%$0.103/M | -86%$0.514/M | — | — | — | |
L1 100% | gemini-3.8-flash | default | -86%$0.103/M | -86%$0.514/M | — | — | — | |
L1 100% | gemini-3.8-flash-high | default | $10.27/M | $30.82/M | — | — | — | |
L1 100% | gemini-3.8-flash | 384 | -95%$0.040/M Cache read$0.0040/MCache write$0.027/MCache write 1h$0.043/M | -95%$0.201/M | — | — | — | |
L1 100% | gemini-3.8-flash-high | 384 | -95%$0.040/M Cache read$0.0040/MCache write$0.027/MCache write 1h$0.043/M | -95%$0.201/M | — | — | — | |
L1 100% | gemini-3.8-flash-low | 384 | -95%$0.040/M Cache read$0.0040/MCache write$0.027/MCache write 1h$0.043/M | -95%$0.201/M | — | — | — | |
L1 100% | gemini-3.8-flash | default | $1.50/M Cache read$0.150/M | $9.00/M | — | — | — | |
L1 100% | gemini-3.8-flash | default | $0.015/request | - | — | — | — | |
L1 100% | gemini-3.8-flash | Anti-Gemini-1 | -98%$0.016/M Cache read$0.0016/M | -98%$0.082/M | — | — | — | |
L1 100% | gemini-3.8-flash | gemini-cli | -96%$0.031/M Cache read$0.0031/M | -96%$0.154/M | — | — | — | |
L1 100% | gemini-3.8-flash | daily-gemini | -85%$0.113/M Cache read$0.011/M | -85%$0.563/M | — | — | — | |
L1 100% | gemini-3.8-flash | gemini-anti | -70%$0.225/M Cache read$0.022/M | -70%$1.13/M | — | — | — | |
L1 100% | gemini-3.8-flash-low | default | -54%$0.342/M | -1%$3.70/M | — | — | — | |
L1 100% | gemini-3.8-flash | default | -54%$0.342/M | -1%$3.70/M | — | — | — | |
L1 100% | gemini-3.8-flash | gemini | -30%$0.525/M Cache read$0.052/M | -30%$2.63/M | — | — | — | |
L1 100% | gemini-3.8-flash-medium | Gemini 智能路由分组 | $2.19/M | -96%$0.165/M | — | — | — | |
L1 100% | gemini-3.8-flash-low | Gemini 智能路由分组 | $2.19/M | -96%$0.165/M | — | — | — |
Alternatives & Similar Models
Claude Opus 4.6
claude-opus-4-6
Anthropic Claude Opus 4.6 is the most capable Claude Opus tier, optimized for complex analysis, long-horizon coding, and high-stakes enterprise reasoning workloads.
Claude Sonnet 4.6
claude-sonnet-4-6
Anthropic Claude Sonnet 4.6 extends the Sonnet line with improved tool use, coding reliability, and long-context performance for everyday production workloads.
GPT-5.6 Sol
gpt-5-6-sol
GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...
GPT-5.6 Terra
gpt-5-6-terra
GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...
GPT-5.6 Luna
gpt-5-6-luna
GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...
GPT-5.4
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
Frequently Asked Questions
- What benchmark data does Gemini 3.8 Flash include?
- LMSpeed shows Gemini 3.8 Flash benchmark context, API price, output speed, first-token latency, and provider data across 261 providers when those signals are available.
- What is the Gemini 3.8 Flash API price?
- Gemini 3.8 Flash has pricing from undefined provider} other undefined providers}}, ranging from $0.0021/request to $1200.00/M. MyDamoxing has the lowest listed price.
- What does the Gemini 3.8 Flash API pricing table include?
- The Gemini 3.8 Flash API pricing table compares 261 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest Gemini 3.8 Flash API pricing?
- MyDamoxing currently has the lowest listed Gemini 3.8 Flash price at $0.0021/request across undefined provider} other undefined providers}}.
- Is Gemini 3.8 Flash API free?
- Yes, Gemini 3.8 Flash free API options are available through 4 providerundefined other undefined} on LMSpeed, including 兔子API, 兔子API, 兔子API, 兔子API. These providers offer free API credits or a free tier with no per-token charges.
- Where can I get Gemini 3.8 Flash free API access?
- LMSpeed currently lists 4 free API providerundefined other undefined} for Gemini 3.8 Flash: 兔子API, 兔子API, 兔子API, 兔子API. Check each provider row before using it because free tier limits can change.
