Gemini 3.5 Flash-Lite API Benchmarks, Pricing & Provider Data
Compare Gemini 3.5 Flash-Lite with another model
Choose a model to open its comparison page.
Gemini 3.5 Flash-Lite benchmark, API pricing, and provider data cover undefined API provider} other undefined API providers}}, with prices starting at $0.0065/M. Gemini 3.5 Flash-Lite free API options are available from undefined provider} other undefined providers}}.
Gemini 3.5 Flash-Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.
Category Performance
Observed-capability estimates with uncertainty reported separately from independent benchmark families.
- Coverage
- 5 / 8
- Methodology
- V3.0
#1Reasoning52Estimated2/4 Measured dimensions
#2Knowledge51Provisional1/4 Measured dimensions
#3Multimodal44.7Provisional1/4 Measured dimensions
#4Coding44.5Estimated3/4 Measured dimensions
#5Agents42.1RatedGlobal rank #573/4 Measured dimensions
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Rankings
Excels at
Falls behind in
Detailed scores
Updated: Sep 11, 2026Overall
undefined metric} other undefined metrics}}
Overall score58.0#60 / 111
Overall
undefined metric} other undefined metrics}}
Speed & latency
undefined metric} other undefined metrics}}
Output speed349.9 tok/s#3 / 81Time to first token7.87 s#66 / 81
Speed & latency
undefined metric} other undefined metrics}}
Pricing
undefined metric} other undefined metrics}}
Input price$0.300/M#55 / 186Output price$2.50/M#89 / 186
Pricing
undefined metric} other undefined metrics}}
Agents
V3.0undefined metric} other undefined metrics}} · Rated
Score42.1#5780% interval33.4–50.93/4 Measured dimensionsAgentic score50.9#53 / 76Terminal-Bench 2.054.0#41 / 53
Agents
V3.0undefined metric} other undefined metrics}} · Rated
Coding
V3.0undefined metric} other undefined metrics}} · Estimated
Score44.580% interval35.2–53.73/4 Measured dimensionsSciCode41.3%#58 / 89Coding score38.5#77 / 87
Coding
V3.0undefined metric} other undefined metrics}} · Estimated
Reasoning
V3.0undefined metric} other undefined metrics}} · Estimated
Score5280% interval41.2–62.82/4 Measured dimensionsGPQA83.8%#77 / 218HLE18.8%#90 / 216
Reasoning
V3.0undefined metric} other undefined metrics}} · Estimated
Knowledge
V3.0undefined metric} other undefined metrics}} · Provisional
Score5180% interval37.0–65.01/4 Measured dimensionsArtificial Analysis Intelligence Index37.4#33 / 117AA-GPQA Diamond83.8#68 / 113
Knowledge
V3.0undefined metric} other undefined metrics}} · Provisional
Math
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Math
V3.0undefined metric} other undefined metrics}} · No data
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
Multimodal
V3.0undefined metric} other undefined metrics}} · Provisional
Score44.780% interval28.3–61.11/4 Measured dimensionsAA-MMMU-Pro79.0#17 / 68Multimodal Grounded score62.3#39 / 56
Multimodal
V3.0undefined metric} other undefined metrics}} · Provisional
Instruction following
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Instruction following
V3.0undefined metric} other undefined metrics}} · No data
OpenRouter endpoints
8 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
Google google-vertex/eu | $0.330/M | $2.75/M | 100% | — | — | undefined tokens / undefined tokens |
Google google-vertex/us | $0.330/M | $2.75/M | 100% | — | — | undefined tokens / undefined tokens |
Google AI Studio google-ai-studio/flex | $0.150/M | $1.25/M | 100.0% | — | — | undefined tokens / undefined tokens |
Google AI Studio google-ai-studio/priority | $0.540/M | $4.50/M | 99.9% | — | — | undefined tokens / undefined tokens |
Google AI Studio google-ai-studio | $0.300/M | $2.50/M | 99.8% | — | — | undefined tokens / undefined tokens |
Google google-vertex/global | $0.300/M | $2.50/M | 99.8% | — | — | undefined tokens / undefined tokens |
Google google-vertex/global/priority | $0.540/M | $4.50/M | 99.3% | — | — | undefined tokens / undefined tokens |
Google google-vertex/global/flex | $0.150/M | $1.25/M | 99.0% | — | — | undefined tokens / undefined tokens |
Pricing Comparison
Compare Gemini 3.5 Flash-Lite API pricing across 163 providers. Prices range from $0.0065/M to $2.19/M. N1N offers the lowest rate at $0.0065/M. 4 providers offer free API credits or a free tier.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 100% | gemini-3.5-flash-lite | 优质gemini | -84%$0.049/M Cache read$0.0049/M | -84%$0.411/M | — | — | — | |
L1 100% | gemini-3.5-flash-lite-maxthinking | default | -86%$0.041/M Cache read$0.0041/MImage$0.041/M | -86%$0.343/M | — | — | — | |
L1 100% | gemini-3.5-flash-lite-nothinking | default | -86%$0.041/M Cache read$0.0041/MImage$0.041/M | -86%$0.343/M | — | — | — | |
L1 100% | gemini-3.5-flash-lite | default | -86%$0.041/M Cache read$0.0041/MImage$0.041/M | -86%$0.343/M | — | — | — | |
L1 100% | gemini-3.5-flash-lite | default | -23%$0.230/M | -23%$1.92/M | — | — | — | |
L1 100% | gemini-3.5-flash-lite | Cli-Gemini-1 | -98%$0.0065/M Cache read$0.0007/M | -98%$0.055/M | — | — | — | |
L1 100% | gemini-3.5-flash-lite | default | $0.0073/request | - | — | — | — | |
L1 100% | gemini-3.5-flash-lite | 优质gemini | -84%$0.049/M Cache read$0.0049/M | -84%$0.411/M | — | — | — | |
L1 100% | gemini-3.5-flash-lite | 优质gemini | -79%$0.062/M | -79%$0.514/M | — | — | — | |
L1 100% | gemini-3.5-flash-lite | gemini-vertex-企业版 | -59%$0.123/M | -59%$1.02/M | — | — | — | |
L1 100% | gemini-3.5-flash-lite | Gemini | -93%$0.021/M Cache read$0.0021/M | -93%$0.171/M | — | — | — | |
L1 100% | gemini-3.5-flash-lite | gemini优质 | -45%$0.164/M | -45%$1.37/M | — | — | — | |
L1 100% | gemini-3.5-flash-lite | default | -32%$0.205/M | -18%$2.05/M | — | — | — | |
L1 100% | gemini-3.5-flash-lite | Gemini | -66%$0.103/M | -66%$0.856/M | — | — | — | |
L1 100% | gemini-3.5-flash-lite | default | $0.300/M Cache read$0.030/M | $2.50/M | — | — | — | |
兔子API Free | L1 99% | gemini-3.5-flash-lite | default | Free | Free | — | — | — |
L1 100% L2 0% | gemini-3.5-flash-lite | default | -50%$0.150/M Cache read$0.015/MCache write$0.500/MCache write 1h$0.800/M | -50%$1.25/M | — | — | — | |
L1 100% | gemini-3.5-flash-lite | 优质gemini | -74%$0.079/M Cache read$0.0079/M | -74%$0.658/M | — | — | — | |
L1 100% | gemini-3.5-flash-lite | Gemini-优质临时分组 | -59%$0.123/M Cache read$0.012/M | -59%$1.03/M | — | — | — | |
L1 100% | gemini-3.5-flash-lite | default | $0.375/M Cache read$0.037/MImage$0.375/M | -10%$2.25/M | — | — | — |
Alternatives & Similar Models
GPT-5.4
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
Gemini 3.6 Flash
gemini-3-6-flash
Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...
GPT-5.6 Terra
gpt-5-6-terra
GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...
GPT-5.6 Luna
gpt-5-6-luna
GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...
Claude Opus 4.6
claude-opus-4-6
Anthropic Claude Opus 4.6 is the most capable Claude Opus tier, optimized for complex analysis, long-horizon coding, and high-stakes enterprise reasoning workloads.
Claude Sonnet 4.6
claude-sonnet-4-6
Anthropic Claude Sonnet 4.6 extends the Sonnet line with improved tool use, coding reliability, and long-context performance for everyday production workloads.
Frequently Asked Questions
- What benchmark data does Gemini 3.5 Flash-Lite include?
- LMSpeed shows Gemini 3.5 Flash-Lite benchmark context, API price, output speed, first-token latency, and provider data across 167 providers when those signals are available.
- What is the Gemini 3.5 Flash-Lite API price?
- Gemini 3.5 Flash-Lite has pricing from undefined provider} other undefined providers}}, ranging from $0.0065/M to $2.19/M. N1N has the lowest listed price.
- What does the Gemini 3.5 Flash-Lite API pricing table include?
- The Gemini 3.5 Flash-Lite API pricing table compares 167 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest Gemini 3.5 Flash-Lite API pricing?
- N1N currently has the lowest listed Gemini 3.5 Flash-Lite price at $0.0065/M across undefined provider} other undefined providers}}.
- Is Gemini 3.5 Flash-Lite API free?
- Yes, Gemini 3.5 Flash-Lite free API options are available through 4 providerundefined other undefined} on LMSpeed, including 兔子API, 兔子API, 兔子API, 兔子API. These providers offer free API credits or a free tier with no per-token charges.
- Where can I get Gemini 3.5 Flash-Lite free API access?
- LMSpeed currently lists 4 free API providerundefined other undefined} for Gemini 3.5 Flash-Lite: 兔子API, 兔子API, 兔子API, 兔子API. Check each provider row before using it because free tier limits can change.
