Gemini 3.1 Flash API Benchmarks, Pricing & Provider Data
Compare Gemini 3.1 Flash with another model
Choose a model to open its comparison page.
Gemini 3.1 Flash API pricing covers 20 API providers, from $0.100/request to $75.00/M.
Google Gemini 3.1 Flash extends the Gemini 3 Flash line with improved reasoning and multimodal accuracy for production assistants and search-augmented apps.
Specifications
Pricing Comparison
Compare Gemini 3.1 Flash API pricing across 20 providers. Prices range from $0.100/request to $75.00/M. 初叶🍂Furry API offers the lowest rate at $0.100/request.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 100% | gemini-3.1-flash-preview | gemini_studio_no_lag_20 | $0.137/M | $0.822/M | — | — | — | |
L1 100% | gemini-3.1-flash-preview-maxthinking | gemini_mix_0329_20 | $0.137/M | $0.411/M | — | — | — | |
L1 100% | gemini-3.1-flash-preview-nothinking | gemini_mix_0329_20 | $0.137/M | $0.411/M | — | — | — | |
L1 100% | gemini-3.1-flash-preview | 优质gemini | $0.103/M | $0.616/M | — | — | — | |
L1 100% | 假流式-gemini-3.1-flash | default | $50.00/request | - | — | — | — | |
L1 99% | gemini-3.1-flash-preview | gemini-号池 | $0.205/M | $1.23/M | — | — | — | |
L1 99% | gemini-3.1-flash | default | $0.400/M Cache read$0.040/M | $2.40/M | — | — | — | |
L1 99% | gemini-3.1-flash-preview | chongzhi | $30.82/M | $92.47/M | — | — | — | |
L1 100% | gemini-3.1-flash-preview | 企业用户 | $75.00/M | $225.00/M | — | — | — | |
L1 0% | gemini-3.1-flash | 速率限制-小叶币 | $0.100/request | - | — | — | — |
Alternatives & Similar Models
GPT-5.4
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
Claude Opus 4.6
claude-opus-4-6
Anthropic Claude Opus 4.6 is the most capable Claude Opus tier, optimized for complex analysis, long-horizon coding, and high-stakes enterprise reasoning workloads.
Claude Sonnet 4.6
claude-sonnet-4-6
Anthropic Claude Sonnet 4.6 extends the Sonnet line with improved tool use, coding reliability, and long-context performance for everyday production workloads.
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
DeepSeek V4 Pro
deepseek-v4-pro
DeepSeek V4 Pro is the professional-tier DeepSeek V4 model, targeting frontier reasoning, coding, and agent workflows with maximum capability.
GLM-5.1
glm-5-1
Zhipu GLM-5.1 is a next-generation GLM model aimed at frontier reasoning, coding, and bilingual agent applications.
Frequently Asked Questions
- What benchmark data does Gemini 3.1 Flash include?
- LMSpeed shows Gemini 3.1 Flash benchmark context, API price, output speed, first-token latency, and provider data across 20 providers when those signals are available.
- What is the Gemini 3.1 Flash API price?
- Gemini 3.1 Flash has pricing from 20 providers, ranging from $0.100/request to $75.00/M. 初叶🍂Furry API has the lowest listed price.
- What does the Gemini 3.1 Flash API pricing table include?
- The Gemini 3.1 Flash API pricing table compares 20 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest Gemini 3.1 Flash API pricing?
- 初叶🍂Furry API currently has the lowest listed Gemini 3.1 Flash price at $0.100/request across 20 providers.
- Is Gemini 3.1 Flash API free?
- Gemini 3.1 Flash does not currently have a free API tier on LMSpeed. All 20 providers charge per token.
