Google
·Released on Apr 2, 2026

Gemma 4 31B API Benchmarks, Pricing & Provider Data

Compare Gemma 4 31B with another model

Choose a model to open its comparison page.

Share on X
LLM

Gemma 4 31B API pricing covers undefined API provider} other undefined API providers}}, from $0.00005/request to $75.00/M.

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, nat...

Cost
$0.00005/ 1M · 8:1 in:out
$0.000008 in · $0.000042 out

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
262.1Ktokens
314.6 pages of text
OUTPUT
16.4Ktokens
8K128K1M4M
262.1K

Features

Technical Details

Input
Output
Total parameters
31B
Released
Apr 2026
Tokenizer
Gemma
Architecture
text+image+video->text
Moderated
No
Supported parameters
frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p

OpenRouter endpoints

14 endpoints

Third-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.

Provider endpointInputOutput1d uptime30m latency30m throughputContext / output
ModelRun
modelrun/fp4
$0.750/M$1/M99.9%undefined tokens / undefined tokens
Parasail
parasail/fp8
$0.150/M$0.400/M99.7%undefined tokens / undefined tokens
Friendli
friendli
$0.140/M$0.400/M99.6%undefined tokens / undefined tokens
DeepInfra
deepinfra/turbo
$0.090/M$0.340/M99.0%undefined tokens / undefined tokens
Venice
venice/bf16
$0.120/M$0.360/M98.2%undefined tokens / undefined tokens
CoreWeave
coreweave/fp4
$0.100/M$0.340/M98.1%undefined tokens / undefined tokens
Crusoe
crusoe
$0.140/M$0.400/M97.3%undefined tokens / undefined tokens
DeepInfra
deepinfra/fp8
$0.130/M$0.380/M97.3%undefined tokens / undefined tokens
Chutes
chutes/fp4
$0.120/M$0.370/M92.4%undefined tokens / undefined tokens
Novita
novita/bf16
$0.140/M$0.400/M90.2%undefined tokens / undefined tokens
Together
together
$0.390/M$0.970/M90.1%undefined tokens / undefined tokens
SambaNova
sambanova
$0.380/M$1.15/M82.7%undefined tokens / undefined tokens
SiliconFlow
siliconflow/fp8
$0.750/M$1/M81.3%undefined tokens / undefined tokens
DeepInfra
deepinfra/ultra
$0.270/M$0.760/M76.5%undefined tokens / undefined tokens

Pricing Comparison

Compare Gemma 4 31B API pricing across 19 providers. Prices range from $0.00005/request to $75.00/M. CM-API 公益站 offers the lowest rate at $0.00005/request.

ProviderHealthModel VariantGroupInput ($/M)Output ($/M)Speed (t/s)First tokenAudit
L1
100%
google/gemma-4-31b-it
default
$0.200/M
$0.200/M
L1
99%
google/gemma-4-31b-it:free
free
$0.00005/request
-
L1
99%
gemma-4-31b-it
gpt
$75.00/M
$75.00/M
L1
99%
gemma-4-31b-it
default
$0.130/M
Cache read$0.026/M
$0.380/M
L1
99%
gemma-4-31b-it
default
$0.137/request
-
L1
97%
google/gemma-4-31B-it
国产模型
$0.390/M
Cache read$0.078/M
$0.970/M
L1
92%
google/gemma-4-31B-it
CC
$2.04/M
$5.84/M
L1
100%
gemma-4-31b-it
default
$0.010/request
-
L1
100%
google/gemma-4-31b-it
default
$0.015/M
Cache read$0.011/M
$0.048/M
L1
100%
google/gemma-4-31b-it
openrouter
$0.022/M
Cache read$0.0044/M
$0.096/M
L1
100%
google/gemma-4-31b-it
default
$36.00/M
$36.00/M
L1
100%
laohuang/google/gemma-4-31b-it
default
$0.020/request
-
L1
100%
google/gemma-4-31b-it
default
$0.020/request
-
L1
0%
google/gemma-4-31b-it
default
$75.00/M
$75.00/M
L1
0%
google/gemma-4-31b-it:free
default
$0.010/request
-
L1
0%
google/gemma-4-31b-it
default
$0.200/M
$0.200/M
L1
0%
google/gemma-4-31b-it:free
OpenAI
$10.27/M
$10.27/M

Alternatives & Similar Models

Frequently Asked Questions

What benchmark data does Gemma 4 31B include?
LMSpeed shows Gemma 4 31B benchmark context, API price, output speed, first-token latency, and provider data across 19 providers when those signals are available.
What is the Gemma 4 31B API price?
Gemma 4 31B has pricing from undefined provider} other undefined providers}}, ranging from $0.00005/request to $75.00/M. CM-API 公益站 has the lowest listed price.
What does the Gemma 4 31B API pricing table include?
The Gemma 4 31B API pricing table compares 19 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
Which provider has the cheapest Gemma 4 31B API pricing?
CM-API 公益站 currently has the lowest listed Gemma 4 31B price at $0.00005/request across undefined provider} other undefined providers}}.
Is Gemma 4 31B API free?
Gemma 4 31B does not currently have a free API tier on LMSpeed. All 19 providers charge per token.

Also known as

gemma-4-31b-itgoogle/gemma-4-31B-itgoogle/gemma-4-31b-itgoogle/gemma-4-31b-it:freelaohuang/google/gemma-4-31b-it

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation