Choose a model to open its comparison page.
Gemma 4 31B API pricing covers 18 API providers, from $0.0010/request to $375.00/M. Gemma 4 31B free API options are available from 2 providers.
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, nat...
Input and output token limits for this model, plus how it ranks on long-context understanding.
Compare Gemma 4 31B API pricing across 16 providers. Prices range from $0.0010/request to $375.00/M. CM-API 公益站 offers the lowest rate at $0.0010/request. 2 providers offer free API credits or a free tier.
| Провайдер | Работоспособность | Вариант модели | Группа | Входные данные ($/M) | Выходные данные ($/M) | Скорость (т/с) | Первый токен | Аудит |
|---|---|---|---|---|---|---|---|---|
L1 100% | gemma-4-31b-it | default | $0.0014/request | - | — | — | — | |
L1 100% | google/gemma-4-31b-it:free | free | $0.0010/request | - | — | — | — | |
L1 100% | gemma-4-31b-it | default | $0.130/M Cache read$0.026/M | $0.380/M | — | — | — | |
Dext API Бесплатно | L1 100% | gemma-4-31b-it | 公益 | Бесплатно | Бесплатно | — | — | — |
L1 100% | gemma-4-31b-it | default | $1.00/M | $3.00/M | — | — | — | |
L1 99% | google/gemma-4-31b-it | default | $10.00/M | $10.00/M | — | — | — | |
L1 99% | gemma-4-31b-it | default | $10.00/M | $10.00/M | — | — | — | |
L1 99% | google/gemma-4-31b-it:free | default | $375.00/M | $375.00/M | — | — | — | |
L1 97% | google/gemma-4-31B-it | 国产模型 | $0.390/M Cache read$0.078/M | $0.970/M | — | — | — | |
L1 77% | google/gemma-4-31B-it | CC | $2.04/M | $5.84/M | — | — | — | |
L1 0% | google/gemma-4-31b-it:free | default | $0.010/request | - | — | — | — | |
L1 100% | google/gemma-4-31b-it | default | $75.00/M | $75.00/M | — | — | — | |
L1 99% | google/gemma-4-31b-it | default | $0.015/M Cache read$0.011/M | $0.048/M | — | — | — | |
91VIP API Бесплатно | L1 95% | google/gemma-4-31b-it | default | Бесплатно | Бесплатно | — | — | — |
L1 0% | google/gemma-4-31b-it | default | $0.300/M | $0.300/M | — | — | — | |
L1 0% | gemma-4-31b-it | Model | $0.0096/M | $0.027/M | — | — | — | |
L1 0% | gemma-4-31b-it | default | $1.00/M | $3.00/M | — | — | — | |
L1 0% | google/gemma-4-31b-it:free | OpenAI | $10.27/M | $10.27/M | — | — | — |
minimax-m2-7
MiniMax M2.7 is a high-tier M2-series model tuned for complex reasoning, long-context dialogue, and production-grade API workloads.
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.
deepseek-v3-2
DeepSeek V3.2 is an upgraded V3-series MoE model with stronger reasoning, coding, and math performance, widely available through OpenAI-compatible API relays.
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
glm-5
Zhipu GLM-5 is Zhipu flagship GLM series model with enhanced reasoning, agent capabilities, and strong performance on Chinese enterprise and coding scenarios.
minimax-m2-5
MiniMax M2.5 is MiniMax's flagship text model for coding and agents, with SOTA-level programming and agentic performance, improved token efficiency, and fast high-TPS API deployment.