Choose a model to open its comparison page.
Gemma 3n e4b It API pricing covers 9 API providers, from $0.010/request to $375.00/M. Gemma 3n e4b It free API options are available from 1 provider.
Gemma 3n E4B-it is optimized for efficient execution on mobile and low-resource devices, such as phones, laptops, and tablets. It supports multimodal inputs—including text, visual data, and audio—enab...
Input and output token limits for this model, plus how it ranks on long-context understanding.
Compare Gemma 3n e4b It API pricing across 8 providers. Prices range from $0.010/request to $375.00/M. 素墨API offers the lowest rate at $0.010/request. 1 provider offers free API credits or a free tier.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 100% | gemma-3n-e4b-it | default | $0.490/M | $0.490/M | — | — | — | |
L1 100% | gemma-3n-e4b-it | default | $75.00/M | $75.00/M | — | — | — | |
Dext API Free | L1 100% | gemma-3n-e4b-it | 公益 | Free | Free | — | — | — |
L1 99% | google/gemma-3n-e4b-it | default | $375.00/M | $375.00/M | — | — | — | |
L1 98% | google/gemma-3n-e4b-it | default | $75.00/M | $75.00/M | — | — | — | |
L1 98% | google/gemma-3n-e4b-it | default | $0.010/request | - | — | — | — | |
L1 0% | gemma-3n-e4b-it | nvidia | $1.00/request | - | — | — | — | |
L1 0% | gemma-3n-e4b-it | Gemini | $0.011/M | $0.011/M | — | — | — |
deepseek-v3-2
DeepSeek V3.2 is an upgraded V3-series MoE model with stronger reasoning, coding, and math performance, widely available through OpenAI-compatible API relays.
gpt-oss
GPT-OSS is an open-weight language model family designed for self-hosted inference, research, and cost-efficient alternatives to proprietary GPT-class models.
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.
minimax-m2-5
MiniMax M2.5 is MiniMax's flagship text model for coding and agents, with SOTA-level programming and agentic performance, improved token efficiency, and fast high-TPS API deployment.
minimax-m2-7
MiniMax M2.7 is a high-tier M2-series model tuned for complex reasoning, long-context dialogue, and production-grade API workloads.
glm-5-1
Zhipu GLM-5.1 is a next-generation GLM model aimed at frontier reasoning, coding, and bilingual agent applications.