Mistral Nemo API Benchmarks, Pricing & Provider Data
Compare Mistral Nemo with another model
Choose a model to open its comparison page.
Mistral Nemo API pricing covers 3 API providers, from $57.53/M to $75.00/M. The page also shows measured API speed and first-token latency.
Mistral Nemo is a 12B open model co-developed by Mistral AI and NVIDIA, optimized for multilingual chat and function calling.
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Technical Details
- Input
- Output
- Released
- Jul 2024
- Knowledge cutoff
- 2024-04-30
- Tokenizer
- Mistral
- Architecture
- text->text
- Instruct type
- mistral
- Moderated
- No
- Supported parameters
- frequency_penaltylogit_biaslogprobsmax_tokensmin_ppresence_penaltyrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
OpenRouter endpoints
4 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
Parasail parasail/fp8 | $0.030/M | $0.030/M | 100.0% | — | — | undefined tokens / undefined tokens |
Io Net io-net/fp16 | $0.044/M | $0.160/M | 96.0% | — | — | undefined tokens / undefined tokens |
Novita novita/fp8 | $0.040/M | $0.170/M | 80.0% | — | — | undefined tokens / undefined tokens |
DeepInfra deepinfra/fp8 | $0.019/M | $0.030/M | 57.3% | — | — | undefined tokens / undefined tokens |
Pricing Comparison
Compare Mistral Nemo API pricing across 3 providers. Prices range from $57.53/M to $75.00/M. KFCV50 offers the lowest rate at $57.53/M.
Alternatives & Similar Models
GPT-OSS
gpt-oss
GPT-OSS is an open-weight language model family designed for self-hosted inference, research, and cost-efficient alternatives to proprietary GPT-class models.
GLM-5
glm-5
Zhipu GLM-5 is Zhipu flagship GLM series model with enhanced reasoning, agent capabilities, and strong performance on Chinese enterprise and coding scenarios.
DeepSeek V3.2
deepseek-v3-2
DeepSeek V3.2 is an upgraded V3-series MoE model with stronger reasoning, coding, and math performance, widely available through OpenAI-compatible API relays.
MiniMax M2.5
minimax-m2-5
MiniMax M2.5 is MiniMax's flagship text model for coding and agents, with SOTA-level programming and agentic performance, improved token efficiency, and fast high-TPS API deployment.
MiniMax M2.7
minimax-m2-7
MiniMax M2.7 is a high-tier M2-series model tuned for complex reasoning, long-context dialogue, and production-grade API workloads.
GLM-4.7
glm-4-7
Zhipu GLM-4.7 is a flagship GLM release from Zhipu AI with advanced Chinese-English reasoning, coding, and agent features.
Frequently Asked Questions
- What benchmark data does Mistral Nemo include?
- LMSpeed shows Mistral Nemo benchmark context, API price, output speed, first-token latency, and provider data across 3 providers when those signals are available.
- What is the Mistral Nemo API price?
- Mistral Nemo has pricing from 3 providers, ranging from $57.53/M to $75.00/M. KFCV50 has the lowest listed price.
- What does the Mistral Nemo API pricing table include?
- The Mistral Nemo API pricing table compares 3 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest Mistral Nemo API pricing?
- KFCV50 currently has the lowest listed Mistral Nemo price at $57.53/M across 3 providers.
- Can I compare Mistral Nemo API price and speed together?
- Yes. LMSpeed shows Mistral Nemo API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
- Is Mistral Nemo API free?
- Mistral Nemo does not currently have a free API tier on LMSpeed. All 3 providers charge per token.
