Choose a model to open its comparison page.
Mistral Nemo API pricing covers 3 API providers, from $57.53/M to $75.00/M. The page also shows measured API speed and first-token latency.
Mistral Nemo is a 12B open model co-developed by Mistral AI and NVIDIA, optimized for multilingual chat and function calling.
Input and output token limits for this model, plus how it ranks on long-context understanding.
Сторонние данные эндпоинтов OpenRouter показаны отдельно от измерений LMSpeed. Некоторые 30-минутные live-метрики появляются только после синхронизации с OpenRouter API key.
| Provider endpoint | Вход | Вывод | Доступность 1 день | Задержка 30м | Пропускная способность 30м | Контекст / вывод |
|---|---|---|---|---|---|---|
DeepInfra deepinfra/fp8 | $0.020/M | $0.040/M | 99.7% | — | — | 131.1K токенов / 16.4K токенов |
Mistral mistral | $0.150/M | $0.150/M | 99.7% | — | — | 131.1K токенов / — |
DekaLLM dekallm/fp8 | $0.020/M | $0.030/M | 95.4% | — | — | 131.1K токенов / — |
Novita novita/fp8 | $0.040/M | $0.170/M | 75.1% | — | — | 60.3K токенов / 16K токенов |
Compare Mistral Nemo API pricing across 3 providers. Prices range from $57.53/M to $75.00/M. KFCV50 offers the lowest rate at $57.53/M.
| Провайдер | Работоспособность | Вариант модели | Группа | Входные данные ($/M) | Выходные данные ($/M) | Скорость (т/с) | Первый токен | Аудит |
|---|---|---|---|---|---|---|---|---|
L1 100% | mistral-nemo | enterprise_emergency_standby | $57.53/M | $172.60/M | — | — | — | |
L1 100% | mistralai/mistral-nemo | default | $75.00/M | $75.00/M | — | — | — | |
L1 100% | mistralai/mistral-nemo:free | default | $75.00/M | $75.00/M | — | — | — |
gpt-oss
GPT-OSS is an open-weight language model family designed for self-hosted inference, research, and cost-efficient alternatives to proprietary GPT-class models.
deepseek-v3-2
DeepSeek V3.2 is an upgraded V3-series MoE model with stronger reasoning, coding, and math performance, widely available through OpenAI-compatible API relays.
minimax-m2-5
MiniMax M2.5 is MiniMax's flagship text model for coding and agents, with SOTA-level programming and agentic performance, improved token efficiency, and fast high-TPS API deployment.
glm-5
Zhipu GLM-5 is Zhipu flagship GLM series model with enhanced reasoning, agent capabilities, and strong performance on Chinese enterprise and coding scenarios.
minimax-m2-7
MiniMax M2.7 is a high-tier M2-series model tuned for complex reasoning, long-context dialogue, and production-grade API workloads.
glm-4-7
Zhipu GLM-4.7 is a flagship GLM release from Zhipu AI with advanced Chinese-English reasoning, coding, and agent features.