Mistral Small 4 API Benchmarks, Pricing & Provider Data
Compare Mistral Small 4 with another model
Choose a model to open its comparison page.
Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Возможности
Technical Details
- Вход
- Выход
- Всего параметров
- 119B
- Выпущена
- Mar 2026
- Токенизатор
- Mistral
- Архитектура
- text+image->text
- Модерация
- Нет
- Поддерживаемые параметры
- frequency_penaltyinclude_reasoningmax_tokenspresence_penaltyreasoningreasoning_effortresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_p
Эндпоинты OpenRouter
4 эндпоинтовСторонние данные эндпоинтов OpenRouter показаны отдельно от измерений LMSpeed. Некоторые 30-минутные live-метрики появляются только после синхронизации с OpenRouter API key.
| Provider endpoint | Вход | Вывод | Доступность 1 день | Задержка 30м | Пропускная способность 30м | Контекст / вывод |
|---|---|---|---|---|---|---|
Venice venice/fp8 | $0.188/M | $0.750/M | 99.9% | — | — | undefined токенов / undefined токенов |
Mistral mistral/eu | $0.165/M | $0.660/M | 99.9% | — | — | undefined токенов / undefined токенов |
Mistral mistral/zdr | $0.150/M | $0.600/M | 99.9% | — | — | undefined токенов / undefined токенов |
Mistral mistral | $0.150/M | $0.600/M | 99.8% | — | — | undefined токенов / undefined токенов |
Альтернативы и похожие модели
GPT-5.4
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
GPT-5.3 Codex
gpt-5-3-codex
OpenAI GPT-5.3 Codex is a code-specialized variant in the GPT-5 series, optimized for code generation, debugging, and software development tasks.
GPT-5.2
gpt-5-2
OpenAI GPT-5.2 is a GPT-5 series model emphasizing advanced reasoning, multimodal understanding, and high-quality outputs for complex enterprise workloads.
GPT-5.4 Mini
gpt-5-4-mini
OpenAI GPT-5.4 Mini is a compact language model in the GPT-5 series, optimized for quick responses and high throughput.
Claude Opus 4.6
claude-opus-4-6
Anthropic Claude Opus 4.6 is the most capable Claude Opus tier, optimized for complex analysis, long-horizon coding, and high-stakes enterprise reasoning workloads.
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
