Mistral Small 3.2 24B API Benchmarks, Pricing & Provider Data
Compare Mistral Small 3.2 24B with another model
Choose a model to open its comparison page.
Mistral-Small-3.2-24B-Instruct-2506 is an updated 24B parameter model from Mistral optimized for instruction following, repetition reduction, and improved function calling. Compared to the 3.1 release...
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Возможности
Technical Details
- Вход
- Выход
- Всего параметров
- 24B
- Выпущена
- Jun 2025
- Срез знаний
- 2023-10-31
- Токенизатор
- Mistral
- Архитектура
- text+image->text
- Модерация
- Нет
- Поддерживаемые параметры
- frequency_penaltylogit_biaslogprobsmax_tokensmin_ppresence_penaltyrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
Эндпоинты OpenRouter
2 эндпоинтовСторонние данные эндпоинтов OpenRouter показаны отдельно от измерений LMSpeed. Некоторые 30-минутные live-метрики появляются только после синхронизации с OpenRouter API key.
| Provider endpoint | Вход | Вывод | Доступность 1 день | Задержка 30м | Пропускная способность 30м | Контекст / вывод |
|---|---|---|---|---|---|---|
Parasail parasail/bf16 | $0.090/M | $0.300/M | 99.9% | — | — | undefined токенов / undefined токенов |
DeepInfra deepinfra/fp8 | $0.075/M | $0.200/M | 99.8% | — | — | undefined токенов / undefined токенов |
Альтернативы и похожие модели
GPT-5.4
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
GPT-5.3 Codex
gpt-5-3-codex
OpenAI GPT-5.3 Codex is a code-specialized variant in the GPT-5 series, optimized for code generation, debugging, and software development tasks.
GPT-5.2
gpt-5-2
OpenAI GPT-5.2 is a GPT-5 series model emphasizing advanced reasoning, multimodal understanding, and high-quality outputs for complex enterprise workloads.
GPT-5.4 Mini
gpt-5-4-mini
OpenAI GPT-5.4 Mini is a compact language model in the GPT-5 series, optimized for quick responses and high throughput.
Claude Opus 4.6
claude-opus-4-6
Anthropic Claude Opus 4.6 is the most capable Claude Opus tier, optimized for complex analysis, long-horizon coding, and high-stakes enterprise reasoning workloads.
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
