Choose a model to open its comparison page.
FLUX.2 Klein 4B API pricing covers 3 API providers, from $1.00/M to $60.00/M. FLUX.2 Klein 4B free API options are available from 1 provider.
FLUX.2 [klein] 4B is the fastest and most cost-effective model in the FLUX.2 family, optimized for high-throughput use cases while maintaining excellent image quality. Pricing is based on the output.....
Input and output token limits for this model, plus how it ranks on long-context understanding.
Compare FLUX.2 Klein 4B API pricing across 2 providers. Prices range from $1.00/M to $60.00/M. zeabur API offers the lowest rate at $1.00/M. 1 provider offers free API credits or a free tier.
| Провайдер | Работоспособность | Вариант модели | Группа | Входные данные ($/M) | Выходные данные ($/M) | Скорость (т/с) | Первый токен | Аудит |
|---|---|---|---|---|---|---|---|---|
Dext API Бесплатно | L1 100% | flux-2-klein-4b | 公益 | Бесплатно | Бесплатно | — | — | — |
L1 100% | flux-2-klein-4b | default | $60.00/M | $60.00/M | — | — | — | |
L1 97% | black-forest-labs/FLUX.2-klein-4B | 国产模型 | $1.00/M | $1.00/M | — | — | — |
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
gpt-5-3-codex
OpenAI GPT-5.3 Codex is a code-specialized variant in the GPT-5 series, optimized for code generation, debugging, and software development tasks.
gpt-5-4-mini
OpenAI GPT-5.4 Mini is a compact language model in the GPT-5 series, optimized for quick responses and high throughput.
claude-opus-4-6
Anthropic Claude Opus 4.6 is the most capable Claude Opus tier, optimized for complex analysis, long-horizon coding, and high-stakes enterprise reasoning workloads.
claude-sonnet-4-6
Anthropic Claude Sonnet 4.6 extends the Sonnet line with improved tool use, coding reliability, and long-context performance for everyday production workloads.
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.