Choose a model to open its comparison page.
Llama 3.2 1B Instruct API pricing covers 18 API providers, from $0.0014/request to $375.00/M. Llama 3.2 1B Instruct free API options are available from 2 providers.
Llama 3.2 1B is a 1-billion-parameter language model focused on efficiently performing natural language tasks, such as summarization, dialogue, and multilingual text analysis. Its smaller size allows ...
Input and output token limits for this model, plus how it ranks on long-context understanding.
Compare Llama 3.2 1B Instruct API pricing across 16 providers. Prices range from $0.0014/request to $375.00/M. 神马中转API offers the lowest rate at $0.0014/request. 2 providers offer free API credits or a free tier.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
兔子API Free | L1 100% | llama-3.2-1b-instruct | default | Free | Free | — | — | — |
L1 100% | llama-3.2-1b-instruct | default | $0.0014/request | - | — | — | — | |
L1 100% | llama-3.2-1b-instruct | default | $0.064/M | $0.016/M | — | — | — | |
L1 100% | llama-3.2-1b-instruct | default | $0.064/M | $0.016/M | — | — | — | |
Dext API Free | L1 100% | llama-3.2-1b-instruct | 公益 | Free | Free | — | — | — |
L1 99% | meta/llama-3.2-1b-instruct | default | $375.00/M | $375.00/M | — | — | — | |
L1 96% | llama-3.2-1b-instruct | 小叶币 | $0.050/request | - | — | — | — | |
L1 100% | llama-3.2-1b-instruct | default | $0.034/M | $0.0086/M | — | — | — |
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
gpt-5-3-codex
OpenAI GPT-5.3 Codex is a code-specialized variant in the GPT-5 series, optimized for code generation, debugging, and software development tasks.
gpt-5-2
OpenAI GPT-5.2 is a GPT-5 series model emphasizing advanced reasoning, multimodal understanding, and high-quality outputs for complex enterprise workloads.
gpt-5-4-mini
OpenAI GPT-5.4 Mini is a compact language model in the GPT-5 series, optimized for quick responses and high throughput.
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
deepseek-v4-pro
DeepSeek V4 Pro is the professional-tier DeepSeek V4 model, targeting frontier reasoning, coding, and agent workflows with maximum capability.