Choose a model to open its comparison page.
Llama 3.2 3B Instruct API pricing covers 20 API providers, from $0.010/request to $375.00/M. Llama 3.2 3B Instruct free API options are available from 2 providers.
Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with ...
Input and output token limits for this model, plus how it ranks on long-context understanding.
Compare Llama 3.2 3B Instruct API pricing across 18 providers. Prices range from $0.010/request to $375.00/M. uglycat offers the lowest rate at $0.010/request. 2 providers offer free API credits or a free tier.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
兔子API Free | L1 100% | llama-3.2-3b-instruct | default | Free | Free | — | — | — |
L1 100% | llama-3.2-3b-instruct | default | $0.027/M | $0.027/M | — | — | — | |
L1 100% | llama-3.2-3b-instruct | default | $0.127/M | $0.064/M | — | — | — | |
L1 100% | llama-3.2-3b-instruct | default | $0.127/M | $0.064/M | — | — | — | |
L1 100% | meta/llama-3.2-3b-instruct | 国产模型 | $75.00/M | $75.00/M | — | — | — | |
Dext API Free | L1 100% | llama-3.2-3b-instruct | 公益 | Free | Free | — | — | — |
L1 99% | meta/llama-3.2-3b-instruct | default | $375.00/M | $375.00/M | — | — | — | |
L1 96% | llama-3.2-3b-instruct | 小叶币 | $0.050/request | - | — | — | — | |
L1 0% | meta-llama/llama-3.2-3b-instruct:free | default | $0.010/request | - | — | — | — | |
L1 100% | llama-3.2-3b-instruct | default | $0.068/M | $0.034/M | — | — | — |
minimax-m2-5
MiniMax M2.5 is MiniMax's flagship text model for coding and agents, with SOTA-level programming and agentic performance, improved token efficiency, and fast high-TPS API deployment.
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.
deepseek-v3-2
DeepSeek V3.2 is an upgraded V3-series MoE model with stronger reasoning, coding, and math performance, widely available through OpenAI-compatible API relays.
glm-4-5-air
Zhipu AI GLM-4.5 Air is a lightweight GLM-4.5 variant optimized for low-latency Chinese and English dialogue, retrieval-augmented apps, and edge deployments.
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.