Meta
·Released on Sep 25, 2024

Llama 3.2 1B Instruct API Benchmarks, Pricing & Provider Data

Compare Llama 3.2 1B Instruct with another model

Choose a model to open its comparison page.

Share on X
LLM

Llama 3.2 1B Instruct API pricing covers undefined API provider} other undefined API providers}}, from $0.0014/request to $0.111/M. Llama 3.2 1B Instruct free API options are available from undefined provider} other undefined providers}}.

Llama 3.2 1B is a 1-billion-parameter language model focused on efficiently performing natural language tasks, such as summarization, dialogue, and multilingual text analysis. Its smaller size allows ...

Cost
$0.0014/ 1M · 8:1 in:out
$0.0002 in · $0.0012 out

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
60Ktokens
72 pages of text
OUTPUT
54Ktokens
8K128K1M4M
60K

Technical Details

Input
Output
Total parameters
1B
Released
Sep 2024
Knowledge cutoff
2023-12-31
Tokenizer
Llama3
Architecture
text->text
Instruct type
llama3
Moderated
No
Supported parameters
frequency_penaltylogit_biasmax_tokensmin_ppresence_penaltyrepetition_penaltyseedstoptemperaturetop_ktop_p

OpenRouter endpoints

1 endpoints

Third-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.

Provider endpointInputOutput1d uptime30m latency30m throughputContext / output
Cloudflare
cloudflare
$0.027/M$0.201/M100%undefined tokens / undefined tokens

Pricing Comparison

Compare Llama 3.2 1B Instruct API pricing across 10 providers. Prices range from $0.0014/request to $0.111/M. 神马中转API offers the lowest rate at $0.0014/request. 2 providers offer free API credits or a free tier.

ProviderHealthModel VariantGroupInput ($/M)Output ($/M)Speed (t/s)First tokenAudit
L1
100%
llama-3.2-1b-instruct
default
$0.0014/request
-
L1
99%
llama-3.2-1b-instruct
default
Free
Free
L1
100%
llama-3.2-1b-instruct
Self-Deployed-2
$0.074/M
$0.019/M
L1
100%
llama-3.2-1b-instruct
Self-Deployed-2
$0.074/M
$0.019/M
L1
52%
llama-3.2-1b-instruct
公益
Free
Free
L1
100%
llama-3.2-1b-instruct
default
$0.034/M
$0.0086/M
L1
99%
laohuang/meta/llama-3.2-1b-instruct
default
$0.020/request
-

Alternatives & Similar Models

Frequently Asked Questions

What benchmark data does Llama 3.2 1B Instruct include?
LMSpeed shows Llama 3.2 1B Instruct benchmark context, API price, output speed, first-token latency, and provider data across 12 providers when those signals are available.
What is the Llama 3.2 1B Instruct API price?
Llama 3.2 1B Instruct has pricing from undefined provider} other undefined providers}}, ranging from $0.0014/request to $0.111/M. 神马中转API has the lowest listed price.
What does the Llama 3.2 1B Instruct API pricing table include?
The Llama 3.2 1B Instruct API pricing table compares 12 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
Which provider has the cheapest Llama 3.2 1B Instruct API pricing?
神马中转API currently has the lowest listed Llama 3.2 1B Instruct price at $0.0014/request across undefined provider} other undefined providers}}.
Is Llama 3.2 1B Instruct API free?
Yes, Llama 3.2 1B Instruct free API options are available through 2 providerundefined other undefined} on LMSpeed, including Dext API, 兔子API. These providers offer free API credits or a free tier with no per-token charges.
Where can I get Llama 3.2 1B Instruct free API access?
LMSpeed currently lists 2 free API providerundefined other undefined} for Llama 3.2 1B Instruct: Dext API, 兔子API. Check each provider row before using it because free tier limits can change.

Also known as

laohuang/meta/llama-3.2-1b-instructllama-3.2-1b-instructmeta-llama/llama-3.2-1b-instructmeta/llama-3.2-1b-instruct

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation