Meta
·Released on Jul 23, 2024

Llama 3.1 8B Instruct API Benchmarks, Pricing & Provider Data

Compare Llama 3.1 8B Instruct with another model

Choose a model to open its comparison page.

Share on X
LLM

Llama 3.1 8B Instruct API pricing covers undefined API provider} other undefined API providers}}, from $0.012/M to $0.027/M.

Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 8B instruct-tuned version is fast and efficient. It has demonstrated strong performance compared to...

Cost
$0.012/ 1M · 8:1 in:out
$0.0020 in · $0.010 out

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
131.1Ktokens
157.3 pages of text
OUTPUT
118.0Ktokens
8K128K1M4M
131.1K

Features

Technical Details

Input
Output
Total parameters
8B
Released
Jul 2024
Knowledge cutoff
2023-12-31
Tokenizer
Llama3
Architecture
text->text
Instruct type
llama3
Moderated
No
Supported parameters
frequency_penaltylogit_biaslogprobsmax_tokensmin_ppresence_penaltyrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p

OpenRouter endpoints

5 endpoints

Third-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.

Provider endpointInputOutput1d uptime30m latency30m throughputContext / output
CoreWeave
coreweave/bf16
$0.220/M$0.220/M100.0%undefined tokens / undefined tokens
Groq
groq
$0.050/M$0.080/M99.9%undefined tokens / undefined tokens
DeepInfra
deepinfra/fp8
$0.020/M$0.040/M99.9%undefined tokens / undefined tokens
Cloudflare
cloudflare/fp8
$0.152/M$0.287/M99.7%undefined tokens / undefined tokens
Novita
novita/fp8
$0.020/M$0.050/M99.5%undefined tokens / undefined tokens

Pricing Comparison

Compare Llama 3.1 8B Instruct API pricing across 4 providers. Prices range from $0.012/M to $0.027/M. 云AI offers the lowest rate at $0.012/M.

ProviderHealthModel VariantGroupInput ($/M)Output ($/M)Speed (t/s)First tokenAudit
L1
100%
llama-3.1-8b-instruct
azure03-0
$0.012/M
$0.026/M
L1
100%
llama-3.1-8b-instruct
default
$0.027/M
$0.041/M
L1
100%
laohuang/meta/llama-3.1-8b-instruct
default
$0.020/request
-

Alternatives & Similar Models

Frequently Asked Questions

What benchmark data does Llama 3.1 8B Instruct include?
LMSpeed shows Llama 3.1 8B Instruct benchmark context, API price, output speed, first-token latency, and provider data across 4 providers when those signals are available.
What is the Llama 3.1 8B Instruct API price?
Llama 3.1 8B Instruct has pricing from undefined provider} other undefined providers}}, ranging from $0.012/M to $0.027/M. 云AI has the lowest listed price.
What does the Llama 3.1 8B Instruct API pricing table include?
The Llama 3.1 8B Instruct API pricing table compares 4 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
Which provider has the cheapest Llama 3.1 8B Instruct API pricing?
云AI currently has the lowest listed Llama 3.1 8B Instruct price at $0.012/M across undefined provider} other undefined providers}}.
Is Llama 3.1 8B Instruct API free?
Llama 3.1 8B Instruct does not currently have a free API tier on LMSpeed. All 4 providers charge per token.

Also known as

laohuang/meta/llama-3.1-8b-instructllama-3.1-8b-instructmeta-llama/llama-3.1-8b-instructmeta/llama-3.1-8b-instruct

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation