Meta
·Released on Jul 23, 2024

Llama 3.1 70B Instruct API Benchmarks, Pricing & Provider Data

Compare Llama 3.1 70B Instruct with another model

Choose a model to open its comparison page.

Share on X
LLM

Llama 3.1 70B Instruct API pricing covers undefined API provider} other undefined API providers}}, from $0.020/request to $0.020/request.

Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 70B instruct-tuned version is optimized for high quality dialogue usecases. It has demonstrated strong...

Cost
$0.020/ 1M · 8:1 in:out
$0.0032 in · $0.017 out

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
131.1Ktokens
157.3 pages of text
OUTPUT
16.4Ktokens
8K128K1M4M
131.1K

Features

Technical Details

Input
Output
Total parameters
70B
Released
Jul 2024
Knowledge cutoff
2023-12-31
Tokenizer
Llama3
Architecture
text->text
Instruct type
llama3
Moderated
No
Supported parameters
frequency_penaltylogit_biasmax_tokensmin_ppresence_penaltyrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_p

OpenRouter endpoints

2 endpoints

Third-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.

Provider endpointInputOutput1d uptime30m latency30m throughputContext / output
Amazon Bedrock
amazon-bedrock
$0.720/M$0.720/M100.0%undefined tokens / undefined tokens
DeepInfra
deepinfra/turbo
$0.400/M$0.400/M99.9%undefined tokens / undefined tokens

Pricing Comparison

Compare Llama 3.1 70B Instruct API pricing across 2 providers. Prices range from $0.020/request to $0.020/request. 91VIP API offers the lowest rate at $0.020/request.

ProviderHealthModel VariantGroupInput ($/M)Output ($/M)Speed (t/s)First tokenAudit
L1
100%
laohuang/meta/llama-3.1-70b-instruct
default
$0.020/request
-

Alternatives & Similar Models

Frequently Asked Questions

What benchmark data does Llama 3.1 70B Instruct include?
LMSpeed shows Llama 3.1 70B Instruct benchmark context, API price, output speed, first-token latency, and provider data across 2 providers when those signals are available.
What is the Llama 3.1 70B Instruct API price?
Llama 3.1 70B Instruct has pricing from undefined provider} other undefined providers}}, ranging from $0.020/request to $0.020/request. 91VIP API has the lowest listed price.
What does the Llama 3.1 70B Instruct API pricing table include?
The Llama 3.1 70B Instruct API pricing table compares 2 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
Which provider has the cheapest Llama 3.1 70B Instruct API pricing?
91VIP API currently has the lowest listed Llama 3.1 70B Instruct price at $0.020/request across undefined provider} other undefined providers}}.
Is Llama 3.1 70B Instruct API free?
Llama 3.1 70B Instruct does not currently have a free API tier on LMSpeed. All 2 providers charge per token.

Also known as

laohuang/meta/llama-3.1-70b-instructllama-3.1-70b-instructmeta-llama/llama-3-1-70b-instructmeta-llama/llama-3.1-70b-instructmeta/llama-3.1-70b-instruct

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation