R1 Distill Llama 70B API Benchmarks, Pricing & Provider Data
Compare R1 Distill Llama 70B with another model
Choose a model to open its comparison page.
R1 Distill Llama 70B benchmark, API pricing, and provider data cover 9 API providers, with prices starting at $0.014/M.
DeepSeek R1 Distill Llama 70B is a distilled large language model based on [Llama-3.3-70B-Instruct](/meta-llama/llama-3.3-70b-instruct), using outputs from [DeepSeek R1](/deepseek/deepseek-r1). The mo...
Category Performance
Observed-capability estimates with uncertainty reported separately from independent benchmark families.
- Coverage
- 3 / 8
- Methodology
- V3.0
#1Math55Estimated2/4 Measured dimensions
#2Reasoning45.3Estimated2/4 Measured dimensions
#3Coding42.6Provisional1/4 Measured dimensions
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Technical Details
- Input
- Output
- Total parameters
- 70B
- Released
- Jan 2025
- Knowledge cutoff
- 2024-07-31
- Tokenizer
- Llama3
- Architecture
- text->text
- Instruct type
- deepseek-r1
- Moderated
- No
- Supported parameters
- frequency_penaltyinclude_reasoningmax_tokenspresence_penaltyreasoningrepetition_penaltyseedstoptemperaturetop_ktop_p
Rankings
Excels at
Detailed scores
Updated: Aug 23, 2026Pricing
undefined metric} other undefined metrics}}
Input price$0.700/M#101 / 180Output price$1.10/M#47 / 180
Pricing
undefined metric} other undefined metrics}}
Agents
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Agents
V3.0undefined metric} other undefined metrics}} · No data
Coding
V3.0undefined metric} other undefined metrics}} · Provisional
Score42.680% interval28.7–56.51/4 Measured dimensionsLiveCodeBench26.6%#99 / 115SciCode31.3%#148 / 204
Coding
V3.0undefined metric} other undefined metrics}} · Provisional
Reasoning
V3.0undefined metric} other undefined metrics}} · Estimated
Score45.380% interval34.2–56.52/4 Measured dimensionsMMLU-Pro79.5%#66 / 129GPQA40.2%#199 / 211
Reasoning
V3.0undefined metric} other undefined metrics}} · Estimated
Knowledge
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Knowledge
V3.0undefined metric} other undefined metrics}} · No data
Math
V3.0undefined metric} other undefined metrics}} · Estimated
Score5580% interval43.2–66.82/4 Measured dimensionsAIME67.0%#24 / 68MATH-50093.5%#25 / 73
Math
V3.0undefined metric} other undefined metrics}} · Estimated
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
Multimodal
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Multimodal
V3.0undefined metric} other undefined metrics}} · No data
Instruction following
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Instruction following
V3.0undefined metric} other undefined metrics}} · No data
OpenRouter endpoints
1 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
Novita novita/bf16 | $0.800/M | $0.800/M | 100% | — | — | undefined tokens / undefined tokens |
Pricing Comparison
Compare R1 Distill Llama 70B API pricing across 9 providers. Prices range from $0.014/M to $10.27/M. Yun API offers the lowest rate at $0.014/M.
Alternatives & Similar Models
GPT-5.4
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
GPT-5.3 Codex
gpt-5-3-codex
OpenAI GPT-5.3 Codex is a code-specialized variant in the GPT-5 series, optimized for code generation, debugging, and software development tasks.
GPT-5.2
gpt-5-2
OpenAI GPT-5.2 is a GPT-5 series model emphasizing advanced reasoning, multimodal understanding, and high-quality outputs for complex enterprise workloads.
GPT-5.4 Mini
gpt-5-4-mini
OpenAI GPT-5.4 Mini is a compact language model in the GPT-5 series, optimized for quick responses and high throughput.
Claude Opus 4.6
claude-opus-4-6
Anthropic Claude Opus 4.6 is the most capable Claude Opus tier, optimized for complex analysis, long-horizon coding, and high-stakes enterprise reasoning workloads.
Claude Sonnet 4.6
claude-sonnet-4-6
Anthropic Claude Sonnet 4.6 extends the Sonnet line with improved tool use, coding reliability, and long-context performance for everyday production workloads.
Frequently Asked Questions
- What benchmark data does R1 Distill Llama 70B include?
- LMSpeed shows R1 Distill Llama 70B benchmark context, API price, output speed, first-token latency, and provider data across 9 providers when those signals are available.
- What is the R1 Distill Llama 70B API price?
- R1 Distill Llama 70B has pricing from 9 providers, ranging from $0.014/M to $10.27/M. Yun API has the lowest listed price.
- What does the R1 Distill Llama 70B API pricing table include?
- The R1 Distill Llama 70B API pricing table compares 9 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest R1 Distill Llama 70B API pricing?
- Yun API currently has the lowest listed R1 Distill Llama 70B price at $0.014/M across 9 providers.
- Is R1 Distill Llama 70B API free?
- R1 Distill Llama 70B does not currently have a free API tier on LMSpeed. All 9 providers charge per token.
