Claude Opus 4.8 (Fast) API Benchmarks, Pricing & Provider Data
Compare Claude Opus 4.8 (Fast) with another model
Choose a model to open its comparison page.
Claude Opus 4.8 (Fast) API pricing covers 1 API provider, from $2.06/M to $2.06/M.
Fast-mode variant of [Opus 4.8](/anthropic/claude-opus-4.8) - identical capabilities with higher output speed at 2x pricing relative to regular Opus 4.8. Learn more in Anthropic's docs: https://platf...
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
OpenRouter endpoints
1 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
Anthropic anthropic | $10/M | $50/M | 100% | — | — | undefined tokens / undefined tokens |
Pricing Comparison
Claude Opus 4.8 (Fast) API pricing starts at $2.06/M from 空悲切b2b API.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 0% | anthropic/claude-opus-4.8-fast | or | $2.06/M Cache read$0.206/MCache write$2.58/MCache write 1h$4.12/M | $10.30/M | — | — | — |
Alternatives & Similar Models
GPT-5.4
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
GPT-5.4 Mini
gpt-5-4-mini
OpenAI GPT-5.4 Mini is a compact language model in the GPT-5 series, optimized for quick responses and high throughput.
Claude Opus 4.6
claude-opus-4-6
Anthropic Claude Opus 4.6 is the most capable Claude Opus tier, optimized for complex analysis, long-horizon coding, and high-stakes enterprise reasoning workloads.
Claude Sonnet 4.6
claude-sonnet-4-6
Anthropic Claude Sonnet 4.6 extends the Sonnet line with improved tool use, coding reliability, and long-context performance for everyday production workloads.
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
DeepSeek V4 Pro
deepseek-v4-pro
DeepSeek V4 Pro is the professional-tier DeepSeek V4 model, targeting frontier reasoning, coding, and agent workflows with maximum capability.
Frequently Asked Questions
- What benchmark data does Claude Opus 4.8 (Fast) include?
- LMSpeed shows Claude Opus 4.8 (Fast) benchmark context, API price, output speed, first-token latency, and provider data across 1 providers when those signals are available.
- What is the Claude Opus 4.8 (Fast) API price?
- Claude Opus 4.8 (Fast) has pricing from 1 providers, ranging from $2.06/M to $2.06/M. 空悲切b2b API has the lowest listed price.
- What does the Claude Opus 4.8 (Fast) API pricing table include?
- The Claude Opus 4.8 (Fast) API pricing table compares 1 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest Claude Opus 4.8 (Fast) API pricing?
- 空悲切b2b API currently has the lowest listed Claude Opus 4.8 (Fast) price at $2.06/M across 1 providers.
- Is Claude Opus 4.8 (Fast) API free?
- Claude Opus 4.8 (Fast) does not currently have a free API tier on LMSpeed. All 1 providers charge per token.
