GPT-5 Chat API Benchmarks, Pricing & Provider Data
Compare GPT-5 Chat with another model
Choose a model to open its comparison page.
GPT-5 Chat API pricing covers undefined API provider} other undefined API providers}}, from $0.034/M to $3.75/M.
GPT-5 Chat is designed for advanced, natural, multimodal, and context-aware conversations for enterprise applications.
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Pricing Comparison
Compare GPT-5 Chat API pricing across 10 providers. Prices range from $0.034/M to $3.75/M. 云AI offers the lowest rate at $0.034/M.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 100% | gpt-5-chat | AzuerPromo | $0.034/M Cache read$0.0034/M | $0.274/M | — | — | — | |
L1 100% | gpt-5-chat-2025-08-07 | azure03-0 | $0.051/M Cache read$0.0051/M | $0.411/M | — | — | — | |
L1 100% | gpt-5-chat | Azure-GPT | $3.75/M Cache read$0.390/M | $30.00/M | — | — | — | |
L1 100% | gpt-5-chat-2025-08-07 | Azure-GPT | $3.75/M | $30.00/M | — | — | — | |
L1 99% | openai/gpt-5-chat | default | $1.10/M | $9.00/M | — | — | — |
Alternatives & Similar Models
GPT-5.4
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
GPT-5.3 Codex
gpt-5-3-codex
OpenAI GPT-5.3 Codex is a code-specialized variant in the GPT-5 series, optimized for code generation, debugging, and software development tasks.
GPT-5.2
gpt-5-2
OpenAI GPT-5.2 is a GPT-5 series model emphasizing advanced reasoning, multimodal understanding, and high-quality outputs for complex enterprise workloads.
GPT-5.4 Mini
gpt-5-4-mini
OpenAI GPT-5.4 Mini is a compact language model in the GPT-5 series, optimized for quick responses and high throughput.
Claude Opus 4.6
claude-opus-4-6
Anthropic Claude Opus 4.6 is the most capable Claude Opus tier, optimized for complex analysis, long-horizon coding, and high-stakes enterprise reasoning workloads.
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
Frequently Asked Questions
- What benchmark data does GPT-5 Chat include?
- LMSpeed shows GPT-5 Chat benchmark context, API price, output speed, first-token latency, and provider data across 10 providers when those signals are available.
- What is the GPT-5 Chat API price?
- GPT-5 Chat has pricing from undefined provider} other undefined providers}}, ranging from $0.034/M to $3.75/M. 云AI has the lowest listed price.
- What does the GPT-5 Chat API pricing table include?
- The GPT-5 Chat API pricing table compares 10 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest GPT-5 Chat API pricing?
- 云AI currently has the lowest listed GPT-5 Chat price at $0.034/M across undefined provider} other undefined providers}}.
- Is GPT-5 Chat API free?
- GPT-5 Chat does not currently have a free API tier on LMSpeed. All 10 providers charge per token.
