Mercury 2 API Benchmarks, Pricing & Provider Data
Compare Mercury 2 with another model
Choose a model to open its comparison page.
Mercury 2 benchmark, API pricing, and provider data cover undefined API provider} other undefined API providers}}, with prices starting at $0.020/request. The page also shows measured API speed and first-token latency.
Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achie...
Category Performance
Observed-capability estimates with uncertainty reported separately from independent benchmark families.
- Coverage
- 2 / 8
- Methodology
- V3.0
#1Reasoning51.1Provisional1/4 Measured dimensions
#2Coding47Provisional1/4 Measured dimensions
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Rankings
Excels at
Detailed scores
Updated: Sep 13, 2026Speed & latency
undefined metric} other undefined metrics}}
Output speed1070.1 tok/s#1 / 80Time to first token4.29 s#59 / 80
Speed & latency
undefined metric} other undefined metrics}}
Pricing
undefined metric} other undefined metrics}}
Input price$0.250/M#44 / 186Output price$0.750/M#35 / 186
Pricing
undefined metric} other undefined metrics}}
Agents
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Agents
V3.0undefined metric} other undefined metrics}} · No data
Coding
V3.0undefined metric} other undefined metrics}} · Provisional
Score4780% interval31.0–63.01/4 Measured dimensionsSciCode37.7%#70 / 89
Coding
V3.0undefined metric} other undefined metrics}} · Provisional
Reasoning
V3.0undefined metric} other undefined metrics}} · Provisional
Score51.180% interval37.2–65.11/4 Measured dimensionsGPQA77.0%#114 / 218HLE17.1%#94 / 216
Reasoning
V3.0undefined metric} other undefined metrics}} · Provisional
Knowledge
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Knowledge
V3.0undefined metric} other undefined metrics}} · No data
Math
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Math
V3.0undefined metric} other undefined metrics}} · No data
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
Multimodal
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Multimodal
V3.0undefined metric} other undefined metrics}} · No data
Instruction following
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Instruction following
V3.0undefined metric} other undefined metrics}} · No data
OpenRouter endpoints
1 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
Inception inception | $0.250/M | $0.750/M | 100.0% | — | — | undefined tokens / undefined tokens |
Pricing Comparison
Mercury 2 API pricing starts at $0.020/request from 91VIP API.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 100% | mercury-2 | default | $0.020/request | - | — | — | — |
Alternatives & Similar Models
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
Kimi K2.5
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.
GLM-5
glm-5
Zhipu GLM-5 is Zhipu flagship GLM series model with enhanced reasoning, agent capabilities, and strong performance on Chinese enterprise and coding scenarios.
MiniMax M2.5
minimax-m2-5
MiniMax M2.5 is MiniMax's flagship text model for coding and agents, with SOTA-level programming and agentic performance, improved token efficiency, and fast high-TPS API deployment.
MiniMax M2.7
minimax-m2-7
MiniMax M2.7 is a high-tier M2-series model tuned for complex reasoning, long-context dialogue, and production-grade API workloads.
GPT-OSS
gpt-oss
GPT-OSS is an open-weight language model family designed for self-hosted inference, research, and cost-efficient alternatives to proprietary GPT-class models.
Frequently Asked Questions
- What benchmark data does Mercury 2 include?
- LMSpeed shows Mercury 2 benchmark context, API price, output speed, first-token latency, and provider data across 1 providers when those signals are available.
- What is the Mercury 2 API price?
- Mercury 2 has pricing from undefined provider} other undefined providers}}, ranging from $0.020/request to $0.020/request. 91VIP API has the lowest listed price.
- What does the Mercury 2 API pricing table include?
- The Mercury 2 API pricing table compares 1 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest Mercury 2 API pricing?
- 91VIP API currently has the lowest listed Mercury 2 price at $0.020/request across undefined provider} other undefined providers}}.
- Can I compare Mercury 2 API price and speed together?
- Yes. LMSpeed shows Mercury 2 API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
- Is Mercury 2 API free?
- Mercury 2 does not currently have a free API tier on LMSpeed. All 1 providers charge per token.
