Mercury 2 API Benchmarks, Pricing & Provider Data
Compare Mercury 2 with another model
Choose a model to open its comparison page.
Mercury 2 is listed across 1 API provider. Mercury 2 free API options are available from 1 provider. The page also shows measured API speed and first-token latency.
Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achie...
Category Performance
Observed-capability estimates with uncertainty reported separately from independent benchmark families.
- Coverage
- 2 / 8
- Methodology
- V3.0
#1Coding51.5Provisional1/4 Measured dimensions
#1Reasoning51.5Provisional1/4 Measured dimensions
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Rankings
Excels at
Detailed scores
Updated: Sep 1, 2026Speed & latency
undefined metric} other undefined metrics}}
Output speed838.3 tok/s#1 / 77Time to first token3.50 s#62 / 77
Speed & latency
undefined metric} other undefined metrics}}
Pricing
undefined metric} other undefined metrics}}
Input price$0.250/M#45 / 181Output price$0.750/M#34 / 181
Pricing
undefined metric} other undefined metrics}}
Agents
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Agents
V3.0undefined metric} other undefined metrics}} · No data
Coding
V3.0undefined metric} other undefined metrics}} · Provisional
Score51.580% interval35.6–67.51/4 Measured dimensionsSciCode38.7%#100 / 206
Coding
V3.0undefined metric} other undefined metrics}} · Provisional
Reasoning
V3.0undefined metric} other undefined metrics}} · Provisional
Score51.580% interval37.6–65.41/4 Measured dimensionsGPQA77.0%#109 / 213HLE17.1%#88 / 210
Reasoning
V3.0undefined metric} other undefined metrics}} · Provisional
Knowledge
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Knowledge
V3.0undefined metric} other undefined metrics}} · No data
Math
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Math
V3.0undefined metric} other undefined metrics}} · No data
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
Multimodal
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Multimodal
V3.0undefined metric} other undefined metrics}} · No data
Instruction following
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Instruction following
V3.0undefined metric} other undefined metrics}} · No data
OpenRouter endpoints
1 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
Inception inception | $0.250/M | $0.750/M | 100.0% | — | — | undefined tokens / undefined tokens |
Pricing Comparison
Mercury 2 is free to use through 1 provider with no per-token charges.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
91VIP API Free | L1 98% | mercury-2 | default | Free | Free | — | — | — |
Alternatives & Similar Models
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
Kimi K2.5
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.
GLM-5
glm-5
Zhipu GLM-5 is Zhipu flagship GLM series model with enhanced reasoning, agent capabilities, and strong performance on Chinese enterprise and coding scenarios.
MiniMax M2.5
minimax-m2-5
MiniMax M2.5 is MiniMax's flagship text model for coding and agents, with SOTA-level programming and agentic performance, improved token efficiency, and fast high-TPS API deployment.
MiniMax M2.7
minimax-m2-7
MiniMax M2.7 is a high-tier M2-series model tuned for complex reasoning, long-context dialogue, and production-grade API workloads.
GPT-OSS
gpt-oss
GPT-OSS is an open-weight language model family designed for self-hosted inference, research, and cost-efficient alternatives to proprietary GPT-class models.
Frequently Asked Questions
- Can I compare Mercury 2 API price and speed together?
- Yes. LMSpeed shows Mercury 2 API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
- Is Mercury 2 API free?
- Yes, Mercury 2 free API options are available through 1 providerundefined other undefined} on LMSpeed, including 91VIP API. These providers offer free API credits or a free tier with no per-token charges.
- Where can I get Mercury 2 free API access?
- LMSpeed currently lists 1 free API providerundefined other undefined} for Mercury 2: 91VIP API. Check each provider row before using it because free tier limits can change.
