Kimi K2 Thinking API Benchmarks, Pricing & Provider Data
Compare Kimi K2 Thinking with another model
Choose a model to open its comparison page.
Kimi K2 Thinking benchmark, API pricing, and provider data cover 57 API providers, with prices starting at $0.010/request. Kimi K2 Thinking free API options are available from 1 provider. The page also shows measured API speed and first-token latency.
Moonshot AI Kimi K2 Thinking is a reasoning model in the Kimi series, designed for complex reasoning, problem-solving, and analytical tasks.
Category Performance
Observed-capability estimates with uncertainty reported separately from independent benchmark families.
- Coverage
- 3 / 8
- Methodology
- V3.0
#1Coding57Provisional1/4 Measured dimensions
#2Reasoning55.5Estimated2/4 Measured dimensions
#3Math54Provisional1/4 Measured dimensions
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Technical Details
- Input
- Output
- Released
- Nov 2025
- Tokenizer
- Other
- Architecture
- text->text
- Moderated
- No
- Supported parameters
- frequency_penaltyinclude_reasoninglogprobsmax_tokenspresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
Rankings
Excels at
Detailed scores
Updated: Sep 1, 2026Pricing
undefined metric} other undefined metrics}}
Input price$0.600/M#94 / 181Output price$2.50/M#87 / 181
Pricing
undefined metric} other undefined metrics}}
Agents
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Agents
V3.0undefined metric} other undefined metrics}} · No data
Coding
V3.0undefined metric} other undefined metrics}} · Provisional
Score5780% interval43.0–70.91/4 Measured dimensionsLiveCodeBench85.3%#8 / 115SciCode42.4%#63 / 206
Coding
V3.0undefined metric} other undefined metrics}} · Provisional
Reasoning
V3.0undefined metric} other undefined metrics}} · Estimated
Score55.580% interval44.4–66.62/4 Measured dimensionsMMLU-Pro84.8%#31 / 129GPQA83.8%#72 / 213
Reasoning
V3.0undefined metric} other undefined metrics}} · Estimated
Knowledge
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Knowledge
V3.0undefined metric} other undefined metrics}} · No data
Math
V3.0undefined metric} other undefined metrics}} · Provisional
Score5480% interval37.8–70.11/4 Measured dimensions
Math
V3.0undefined metric} other undefined metrics}} · Provisional
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
Multimodal
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Multimodal
V3.0undefined metric} other undefined metrics}} · No data
Instruction following
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Instruction following
V3.0undefined metric} other undefined metrics}} · No data
OpenRouter endpoints
2 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
Google google-vertex | $0.600/M | $2.50/M | 100% | — | — | undefined tokens / undefined tokens |
Novita novita/bf16 | $0.600/M | $2.50/M | 99.6% | — | — | undefined tokens / undefined tokens |
Pricing Comparison
Compare Kimi K2 Thinking API pricing across 56 providers. Prices range from $0.010/request to $100.00/request. 素墨API offers the lowest rate at $0.010/request. 1 provider offers free API credits or a free tier.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 100% | kimi-k2-thinking | default | -36%$0.384/M | -36%$1.61/M | — | — | — | |
L1 100% | kimi-k2-thinking | default | -9%$0.548/M | -12%$2.19/M | — | — | — | |
L1 100% | kimi-k2-thinking-251104 | default | -9%$0.548/M | -12%$2.19/M | — | — | — | |
L1 100% | kimi-k2-thinking | deepseek | -63%$0.219/M | -65%$0.877/M | — | — | — | |
L1 99% | kimi-k2-thinking | other | -72%$0.167/M Cache read$0.167/M | -58%$1.04/M | — | — | — | |
L1 100% | kimi-k2-thinking | default | -9%$0.548/M Cache read$0.137/M | -12%$2.19/M | — | — | — | |
L1 100% | kimi-k2-thinking | default | $100.00/request | - | — | — | — | |
L1 99% | kimi-k2-thinking | kimi | -27%$0.438/M | -30%$1.75/M | — | — | — | |
L1 100% | kimi-k2-thinking | default | -95%$0.027/M Cache read$0.027/M | -93%$0.171/M | — | — | — | |
L1 100% | kimi-k2-thinking | default | -54%$0.274/M | $2.74/M | — | — | — | |
L1 99% | kimi-k2-thinking | default | -9%$0.548/M | -12%$2.19/M | — | — | — | |
L1 99% | kimi-k2-thinking | default | $1.96/M | $7.83/M | — | — | — | |
L1 99% | kimi-k2-thinking-251104 | default | $1.96/M | $7.83/M | — | — | — | |
L1 99% | Pro/moonshotai/Kimi-K2-Thinking | default | $1.96/M | $7.83/M | — | — | — | |
L1 100% | kimi-k2-thinking | default | $4.00/M | $16.00/M | — | — | — | |
L1 99% | kimi-k2-thinking | default | $3.20/M | $12.80/M | — | — | — | |
L1 100% | kimi-k2-thinking-tidy | default | $0.010/request | - | — | — | — | |
L1 100% | moonshotai/kimi-k2-thinking | default | $0.010/request | - | — | — | — | |
L1 100% | kimi-k2-thinking | default | $0.010/request | - | — | — | — | |
L1 99% | moonshotai/kimi-k2-thinking | default | $75.00/M | $75.00/M | — | — | — |
Alternatives & Similar Models
DeepSeek V3.2
deepseek-v3-2
DeepSeek V3.2 is an upgraded V3-series MoE model with stronger reasoning, coding, and math performance, widely available through OpenAI-compatible API relays.
Kimi K2.5
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.
MiniMax M2.5
minimax-m2-5
MiniMax M2.5 is MiniMax's flagship text model for coding and agents, with SOTA-level programming and agentic performance, improved token efficiency, and fast high-TPS API deployment.
GLM-5.1
glm-5-1
Zhipu GLM-5.1 is a next-generation GLM model aimed at frontier reasoning, coding, and bilingual agent applications.
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
DeepSeek V4 Pro
deepseek-v4-pro
DeepSeek V4 Pro is the professional-tier DeepSeek V4 model, targeting frontier reasoning, coding, and agent workflows with maximum capability.
Frequently Asked Questions
- What benchmark data does Kimi K2 Thinking include?
- LMSpeed shows Kimi K2 Thinking benchmark context, API price, output speed, first-token latency, and provider data across 57 providers when those signals are available.
- What is the Kimi K2 Thinking API price?
- Kimi K2 Thinking has pricing from 57 providers, ranging from $0.010/request to $100.00/request. 素墨API has the lowest listed price.
- What does the Kimi K2 Thinking API pricing table include?
- The Kimi K2 Thinking API pricing table compares 57 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest Kimi K2 Thinking API pricing?
- 素墨API currently has the lowest listed Kimi K2 Thinking price at $0.010/request across 57 providers.
- Can I compare Kimi K2 Thinking API price and speed together?
- Yes. LMSpeed shows Kimi K2 Thinking API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
- Is Kimi K2 Thinking API free?
- Yes, Kimi K2 Thinking free API options are available through 1 providerundefined other undefined} on LMSpeed, including Zero API. These providers offer free API credits or a free tier with no per-token charges.
- Where can I get Kimi K2 Thinking free API access?
- LMSpeed currently lists 1 free API providerundefined other undefined} for Kimi K2 Thinking: Zero API. Check each provider row before using it because free tier limits can change.
