Kimi K2.6 API Benchmarks, Pricing & Provider Data
Compare Kimi K2.6 with another model
Choose a model to open its comparison page.
Kimi K2.6 benchmark, API pricing, and provider data cover undefined API provider} other undefined API providers}}, with prices starting at $0.00005/request. Kimi K2.6 free API options are available from undefined provider} other undefined providers}}. The page also shows measured API speed and first-token latency.
Moonshot Kimi K2.6 is an open-source native multimodal agent model with 1 trillion MoE parameters, a 256K context window, and state-of-the-art coding, vision, and long-horizon agent swarm capabilities...
Category Performance
Observed-capability estimates with uncertainty reported separately from independent benchmark families.
- Coverage
- 7 / 8
- Methodology
- V3.0
#1Math58.4Estimated3/4 Measured dimensions
#2Reasoning55.3Estimated2/4 Measured dimensions
#3Instruction following54.8Provisional1/4 Measured dimensions
#4Coding54.5RatedGlobal rank #174/4 Measured dimensions
#5Agents53.6RatedGlobal rank #274/4 Measured dimensions
#6Multimodal53.4Estimated3/4 Measured dimensions
#7Knowledge49.3Provisional1/4 Measured dimensions
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Technical Details
- Input
- Output
- Released
- Apr 2026
- Tokenizer
- Other
- Architecture
- text+image->text
- Moderated
- No
- Supported parameters
- frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_pparallel_tool_callspresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
Rankings
Excels at
It's decent at
Falls behind in
Detailed scores
Updated: Sep 13, 2026Overall
undefined metric} other undefined metrics}}
Overall score60.0#48 / 112
Overall
undefined metric} other undefined metrics}}
Pricing
undefined metric} other undefined metrics}}
Input price$0.950/M#111 / 186Output price$4.00/M#112 / 186
Pricing
undefined metric} other undefined metrics}}
Agents
V3.0undefined metric} other undefined metrics}} · Rated
Score53.6#2780% interval48.3–58.84/4 Measured dimensionsAgentic score61.0#41 / 77Terminal-Bench 2.066.7#21 / 53
Agents
V3.0undefined metric} other undefined metrics}} · Rated
Coding
V3.0undefined metric} other undefined metrics}} · Rated
Score54.5#1780% interval48.4–60.54/4 Measured dimensionsSciCode51.5%#28 / 89Coding score58.2#29 / 87
Coding
V3.0undefined metric} other undefined metrics}} · Rated
Reasoning
V3.0undefined metric} other undefined metrics}} · Estimated
Score55.380% interval44.5–66.02/4 Measured dimensionsGPQA91.1%#25 / 218HLE37.5%#36 / 216
Reasoning
V3.0undefined metric} other undefined metrics}} · Estimated
Knowledge
V3.0undefined metric} other undefined metrics}} · Provisional
Score49.380% interval35.3–63.31/4 Measured dimensionsKnowledge score56.6#63 / 83GPQA-D90.5#15 / 32
Knowledge
V3.0undefined metric} other undefined metrics}} · Provisional
Math
V3.0undefined metric} other undefined metrics}} · Estimated
Score58.480% interval49.4–67.53/4 Measured dimensionsAIME2696.4#3 / 14HMMT Feb 202692.7#6 / 18
Math
V3.0undefined metric} other undefined metrics}} · Estimated
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
Multimodal
V3.0undefined metric} other undefined metrics}} · Estimated
Score53.480% interval44.4–62.43/4 Measured dimensionsMultimodal Grounded score64.0#35 / 56MMMU-Pro79.4#12 / 31
Multimodal
V3.0undefined metric} other undefined metrics}} · Estimated
Instruction following
V3.0undefined metric} other undefined metrics}} · Provisional
Score54.880% interval38.8–70.81/4 Measured dimensionsAA-IFBench76.0#12 / 84
Instruction following
V3.0undefined metric} other undefined metrics}} · Provisional
OpenRouter endpoints
20 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
DigitalOcean digitalocean | $0.950/M | $4/M | 100% | — | — | undefined tokens / undefined tokens |
SiliconFlow siliconflow/fp8 | $0.770/M | $3.40/M | 100.0% | — | — | undefined tokens / undefined tokens |
Cloudflare cloudflare | $0.950/M | $4/M | 99.9% | — | — | undefined tokens / undefined tokens |
Decart decart/fp4 | $0.587/M | $2.47/M | 99.8% | — | — | undefined tokens / undefined tokens |
Baidu baidu/fp4 | $0.580/M | $2.44/M | 99.7% | — | — | undefined tokens / undefined tokens |
CoreWeave coreweave/fp4 | $0.650/M | $3.41/M | 99.6% | — | — | undefined tokens / undefined tokens |
Moonshot AI moonshotai/int4 | $0.950/M | $4/M | 99.6% | — | — | undefined tokens / undefined tokens |
Parasail parasail/int4 | $0.750/M | $3.50/M | 99.6% | — | — | undefined tokens / undefined tokens |
Inceptron inceptron/int4 | $0.590/M | $2.45/M | 99.4% | — | — | undefined tokens / undefined tokens |
DeepInfra deepinfra/fp4 | $0.750/M | $3.50/M | 99.1% | — | — | undefined tokens / undefined tokens |
Novita novita | $0.800/M | $3.40/M | 98.6% | — | — | undefined tokens / undefined tokens |
AtlasCloud atlas-cloud/int4 | $0.950/M | $4/M | 97.5% | — | — | undefined tokens / undefined tokens |
StreamLake streamlake/fp8 | $0.598/M | $2.52/M | 97.1% | — | — | undefined tokens / undefined tokens |
Phala phala | $1.09/M | $4.60/M | 95.5% | — | — | undefined tokens / undefined tokens |
Venice venice/int4 | $0.750/M | $3.50/M | 94.4% | — | — | undefined tokens / undefined tokens |
Crusoe crusoe/bf16 | $0.700/M | $3.50/M | 93.1% | — | — | undefined tokens / undefined tokens |
Chutes chutes/int4 | $0.580/M | $3.40/M | 92.5% | — | — | undefined tokens / undefined tokens |
GMICloud gmicloud/fp8 | $0.855/M | $3.60/M | 90.7% | — | — | undefined tokens / undefined tokens |
BaseTen baseten/fp4 | $0.950/M | $4/M | 51.8% | — | — | undefined tokens / undefined tokens |
Fireworks fireworks | $0.950/M | $4/M | 0% | — | — | undefined tokens / undefined tokens |
Pricing Comparison
Compare Kimi K2.6 API pricing across 224 providers. Prices range from $0.00005/request to $100.00/request. CM-API 公益站 offers the lowest rate at $0.00005/request. 9 providers offer free API credits or a free tier.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 99% L2 9% | moonshot/kimi-k2.6 | default | -20%$0.760/M Cache read$0.128/M | -20%$3.20/M | 119.6 t/s | 19.53 s | ||
S3AI API Free | L1 94% L2 97% | Kimi-K2.6-free | free | Free | Free | 115.8 t/s | 18.26 s | — |
L1 94% L2 97% | Kimi-K2.6 | default | -29%$0.670/M Cache read$0.113/M | -30%$2.78/M | — | — | — | |
L1 100% | kimi-k2.6 | default | -41%$0.557/M Cache read$0.094/M | -42%$2.31/M | 75.1 t/s | 14.12 s | — | |
L1 0% | kimi-k2.6 | 2api | $0.950/M Cache read$0.160/M | $4.00/M | 56.6 t/s | 9.32 s | ||
L1 0% | moonshotai/kimi-k2.6 | 懒人 | $0.950/M Cache read$0.200/M | $4.00/M | — | — | — | |
L1 100% | kimi-k2.6 | sale | $3.25/M | $13.50/M | — | — | — | |
L1 100% | kimi-k2.6 | default | -53%$0.445/M | -54%$1.85/M | — | — | — | |
L1 100% | kimi-k2.6 | kimi-officially | -7%$0.882/M Cache read$0.149/M | -8%$3.66/M | — | — | — | |
L1 100% | kimi-k2.6 | default | -6%$0.890/M Cache read$0.151/M | -8%$3.70/M | — | — | — | |
L1 100% | kimi-k2.6 | default | -53%$0.445/M | -54%$1.85/M | — | — | — | |
L1 100% | KIMI-K2.6 | default | -6%$0.890/M | -7%$3.70/M | — | — | — | |
L1 100% | kimi-k2.6 | default | -6%$0.890/M Cache read$0.150/M | -7%$3.70/M | — | — | — | |
L1 100% | Kimi-K2.6 | default | -6%$0.890/M | -7%$3.70/M | — | — | — | |
L1 100% | Pro/moonshotai/Kimi-K2.6 | default | $10.27/M | $10.27/M | — | — | — | |
L1 100% | kimi-k2.6 | default | -25%$0.712/M | -26%$2.96/M | — | — | — | |
L1 100% | kimi-k2.6 | default | -93%$0.065/M Cache read$0.011/M | -93%$0.274/M | — | — | — | |
L1 100% | kimi-k2.6 | Model-vip | -72%$0.267/M Cache read$0.045/M | -72%$1.11/M | — | — | — | |
L1 100% | kimi-k2.6 | default | -64%$0.342/M | -49%$2.05/M | — | — | — | |
L1 100% | kimi-k2.6 | deepseek | -63%$0.356/M Cache read$0.061/M | -63%$1.48/M | — | — | — |
Alternatives & Similar Models
DeepSeek V4 Pro
deepseek-v4-pro
DeepSeek V4 Pro is the professional-tier DeepSeek V4 model, targeting frontier reasoning, coding, and agent workflows with maximum capability.
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
GLM-5.1
glm-5-1
Zhipu GLM-5.1 is a next-generation GLM model aimed at frontier reasoning, coding, and bilingual agent applications.
Kimi K2.5
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.
GPT-5.4
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
GLM-5.2
glm-5-2
Zhipu GLM-5.2 is Zhipu latest flagship coding and agentic model with a 1M-token context window, enhanced reasoning modes, and long-horizon software engineering capabilities.
Frequently Asked Questions
- What benchmark data does Kimi K2.6 include?
- LMSpeed shows Kimi K2.6 benchmark context, API price, output speed, first-token latency, and provider data across 233 providers when those signals are available.
- What is the Kimi K2.6 API price?
- Kimi K2.6 has pricing from undefined provider} other undefined providers}}, ranging from $0.00005/request to $100.00/request. CM-API 公益站 has the lowest listed price.
- What does the Kimi K2.6 API pricing table include?
- The Kimi K2.6 API pricing table compares 233 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest Kimi K2.6 API pricing?
- CM-API 公益站 currently has the lowest listed Kimi K2.6 price at $0.00005/request across undefined provider} other undefined providers}}.
- Can I compare Kimi K2.6 API price and speed together?
- Yes. LMSpeed shows Kimi K2.6 API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
- Is Kimi K2.6 API free?
- Yes, Kimi K2.6 free API options are available through 9 providerundefined other undefined} on LMSpeed, including Zero API, 兔子API, DeadlySignal API, S3AI API, 兔子API. These providers offer free API credits or a free tier with no per-token charges.
- Where can I get Kimi K2.6 free API access?
- LMSpeed currently lists 9 free API providerundefined other undefined} for Kimi K2.6: Zero API, 兔子API, DeadlySignal API, S3AI API, 兔子API. Check each provider row before using it because free tier limits can change.
