数据点: 82
Model compare
DeepSeek V3.2 和 Kimi K2 Thinking 的结论先放在这里,方便先判断是否值得继续看明细。
综合加权结果:Kimi K2 Thinking。Benchmark 能力分类占 80%,价格、API 性能和可用性占 20%。
结论
Kimi K2 Thinking
Kimi K2 Thinking 当前综合加权更高;模型 A / B 得分为 16 对 84。
证据覆盖
82 个数据点
包含 8 个 benchmark、0 个 audit 样本和 7 个 provider 样本。
选择依据
优先看 Kimi K2 Thinking
下方图表把 15 组高信号样本拆开,便于核对速度、跑分和安全分。
左右两边都可以换成其他模型,页面会打开新的 LMSpeed 对比 URL。
选择其他模型后会打开新的对比页。
这份报告只使用 LMSpeed 已有数据:DeepSeek V3.2 和 Kimi K2 Thinking 的价格、测速聚合、第三方跑分与共同服务商样本。
| Model compare | DeepSeek V3.2 | Kimi K2 Thinking |
|---|
| 综合领先 | 对照 | 领先 |
|---|---|---|
| 综合加权得分 | 16.0 分 | 84.0 分 |
| Benchmark 分类领先 | 0 类 | 3 类 |
| 运营维度优势 | 最低输入价格、首 token 延迟、免费服务商、服务商覆盖 | 平均速度 |
| 上下文窗口 | 163.8K tokens | 262.1K tokens |
| 最大输出 | 65.5K tokens | 100.4K tokens |
| 模态 | 输入 文本 输出 文本 | 输入 文本 |
总体结果按 80% benchmark 能力分类和 20% 价格、API 速度/延迟与可用性加权;近期测试数量不参与胜负,缺失的 benchmark 分类不计分。
| Model compare | DeepSeek V3.2 | Kimi K2 Thinking |
|---|---|---|
| 开发者 | DeepSeek | MoonshotAI |
| 发布日期 | 2025年12月 | 2025年11月 |
| 参数量 | 暂无数据 | 暂无数据 |
| Tokenizer | DeepSeek | Other |
| 知识截止 | 暂无数据 | 暂无数据 |
| OpenRouter ID | deepseek/deepseek-v3.2 | moonshotai/kimi-k2-thinking |
| 来源链接 | 暂无数据 | 暂无数据 |
这份报告只使用 LMSpeed 已有数据:DeepSeek V3.2 和 Kimi K2 Thinking 的价格、测速聚合、第三方跑分与共同服务商样本。
DeepSeek V3.2
DeepSeek V3.2 的运营优势是:最低输入价格、首 token 延迟、免费服务商、服务商覆盖。
Kimi K2 Thinking
Kimi K2 Thinking 在 benchmark 分类(代码、推理、数学)和运营维度(平均速度)更强。
来自 LMSpeed 同步的第三方 benchmark profile;只展示两个模型都有数值的指标。
按 0-100 分对比 benchmark 分类表现;点击分类可以聚焦查看差距。
平均分
DeepSeek V3.2
46.3
平均分
Kimi K2 Thinking
55.7
DeepSeek V3.2
Kimi K2 Thinking 领先 6.2
Kimi K2 Thinking 领先 4.6
DeepSeek V3.2
Kimi K2 Thinking 领先 4.6
暂无数据
暂无数据
DeepSeek V3.2
按具体 benchmark 指标对比两个模型,展示来源、排名覆盖、置信度、误差和评测日期等上下文。
DeepSeek V3.2
领先$0.420/M
排名 #17/162 · confidence 4
Kimi K2 Thinking
$2.50/M
排名 #75/162 · confidence 4
DeepSeek V3.2
领先$0.315/M
排名 #26/162 · confidence 4
Kimi K2 Thinking
$1.07/M
排名 #78/162 · confidence 4
DeepSeek V3.2
领先$0.280/M
排名 #44/162 · confidence 4
Kimi K2 Thinking
$0.600/M
排名 #80/162 · confidence 4
DeepSeek V3.2
领先86.2%
排名 #15/125 · confidence 4
Kimi K2 Thinking
84.8%
排名 #28/125 · confidence 4
DeepSeek V3.2
22.2%
排名 #50/187 · confidence 4
Kimi K2 Thinking
领先22.3%
排名 #49/187 · confidence 4
DeepSeek V3.2
领先84.0%
排名 #53/188 · confidence 4
Kimi K2 Thinking
83.8%
排名 #55/188 · confidence 4
DeepSeek V3.2
领先86.2%
排名 #5/115 · confidence 4
Kimi K2 Thinking
85.3%
排名 #8/115 · confidence 4
DeepSeek V3.2
38.9%
排名 #82/185 · confidence 4
Kimi K2 Thinking
领先42.4%
排名 #49/185 · confidence 4
来自共同 provider 的最近完成 audit,展示四个安全/完整性分组分数和报告入口。
| Provider | DeepSeek V3.2 | Kimi K2 Thinking |
|---|---|---|
| 暂无共同 provider 的已完成 audit。 | ||
把同一 provider 的测速聚合和 input/output 价格放进同一行,便于判断实际 API 表现和迁移成本。
| Provider | DeepSeek V3.2 | Kimi K2 Thinking |
|---|---|---|
126 次测试 | DeepSeek V3.2 speed / latency 26 tok/s / 11125ms input / output 暂无数据 | Kimi K2 Thinking speed / latency N/A / N/A input / output 暂无数据 |
75 次测试 | DeepSeek V3.2 speed / latency 18 tok/s / 4488ms input / output 暂无数据 | Kimi K2 Thinking speed / latency 41 tok/s / 19451ms input / output 暂无数据 |
65 次测试 | DeepSeek V3.2 speed / latency 27 tok/s / 1295ms input / output 暂无数据 | Kimi K2 Thinking speed / latency 34 tok/s / 16242ms input / output 暂无数据 |
45 次测试 | DeepSeek V3.2 speed / latency 235 tok/s / 1787ms input / output 暂无数据 | Kimi K2 Thinking speed / latency N/A / N/A input / output 暂无数据 |
40 次测试 | DeepSeek V3.2 speed / latency 29 tok/s / 1828ms input / output 暂无数据 | Kimi K2 Thinking speed / latency N/A / N/A input / output 暂无数据 |
DeepSeek V3.2 deepseek-v3.2 speed / latency 暂无数据 input / output $0/M/$0/M | Kimi K2 Thinking kimi-k2-thinking speed / latency 暂无数据 input / output $0.027/M | |
DeepSeek V3.2 次DeepSeek-V3.2 speed / latency 暂无数据 input / output $0.0026/request | Kimi K2 Thinking kimi-k2-thinking speed / latency 暂无数据 input / output $0.274/M |
综合加权结果:Kimi K2 Thinking。Benchmark 能力分类占 80%,价格、API 性能和可用性占 20%。
输出
| 能力 | 文本输入文本输出工具调用结构化输出JSON 模式推理 | 文本输入文本输出工具调用结构化输出JSON 模式推理 |
|---|