选择一个模型后会直接打开对应的对比页面。
Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion tot...
基于独立 benchmark family 的实测能力估计,不确定性单独展示。
Input and output token limits for this model, plus how it ranks on long-context understanding.
2 个指标
2 个指标
0 个指标 · 暂无数据
1 个指标 · 暂定评分
2 个指标 · 暂定评分
0 个指标 · 暂无数据
0 个指标 · 暂无数据
0 个指标 · 暂无数据
0 个指标 · 暂无数据
0 个指标 · 暂无数据
来自 OpenRouter 的第三方端点数据,和 LMSpeed 自测数据分开展示。部分 30 分钟实时性能字段需要配置 OpenRouter API key 后同步才会出现。
| Provider endpoint | 输入 | 输出 | 1 天在线率 | 30m 延迟 | 30m 吞吐 | 上下文 / 输出 |
|---|---|---|---|---|---|---|
Alibaba alibaba | $2/M | $6/M | 100.0% | — | — | 1M tokens / 131.1K tokens |
DigitalOcean digitalocean | $2/M | $6/M | 99.7% | — | — | 262.1K tokens / 52.4K tokens |
SiliconFlow siliconflow/fp8 | $2/M | $6/M | 99.2% | — | — | 1.0M tokens / 131.1K tokens |
Together together | $2.50/M | $6.25/M | 98.2% | — | — | 1.0M tokens / — |
Modal modal/nvfp4 | $2/M | $6/M | 98.2% | — | — | 1M tokens / 262.1K tokens |
Venice venice | $2.50/M | $7.50/M | 95.5% | — | — | 262.1K tokens / 65.5K tokens |
DeepInfra deepinfra/fp4 | $2/M | $6/M | 94.5% | — | — | 262.1K tokens / 131.1K tokens |