GLM 5.3 Flash API 基准测试 价格和服务商数据
选择与 GLM 5.3 Flash 对比的模型
选择一个模型后会直接打开对应的对比页面。
页面同时展示实测速度和首字延迟。
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-contex...
分类性能
基于独立 benchmark family 的实测能力估计,不确定性单独展示。
- 覆盖
- 2 / 8
- 方法论
- V3.0
#1综合推理60.7暂定评分1/4 实测维度
#2代码56.5暂定评分1/4 实测维度
技术规格
Input and output token limits for this model, plus how it ranks on long-context understanding.
能力
Technical Details
- 输入
- 输出
- 发布日期
- Aug 2026
- Tokenizer
- Other
- 架构
- text+image+video->text
- 内容审核
- 否
- 到期日期
- 2098-12-31
- 支持参数
- frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
排名
擅长
详细分数
更新时间: 2026年8月27日速度与延迟
2 个指标
输出速度41.8 tok/s#68 / 76首字延迟1.16 s#39 / 76
速度与延迟
2 个指标
价格
2 个指标
输入价格$0.150/M#21 / 180输出价格$0.500/M#21 / 180
价格
2 个指标
智能体
V3.00 个指标 · 暂无数据
暂无数据80% 区间30.8–69.20/4 实测维度
智能体
V3.00 个指标 · 暂无数据
代码
V3.01 个指标 · 暂定评分
分数56.580% 区间40.5–72.41/4 实测维度SciCode46.1%#43 / 205
代码
V3.01 个指标 · 暂定评分
综合推理
V3.02 个指标 · 暂定评分
分数60.780% 区间46.8–74.61/4 实测维度GPQA91.2%#21 / 212HLE39.9%#27 / 209
综合推理
V3.02 个指标 · 暂定评分
知识
V3.00 个指标 · 暂无数据
暂无数据80% 区间30.8–69.20/4 实测维度
知识
V3.00 个指标 · 暂无数据
数学
V3.00 个指标 · 暂无数据
暂无数据80% 区间30.8–69.20/4 实测维度
数学
V3.00 个指标 · 暂无数据
多语言
V3.00 个指标 · 暂无数据
暂无数据80% 区间30.8–69.20/4 实测维度
多语言
V3.00 个指标 · 暂无数据
多模态
V3.00 个指标 · 暂无数据
暂无数据80% 区间30.8–69.20/4 实测维度
多模态
V3.00 个指标 · 暂无数据
指令遵循
V3.00 个指标 · 暂无数据
暂无数据80% 区间30.8–69.20/4 实测维度
指令遵循
V3.00 个指标 · 暂无数据
替代方案与相似模型
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek V4 Flash是DeepSeek V4 系列的语言模型,专为通用对话与文本生成任务优化。
DeepSeek V4 Pro
deepseek-v4-pro
DeepSeek V4 Pro是DeepSeek V4 系列的语言模型,专为通用对话与文本生成任务优化。
GLM-5.1
glm-5-1
Zhipu 的 GLM-5.1是语言模型,专为通用对话与文本生成任务优化。
GLM-5
glm-5
Zhipu 的 GLM-5是旗舰大语言模型,专为高复杂度通用任务与多步推理优化。
GLM-4.7
glm-4-7
Zhipu 的 GLM-4.7是旗舰大语言模型,专为高复杂度通用任务与多步推理优化。
GLM-5.2
glm-5-2
Zhipu GLM-5.2 is Zhipu latest flagship coding and agentic model with a 1M-token context window, enhanced reasoning modes, and long-horizon software engineering capabilities.
