Z.ai
·Released on Feb 11, 2026

GLM-5 API Benchmarks, Pricing & Provider Data

Compare GLM-5 with another model

Choose a model to open its comparison page.

Share on X
LLM

GLM-5 benchmark, API pricing, and provider data cover undefined API provider} other undefined API providers}}, with prices starting at $0.010/request. GLM-5 free API options are available from undefined provider} other undefined providers}}. The page also shows measured API speed and first-token latency.

Zhipu GLM-5 is Zhipu flagship GLM series model with enhanced reasoning, agent capabilities, and strong performance on Chinese enterprise and coding scenarios.

Quality
#72of 112
54.0
LMSpeed score
Speed
51char/s
21.59 s
Cost
#113of 186
$0.010/ 1M · 8:1 in:out
$0.0016 in · $0.0084 out

Category Performance

Observed-capability estimates with uncertainty reported separately from independent benchmark families.

Coverage
7 / 8
Methodology
V3.0
Category PerformanceObserved-capability estimates with uncertainty reported separately from independent benchmark families.Agents41.4Coding51.7Reasoning48.8Knowledge43.2Math51Multilingual43.8Multimodal-Instruction following52.4
#1Instruction following52.4Estimated2/4 Measured dimensions
80% interval: 40.564.4
constraint following51.5
ifeval · ifeval
structured outputPrior only
novel instruction generalization53.4
ifbench · aa_if_bench
long multiturn instructionPrior only
#2Coding51.7RatedGlobal rank #223/4 Measured dimensions
80% interval: 43.559.8
code generation53.4
scicode · aa_sci_code
repository engineering46.6
react_native_bench · react_native_evals / swe_pro · swe_pro / vibecode · vibe_code_bench
debugging testing55
swe_multilingual · benchlm_coding_swe_multilingual / swe_rebench · swe_rebench / swe_verified · swe_verified
tooling qualityPrior only
#3Math51Estimated3/4 Measured dimensions
80% interval: 41.960.0
foundational mathPrior only
competition math57.9
aime · aime2026 / hmmt · hmmt_feb2025 / hmmt · hmmt_feb2026 / hmmt · hmmt_nov2025
proof frontier47.9
frontiermath · frontier_math_v2_tier4 / frontiermath · frontier_math_v2_tiers13
applied tool math47.2
mmanswer · mm_answer_bench
#4Reasoning48.8RatedGlobal rank #453/4 Measured dimensions
80% interval: 40.757.0
abstract logicPrior only
scientific causal47
critpt · critpt / gpqa · gpqa / hle · hle
multistep constraints56.9
mmlu_pro · mmlu_pro
evidence verification42.6
ai_needle · ai_needle / lcr · lcr / longbench · long_bench_v2
#5Multilingual43.8Estimated2/4 Measured dimensions
80% interval: 31.656.0
cross language understanding39
nova · nova63
multilingual generationPrior only
reasoning transfer48.5
mmlu_prox · mmlu_pro_x
low resource robustnessPrior only
#6Knowledge43.2Provisional1/4 Measured dimensions
80% interval: 26.959.5
broad knowledgePrior only
professional knowledge43.2
supergpqa · super_gpqa
factualityPrior only
retrieval open bookPrior only
#7Agents41.4RatedGlobal rank #584/4 Measured dimensions
80% interval: 35.747.0
planning51
deep_planning · gert_labs
tool use34.1
mcp_atlas · mcp_atlas / mcp_tasks · mcp_tasks / tau · tau2_bench / tau · tau3_bench / toolathlon · toolathlon
environment execution42.4
terminalbench · benchlm_agentic_terminal_bench2 / wide_research · wide_research
recovery reliability37.9
apex_agents · apex_agents_aa / claw_eval · claw_eval / cyber_gym · cyber_gym
No data:Multimodal

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
204.8Ktokens
245.8 pages of text
OUTPUT
128Ktokens
8K128K1M4M
204.8K

Features

Technical Details

Input
Output
Released
Feb 2026
Tokenizer
Other
Architecture
text->text
Moderated
No
Supported parameters
frequency_penaltyinclude_reasoninglogprobsmax_tokenspresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p

Rankings

Excels at

It's decent at

Falls behind in

Detailed scores

Updated: Sep 13, 2026

Overall

undefined metric} other undefined metrics}}

Overall score54.0#72 / 112
Overall score
LMSpeed rank#72
Score54.0
UpdatedSep 13, 2026
Confidence4

Pricing

undefined metric} other undefined metrics}}

Input price$1.00/M#113 / 186Output price$3.20/M#105 / 186
Input price
LMSpeed rank#113
Score$1.00/M
UpdatedSep 13, 2026
Confidence4
Output price
LMSpeed rank#105
Score$3.20/M
UpdatedSep 13, 2026
Confidence4

Agents

V3.0

undefined metric} other undefined metrics}} · Rated

Score41.4#5880% interval35.7–47.04/4 Measured dimensionsAgentic score42.8#62 / 77Terminal-Bench 2.056.2#39 / 53
Dimensions and evidence
Planning & decomposition51
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
deep_planning · gert_labs · z 0.15 · q 1.00
Tool use34.1
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
mcp_atlas · mcp_atlas · z -3.00 · q 1.00
mcp_tasks · mcp_tasks · z -1.09 · q 0.63
tau · tau2_bench · z 1.22 · q 1.00
tau · tau3_bench · z -1.74 · q 1.00
toolathlon · toolathlon · z -1.03 · q 1.00
Environment & long-horizon execution42.4
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
terminalbench · benchlm_agentic_terminal_bench2 · z -0.47 · q 1.00
wide_research · wide_research · z -0.77 · q 1.00
Recovery & completion reliability37.9
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
apex_agents · apex_agents_aa · z -0.84 · q 1.00
claw_eval · claw_eval · z -0.31 · q 1.00
cyber_gym · cyber_gym · z -2.78 · q 1.00
Agentic score
LMSpeed rank#62
Score42.8
UpdatedSep 13, 2026
Confidence4
Terminal-Bench 2.0V3 evidence
LMSpeed rank#39
Score56.2
UpdatedSep 13, 2026
Confidence4
Claw-EvalV3 evidence
LMSpeed rank#19
Score57.7
UpdatedSep 13, 2026
Confidence4
QwenClawBench
LMSpeed rank#6
Score54.1
UpdatedSep 13, 2026
Confidence4
Τ³-bench resultsV3 evidence
LMSpeed rank#9
Score65.6
UpdatedSep 13, 2026
Confidence4
DeepPlanningV3 evidence
LMSpeed rank#6
Score14.6
UpdatedSep 13, 2026
Confidence4
ToolathlonV3 evidence
LMSpeed rank#20
Score38.0
UpdatedSep 13, 2026
Confidence4
MCP AtlasV3 evidence
LMSpeed rank#30
Score31.1
UpdatedSep 13, 2026
Confidence4
MCP-TasksV3 evidence
LMSpeed rank#4
Score60.8
UpdatedSep 13, 2026
Confidence4
WideResearchV3 evidence
LMSpeed rank#8
Score69.8
UpdatedSep 13, 2026
Confidence4
Τ²-bench resultsV3 evidence
LMSpeed rank#7
Score98.2
UpdatedSep 13, 2026
Confidence4
CyberGymV3 evidence
LMSpeed rank#15
Score43.2
UpdatedSep 13, 2026
Confidence4
APEX-Agents-AAV3 evidence
LMSpeed rank#18
Score14.5
UpdatedSep 13, 2026
Confidence4
Gert LabsV3 evidence
LMSpeed rank#21
Score51.0
UpdatedSep 13, 2026
Confidence4

Coding

V3.0

undefined metric} other undefined metrics}} · Rated

Score51.7#2280% interval43.5–59.83/4 Measured dimensionsCoding score49.5#57 / 87SWE-bench Verified77.8#21 / 49
Dimensions and evidence
Code generation53.4
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
scicode · aa_sci_code · z 0.29 · q 1.00
Repository engineering46.6
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
react_native_bench · react_native_evals · z -0.36 · q 1.00
swe_pro · swe_pro · z -0.35 · q 1.00
vibecode · vibe_code_bench · z -0.07 · q 1.00
Debugging & testing55
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
swe_multilingual · benchlm_coding_swe_multilingual · z -0.58 · q 1.00
swe_rebench · swe_rebench · z 1.53 · q 1.00
swe_verified · swe_verified · z 0.20 · q 1.00
Tool-assisted development & qualityPrior only
Coding score
LMSpeed rank#57
Score49.5
UpdatedSep 13, 2026
Confidence4
SWE-bench VerifiedV3 evidence
LMSpeed rank#21
Score77.8
UpdatedSep 13, 2026
Confidence4
SWE-bench Verified*
LMSpeed rank#3
Score72.8
UpdatedSep 13, 2026
Confidence4
SWE-bench ProV3 evidence
LMSpeed rank#35
Score55.1
UpdatedSep 13, 2026
Confidence4
SWE MultilingualV3 evidence
LMSpeed rank#15
Score73.3
UpdatedSep 13, 2026
Confidence4
SWE-RebenchV3 evidence
LMSpeed rank#2
Score62.8
UpdatedSep 13, 2026
Confidence4
React Native EvalsV3 evidence
LMSpeed rank#8
Score74.8
UpdatedSep 13, 2026
Confidence4
AA-SciCodeV3 evidence
LMSpeed rank#52
Score46.2
UpdatedSep 8, 2026
Confidence4
Vibe Code BenchV3 evidence
LMSpeed rank#19
Score23.4
UpdatedAug 20, 2026
Confidence1

Reasoning

V3.0

undefined metric} other undefined metrics}} · Rated

Score48.8#4580% interval40.7–57.03/4 Measured dimensionsMMLU-Pro85.7%#26 / 129GPQA82.0%#90 / 218
Dimensions and evidence
Abstract logicPrior only
Scientific & causal reasoning47
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
critpt · critpt · z -0.17 · q 1.00
gpqa · gpqa · z -0.54 · q 1.00
hle · hle · z -0.48 · q 1.00
Multi-step constraints56.9
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
mmlu_pro · mmlu_pro · z 0.72 · q 1.00
Evidence integration & verification42.6
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
ai_needle · ai_needle · z -2.25 · q 0.50
lcr · lcr · z 0.14 · q 1.00
longbench · long_bench_v2 · z -0.10 · q 1.00
MMLU-ProV3 evidence
LMSpeed rank#26
Score85.7%
UpdatedSep 13, 2026
Confidence4
GPQAV3 evidence
LMSpeed rank#90
Score82.0%
UpdatedSep 13, 2026
Confidence4
HLEV3 evidence
LMSpeed rank#59
Score29.3%
UpdatedSep 13, 2026
Confidence4
Reasoning score
LMSpeed rank#56
Score52.1
UpdatedSep 13, 2026
Confidence4
LongBench v2V3 evidence
LMSpeed rank#6
Score60.8
UpdatedSep 13, 2026
Confidence4
AI-NeedleV3 evidence
LMSpeed rank#4
Score63.3
UpdatedSep 13, 2026
Confidence4
AA-LCRV3 evidence
LMSpeed rank#48
Score75.7
UpdatedSep 13, 2026
Confidence4
CritPtV3 evidence
LMSpeed rank#61
Score2.0
UpdatedSep 13, 2026
Confidence4

Knowledge

V3.0

undefined metric} other undefined metrics}} · Provisional

Score43.280% interval26.9–59.51/4 Measured dimensionsKnowledge score72.8#32 / 83GPQA-D86.0#30 / 32
Dimensions and evidence
Broad knowledgePrior only
Professional knowledge43.2
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
supergpqa · super_gpqa · z -0.64 · q 1.00
FactualityPrior only
Retrieval & open-book usePrior only
Knowledge score
LMSpeed rank#32
Score72.8
UpdatedSep 13, 2026
Confidence4
GPQA-D
LMSpeed rank#30
Score86.0
UpdatedSep 13, 2026
Confidence4
SuperGPQAV3 evidence
LMSpeed rank#11
Score66.8
UpdatedSep 13, 2026
Confidence4
MMLU-Pro (Arcee)
LMSpeed rank#3
Score85.8
UpdatedSep 13, 2026
Confidence4
Artificial Analysis Intelligence Index
LMSpeed rank#54
Score27.9
UpdatedSep 13, 2026
Confidence4
AA-GPQA Diamond
LMSpeed rank#72
Score82.0
UpdatedSep 13, 2026
Confidence4
AA-HLE
LMSpeed rank#51
Score29.3
UpdatedSep 13, 2026
Confidence4
AA-Omniscience Index
LMSpeed rank#44
Score0.3
UpdatedSep 13, 2026
Confidence4
AA-Omniscience Accuracy
LMSpeed rank#66
Score26.3
UpdatedSep 13, 2026
Confidence4
AA-Omniscience Hallucination Rate
LMSpeed rank#91
Score35.3
UpdatedSep 13, 2026
Confidence4

Math

V3.0

undefined metric} other undefined metrics}} · Estimated

Score5180% interval41.9–60.03/4 Measured dimensionsMath score56.9#32 / 63AIME2695.8#4 / 14
Dimensions and evidence
Foundational mathPrior only
Competition math57.9
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
aime · aime2026 · z 0.48 · q 1.00
hmmt · hmmt_feb2025 · z 1.83 · q 0.88
hmmt · hmmt_feb2026 · z -0.15 · q 1.00
hmmt · hmmt_nov2025 · z 2.19 · q 1.00
Advanced proofs47.9
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
frontiermath · frontier_math_v2_tier4 · z -0.60 · q 1.00
frontiermath · frontier_math_v2_tiers13 · z -0.27 · q 1.00
Applied & tool-assisted math47.2
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
mmanswer · mm_answer_bench · z -0.32 · q 1.00
Math score
LMSpeed rank#32
Score56.9
UpdatedSep 13, 2026
Confidence4
AIME26V3 evidence
LMSpeed rank#4
Score95.8
UpdatedSep 13, 2026
Confidence4
AIME25 (Arcee)
LMSpeed rank#4
Score93.3
UpdatedSep 13, 2026
Confidence4
HMMT Feb 2025V3 evidence
LMSpeed rank#1
Score97.5
UpdatedSep 13, 2026
Confidence4
HMMT Nov 2025V3 evidence
LMSpeed rank#1
Score96.9
UpdatedSep 13, 2026
Confidence4
HMMT Feb 2026V3 evidence
LMSpeed rank#14
Score86.4
UpdatedSep 13, 2026
Confidence4
MMAnswerBenchV3 evidence
LMSpeed rank#6
Score82.5
UpdatedSep 13, 2026
Confidence4
FrontierMath v2 (Tiers 1-3)V3 evidence
LMSpeed rank#31
Score16.4
UpdatedSep 13, 2026
Confidence4
FrontierMath v2 (Tier 4)V3 evidence
LMSpeed rank#31
Score2.1
UpdatedSep 13, 2026
Confidence4

Multilingual

V3.0

undefined metric} other undefined metrics}} · Estimated

Score43.880% interval31.6–56.02/4 Measured dimensionsMultilingual score48.7#6 / 11MMLU-ProX83.1#6 / 11
Dimensions and evidence
Cross-language understanding39
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
nova · nova63 · z -1.47 · q 0.88
Multilingual generationPrior only
Reasoning transfer48.5
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
mmlu_prox · mmlu_pro_x · z 0.00 · q 1.00
Low-resource robustnessPrior only
Multilingual score
LMSpeed rank#6
Score48.7
UpdatedSep 13, 2026
Confidence4
MMLU-ProXV3 evidence
LMSpeed rank#6
Score83.1
UpdatedSep 13, 2026
Confidence4
NOVA-63V3 evidence
LMSpeed rank#7
Score55.1
UpdatedSep 13, 2026
Confidence4

Multimodal

V3.0

undefined metric} other undefined metrics}} · No data

80% interval30.8–69.20/4 Measured dimensionsDesign Arena Website1260.0#34 / 78
Dimensions and evidence
Perception & OCRPrior only
Document & spatial understandingPrior only
Visual reasoningPrior only
Video & grounded actionPrior only
Design Arena Website
LMSpeed rank#34
Score1260.0
UpdatedSep 13, 2026
Confidence4

Instruction following

V3.0

undefined metric} other undefined metrics}} · Estimated

Score52.480% interval40.5–64.42/4 Measured dimensionsInstruction Following score88.5#23 / 52IFEval92.6#8 / 16
Dimensions and evidence
Constraint following51.5
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
ifeval · ifeval · z 0.00 · q 1.00
Structured outputPrior only
Novel-instruction generalization53.4
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
ifbench · aa_if_bench · z 0.16 · q 1.00
Multi-turn & long instructionsPrior only
Instruction Following score
LMSpeed rank#23
Score88.5
UpdatedSep 13, 2026
Confidence4
IFEvalV3 evidence
LMSpeed rank#8
Score92.6
UpdatedSep 13, 2026
Confidence4
AA-IFBenchV3 evidence
LMSpeed rank#29
Score72.3
UpdatedSep 13, 2026
Confidence4

OpenRouter endpoints

8 endpoints

Third-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.

Provider endpointInputOutput1d uptime30m latency30m throughputContext / output
Novita
novita/fp8
$1/M$3.20/M100%undefined tokens / undefined tokens
Z.AI
z-ai/fp8
$1/M$3.20/M100.0%undefined tokens / undefined tokens
SiliconFlow
siliconflow/fp8
$0.950/M$2.55/M99.9%undefined tokens / undefined tokens
StreamLake
streamlake/fp8
$0.600/M$1.92/M99.7%undefined tokens / undefined tokens
GMICloud
gmicloud/fp8
$0.600/M$1.92/M99.6%undefined tokens / undefined tokens
Baidu
baidu/fp8
$0.700/M$2.24/M99.5%undefined tokens / undefined tokens
Venice
venice/fp8
$1/M$3.20/M98.9%undefined tokens / undefined tokens
Amazon Bedrock
amazon-bedrock
$1/M$3.20/M97.4%undefined tokens / undefined tokens

Pricing Comparison

Compare GLM-5 API pricing across 150 providers. Prices range from $0.010/request to $150.00/M. 素墨API offers the lowest rate at $0.010/request. 3 providers offer free API credits or a free tier.

ProviderHealthModel VariantGroupInput ($/M)Output ($/M)Speed (t/s)First tokenAudit
L1
100%
L2
8%
z-ai/glm-5
default
-20%$0.800/M
Cache read$0.160/M
-20%$2.56/M
139.0 t/s
13.57 s
L1
100%
L2
0%
z-ai/glm-5
default
-11%$0.890/M
Cache read$0.223/M
$3.26/M
42.9 t/s
24.10 s
L1
100%
glm-5
限时特价
$75.00/M
$75.00/M
36.1 t/s
21.94 s
L1
100%
glm-5
sale
$2.00/M
$9.00/M
L1
100%
glm-5
default
-73%$0.274/M
-61%$1.23/M
L1
100%
glm-5
bailian
-71%$0.286/M
Cache read$0.057/M
-60%$1.29/M
L1
100%
glm-5
default
-73%$0.274/M
-61%$1.23/M
L1
100%
glm-5
default
-45%$0.548/M
Cache read$0.137/M
-23%$2.47/M
L1
100%
glm-5
default
-18%$0.822/M
$3.29/M
L1
100%
glm-5
default
-23%$0.767/M
-62%$1.23/M
L1
100%
glm-5
default
-93%$0.068/M
Cache read$0.014/M
-93%$0.219/M
L1
100%
glm-5
Self-Deployed-2
-85%$0.148/M
-79%$0.667/M
L1
100%
glm-5
default
-73%$0.274/M
-14%$2.74/M
L1
100%
glm-5
国产模型
-11%$0.891/M
Cache read$0.891/MCache write$0.178/MCache write 1h$0.285/M
-11%$2.85/M
L1
100%
glm-5
default
$4.00/M
Cache read$0.800/M
$18.00/M
L1
100%
GLM-5
Zai-officially
$9.25/M
$9.25/M
L1
99%
L2
100%
glm-5
diamond-glm
-56%$0.438/M
Cache read$0.088/M
-41%$1.90/M
L1
100%
glm-5
default
-45%$0.551/M
-27%$2.34/M
L1
100%
glm-5
default
$1.64/M
$6.03/M
L1
99%
glm-5
国产模型渠道
$4.00/M
$18.00/M
Showing 20 model IDs of 72.

Alternatives & Similar Models

Frequently Asked Questions

What benchmark data does GLM-5 include?
LMSpeed shows GLM-5 benchmark context, API price, output speed, first-token latency, and provider data across 153 providers when those signals are available.
What is the GLM-5 API price?
GLM-5 has pricing from undefined provider} other undefined providers}}, ranging from $0.010/request to $150.00/M. 素墨API has the lowest listed price.
What does the GLM-5 API pricing table include?
The GLM-5 API pricing table compares 153 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
Which provider has the cheapest GLM-5 API pricing?
素墨API currently has the lowest listed GLM-5 price at $0.010/request across undefined provider} other undefined providers}}.
Can I compare GLM-5 API price and speed together?
Yes. LMSpeed shows GLM-5 API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
Is GLM-5 API free?
Yes, GLM-5 free API options are available through 3 providerundefined other undefined} on LMSpeed, including 兔子API, 兔子API, Zero API. These providers offer free API credits or a free tier with no per-token charges.
Where can I get GLM-5 free API access?
LMSpeed currently lists 3 free API providerundefined other undefined} for GLM-5: 兔子API, 兔子API, Zero API. Check each provider row before using it because free tier limits can change.

Also known as

GLM-5GLM-5-2ccGLM-5-fastPro/zai-org/GLM-5Z-AI/GLM-5

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation