Z.ai
·Released on Apr 7, 2026

GLM-5.1 API Benchmarks, Pricing & Provider Data

Compare GLM-5.1 with another model

Choose a model to open its comparison page.

Share on X
LLM

GLM-5.1 benchmark, API pricing, and provider data cover undefined API provider} other undefined API providers}}, with prices starting at $0.0008/M. GLM-5.1 free API options are available from undefined provider} other undefined providers}}. The page also shows measured API speed and first-token latency.

Zhipu GLM-5.1 is a next-generation GLM model aimed at frontier reasoning, coding, and bilingual agent applications.

Quality
#48of 112
60.0
LMSpeed score
Speed
50char/s
14.60 s
Cost
#120of 186
$0.0008/ 1M · 8:1 in:out
$0.0001 in · $0.0007 out

Category Performance

Observed-capability estimates with uncertainty reported separately from independent benchmark families.

Coverage
6 / 8
Methodology
V3.0
Category PerformanceObserved-capability estimates with uncertainty reported separately from independent benchmark families.Agents52.5Coding54.3Reasoning54Knowledge52.9Math51.5Multilingual-Multimodal-Instruction following54.9
#1Instruction following54.9Provisional1/4 Measured dimensions
80% interval: 38.970.9
constraint followingPrior only
structured outputPrior only
novel instruction generalization54.9
ifbench · aa_if_bench
long multiturn instructionPrior only
#2Coding54.3RatedGlobal rank #183/4 Measured dimensions
80% interval: 45.663.1
code generation52.4
scicode · scicode
repository engineering50.2
nl2repo · nl2_repo / swe_pro · swe_pro / vibecode · vibe_code_bench
debugging testing60.4
swe_rebench · swe_rebench
tooling qualityPrior only
#3Reasoning54Estimated2/4 Measured dimensions
80% interval: 43.364.8
abstract logicPrior only
scientific causal55.4
critpt · critpt / gpqa · gpqa / hle · hle
multistep constraintsPrior only
evidence verification52.7
lcr · lcr
#4Knowledge52.9Provisional1/4 Measured dimensions
80% interval: 39.066.9
broad knowledge52.9
aa_omniscience · aa_omniscience_index / benchlm_category_knowledge · benchlm_category_knowledge
professional knowledgePrior only
factualityPrior only
retrieval open bookPrior only
#5Agents52.5RatedGlobal rank #314/4 Measured dimensions
80% interval: 46.858.2
planning55.7
deep_planning · gert_labs
tool use54.3
mcp_atlas · mcp_atlas / tau · tau2_bench / tau · tau3_bench
environment execution48
browsecomp · browse_comp / gdpval_aa · benchlm_agentic_gdpval_aa / terminalbench · benchlm_agentic_terminal_bench2
recovery reliability52.1
claw_eval · claw_eval / cyber_gym · cyber_gym / researchclaw · research_claw_bench
#6Math51.5Estimated3/4 Measured dimensions
80% interval: 42.460.5
foundational mathPrior only
competition math48.7
aime · aime2026 / hmmt · hmmt_feb2026 / hmmt · hmmt_nov2025
proof frontier53.6
frontiermath · frontier_math_v2_tier4 / frontiermath · frontier_math_v2_tiers13
applied tool math52
mmanswer · mm_answer_bench
No data:MultilingualMultimodal

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
204.8Ktokens
245.8 pages of text
OUTPUT
128Ktokens
8K128K1M4M
204.8K

Features

Technical Details

Input
Output
Released
Apr 2026
Tokenizer
Other
Architecture
text->text
Moderated
No
Supported parameters
frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p

Rankings

Excels at

It's decent at

Falls behind in

Detailed scores

Updated: Sep 13, 2026

Overall

undefined metric} other undefined metrics}}

Overall score60.0#48 / 112
Overall score
LMSpeed rank#48
Score60.0
UpdatedSep 13, 2026
Confidence3

Pricing

undefined metric} other undefined metrics}}

Input price$1.20/M#120 / 186Output price$4.40/M#119 / 186
Input price
LMSpeed rank#120
Score$1.20/M
UpdatedSep 13, 2026
Confidence4
Output price
LMSpeed rank#119
Score$4.40/M
UpdatedSep 13, 2026
Confidence4

Agents

V3.0

undefined metric} other undefined metrics}} · Rated

Score52.5#3180% interval46.8–58.24/4 Measured dimensionsAgentic score46.9#60 / 77Terminal-Bench 2.063.5#28 / 53
Dimensions and evidence
Planning & decomposition55.7
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
deep_planning · gert_labs · z 0.77 · q 1.00
Tool use54.3
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
mcp_atlas · mcp_atlas · z -0.15 · q 1.00
tau · tau2_bench · z 1.08 · q 1.00
tau · tau3_bench · z 0.16 · q 1.00
Environment & long-horizon execution48
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
browsecomp · browse_comp · z -1.03 · q 1.00
gdpval_aa · benchlm_agentic_gdpval_aa · z 0.04 · q 1.00
terminalbench · benchlm_agentic_terminal_bench2 · z -0.02 · q 1.00
Recovery & completion reliability52.1
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
claw_eval · claw_eval · z 0.45 · q 1.00
cyber_gym · cyber_gym · z -0.82 · q 1.00
researchclaw · research_claw_bench · z 0.06 · q 1.00
Agentic score
LMSpeed rank#60
Score46.9
UpdatedSep 13, 2026
Confidence3
Terminal-Bench 2.0V3 evidence
LMSpeed rank#28
Score63.5
UpdatedSep 13, 2026
Confidence3
BrowseCompV3 evidence
LMSpeed rank#24
Score68.0
UpdatedSep 13, 2026
Confidence3
Τ³-bench resultsV3 evidence
LMSpeed rank#4
Score70.6
UpdatedSep 13, 2026
Confidence3
MCP AtlasV3 evidence
LMSpeed rank#18
Score71.8
UpdatedSep 13, 2026
Confidence3
CyberGymV3 evidence
LMSpeed rank#10
Score68.7
UpdatedSep 13, 2026
Confidence3
Claw-EvalV3 evidence
LMSpeed rank#10
Score62.3
UpdatedSep 13, 2026
Confidence3
AA Agentic Index
LMSpeed rank#36
Score25.2
UpdatedSep 13, 2026
Confidence3
Τ²-bench resultsV3 evidence
LMSpeed rank#9
Score97.7
UpdatedSep 13, 2026
Confidence3
GDPval-AA
LMSpeed rank#34
Score34.0
UpdatedSep 13, 2026
Confidence3
Gert LabsV3 evidence
LMSpeed rank#12
Score60.1
UpdatedSep 13, 2026
Confidence3
GDPval-AAV3 evidence
LMSpeed rank#35
Score1181.0
UpdatedSep 13, 2026
Confidence3
ResearchClawBenchV3 evidence
LMSpeed rank#7
Score18.2
UpdatedSep 13, 2026
Confidence3
Terminal-Bench 2.1 (Vals)
LMSpeed rank#27
Score56.9
UpdatedSep 13, 2026
Confidence3

Coding

V3.0

undefined metric} other undefined metrics}} · Rated

Score54.3#1880% interval45.6–63.13/4 Measured dimensionsSciCode44.8%#46 / 89Coding score52.1#49 / 87
Dimensions and evidence
Code generation52.4
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
scicode · scicode · z 0.15 · q 1.00
Repository engineering50.2
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
nl2repo · nl2_repo · z -0.04 · q 1.00
swe_pro · swe_pro · z 0.23 · q 1.00
vibecode · vibe_code_bench · z 0.25 · q 1.00
Debugging & testing60.4
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
swe_rebench · swe_rebench · z 1.49 · q 1.00
Tool-assisted development & qualityPrior only
SciCodeV3 evidence
LMSpeed rank#46
Score44.8%
UpdatedSep 13, 2026
Confidence4
Coding score
LMSpeed rank#49
Score52.1
UpdatedSep 13, 2026
Confidence3
SWE-bench ProV3 evidence
LMSpeed rank#20
Score58.4
UpdatedSep 13, 2026
Confidence3
NL2RepoV3 evidence
LMSpeed rank#9
Score42.7
UpdatedSep 13, 2026
Confidence3
SWE-RebenchV3 evidence
LMSpeed rank#3
Score62.7
UpdatedSep 13, 2026
Confidence3
Vibe Code BenchV3 evidence
LMSpeed rank#14
Score31.5
UpdatedSep 13, 2026
Confidence3
AA Coding Index
LMSpeed rank#39
Score55.8
UpdatedSep 13, 2026
Confidence3
AA-SciCodeV3 evidence
LMSpeed rank#57
Score44.8
UpdatedSep 13, 2026
Confidence3
OpenHarmony Bench
LMSpeed rank#6
Score52.3
UpdatedSep 13, 2026
Confidence3
LiveCodeBench (Vals)
LMSpeed rank#32
Score81.4
UpdatedSep 13, 2026
Confidence3
SWE-bench (Vals)
LMSpeed rank#27
Score76.4
UpdatedSep 13, 2026
Confidence3

Reasoning

V3.0

undefined metric} other undefined metrics}} · Estimated

Score5480% interval43.3–64.82/4 Measured dimensionsGPQA86.8%#55 / 218HLE30.1%#56 / 216
Dimensions and evidence
Abstract logicPrior only
Scientific & causal reasoning55.4
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
critpt · critpt · z 0.16 · q 1.00
gpqa · gpqa · z 0.37 · q 1.00
hle · hle · z 0.51 · q 1.00
Multi-step constraintsPrior only
Evidence integration & verification52.7
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
lcr · lcr · z 0.01 · q 1.00
GPQAV3 evidence
LMSpeed rank#55
Score86.8%
UpdatedSep 13, 2026
Confidence4
HLEV3 evidence
LMSpeed rank#56
Score30.1%
UpdatedSep 13, 2026
Confidence4
AA-LCRV3 evidence
LMSpeed rank#54
Score73.7
UpdatedSep 13, 2026
Confidence3
CritPtV3 evidence
LMSpeed rank#51
Score4.6
UpdatedSep 13, 2026
Confidence3
Reasoning score
LMSpeed rank#33
Score71.7
UpdatedSep 13, 2026
Confidence3

Knowledge

V3.0

undefined metric} other undefined metrics}} · Provisional

Score52.980% interval39.0–66.91/4 Measured dimensionsKnowledge score71.2#37 / 83GPQA-D86.2#29 / 32
Dimensions and evidence
Broad knowledge52.9
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
aa_omniscience · aa_omniscience_index · z 0.15 · q 1.00
benchlm_category_knowledge · benchlm_category_knowledge · z 0.09 · q 1.00
Professional knowledgePrior only
FactualityPrior only
Retrieval & open-book usePrior only
Knowledge scoreV3 evidence
LMSpeed rank#37
Score71.2
UpdatedSep 13, 2026
Confidence3
GPQA-D
LMSpeed rank#29
Score86.2
UpdatedSep 13, 2026
Confidence3
Artificial Analysis Intelligence IndexV3 evidence
LMSpeed rank#56
Score26.4
UpdatedSep 13, 2026
Confidence3
AA-GPQA Diamond
LMSpeed rank#51
Score86.8
UpdatedSep 13, 2026
Confidence3
AA-HLE
LMSpeed rank#49
Score30.1
UpdatedSep 13, 2026
Confidence3
AA-Omniscience IndexV3 evidence
LMSpeed rank#40
Score0.9
UpdatedSep 13, 2026
Confidence3
AA-Omniscience Accuracy
LMSpeed rank#79
Score23.7
UpdatedSep 13, 2026
Confidence3
AA-Omniscience Hallucination Rate
LMSpeed rank#99
Score29.9
UpdatedSep 13, 2026
Confidence3
GPQA Diamond (Vals)
LMSpeed rank#33
Score84.5
UpdatedSep 13, 2026
Confidence3
MMLU-Pro (Vals)
LMSpeed rank#25
Score86.9
UpdatedSep 13, 2026
Confidence3

Math

V3.0

undefined metric} other undefined metrics}} · Estimated

Score51.580% interval42.4–60.53/4 Measured dimensionsMath score64.1#25 / 63AIME2695.3#7 / 14
Dimensions and evidence
Foundational mathPrior only
Competition math48.7
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
aime · aime2026 · z 0.00 · q 1.00
hmmt · hmmt_feb2026 · z -0.62 · q 1.00
hmmt · hmmt_nov2025 · z 0.32 · q 1.00
Advanced proofs53.6
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
frontiermath · frontier_math_v2_tier4 · z 0.40 · q 1.00
frontiermath · frontier_math_v2_tiers13 · z 0.27 · q 1.00
Applied & tool-assisted math52
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
mmanswer · mm_answer_bench · z 0.32 · q 1.00
Math score
LMSpeed rank#25
Score64.1
UpdatedSep 13, 2026
Confidence3
AIME26V3 evidence
LMSpeed rank#7
Score95.3
UpdatedSep 13, 2026
Confidence3
HMMT Nov 2025V3 evidence
LMSpeed rank#4
Score94.0
UpdatedSep 13, 2026
Confidence3
HMMT Feb 2026V3 evidence
LMSpeed rank#18
Score82.6
UpdatedSep 13, 2026
Confidence3
MMAnswerBenchV3 evidence
LMSpeed rank#4
Score83.8
UpdatedSep 13, 2026
Confidence3
FrontierMath v2 (Tiers 1-3)V3 evidence
LMSpeed rank#17
Score33.4
UpdatedSep 13, 2026
Confidence3
FrontierMath v2 (Tier 4)V3 evidence
LMSpeed rank#17
Score12.5
UpdatedSep 13, 2026
Confidence3

Multilingual

V3.0

undefined metric} other undefined metrics}} · No data

No data80% interval30.8–69.20/4 Measured dimensions
Dimensions and evidence
Cross-language understandingPrior only
Multilingual generationPrior only
Reasoning transferPrior only
Low-resource robustnessPrior only

Multimodal

V3.0

undefined metric} other undefined metrics}} · No data

80% interval30.8–69.20/4 Measured dimensionsDesign Arena Website1290.0#18 / 78
Dimensions and evidence
Perception & OCRPrior only
Document & spatial understandingPrior only
Visual reasoningPrior only
Video & grounded actionPrior only
Design Arena Website
LMSpeed rank#18
Score1290.0
UpdatedSep 13, 2026
Confidence3

Instruction following

V3.0

undefined metric} other undefined metrics}} · Provisional

Score54.980% interval38.9–70.91/4 Measured dimensionsAA-IFBench76.3#10 / 84Instruction Following score93.7#2 / 52
Dimensions and evidence
Constraint followingPrior only
Structured outputPrior only
Novel-instruction generalization54.9
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
ifbench · aa_if_bench · z 0.37 · q 1.00
Multi-turn & long instructionsPrior only
AA-IFBenchV3 evidence
LMSpeed rank#10
Score76.3
UpdatedSep 13, 2026
Confidence3
Instruction Following score
LMSpeed rank#2
Score93.7
UpdatedSep 13, 2026
Confidence3

OpenRouter endpoints

15 endpoints

Third-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.

Provider endpointInputOutput1d uptime30m latency30m throughputContext / output
Alibaba
alibaba/fp8
$1.33/M$4.18/M100.0%undefined tokens / undefined tokens
SiliconFlow
siliconflow/fp8
$1.19/M$3.74/M100.0%undefined tokens / undefined tokens
Friendli
friendli
$1.40/M$4.40/M100.0%undefined tokens / undefined tokens
DeepInfra
deepinfra/fp4
$1.05/M$3.50/M100.0%undefined tokens / undefined tokens
StreamLake
streamlake/fp8
$0.966/M$3.04/M99.8%undefined tokens / undefined tokens
Z.AI
z-ai/fp8
$1.40/M$4.40/M99.8%undefined tokens / undefined tokens
GMICloud
gmicloud/fp8
$1.40/M$4.40/M99.7%undefined tokens / undefined tokens
Novita
novita/fp8
$1.38/M$4.40/M99.6%undefined tokens / undefined tokens
Baidu
baidu/fp8
$0.965/M$3.03/M99.3%undefined tokens / undefined tokens
AtlasCloud
atlas-cloud/fp8
$1.26/M$3.96/M99.2%undefined tokens / undefined tokens
Venice
venice/fp8
$1.40/M$4.40/M99.2%undefined tokens / undefined tokens
Crusoe
crusoe/fp8
$1.20/M$4.40/M98.5%undefined tokens / undefined tokens
Nebius
nebius/fp8
$1.40/M$4.40/M94.0%undefined tokens / undefined tokens
Phala
phala
$1.21/M$4.20/M90.2%undefined tokens / undefined tokens
Chutes
chutes/fp8
$0.980/M$3.08/M89.8%undefined tokens / undefined tokens

Pricing Comparison

Compare GLM-5.1 API pricing across 183 providers. Prices range from $0.0008/M to $97.50/M. S3AI API offers the lowest rate at $0.0008/M. 5 providers offer free API credits or a free tier.

ProviderHealthModel VariantGroupInput ($/M)Output ($/M)Speed (t/s)First tokenAudit
L1
99%
glm-5.1
临时渠道
-93%$0.082/M
Cache read$0.018/M
-93%$0.329/M
148.1 t/s
5.30 s
L1
99%
z-ai/glm-5.1
临时渠道
-99%$0.017/M
Cache read$0.0037/M
-99%$0.054/M
L1
99%
L2
8%
z-ai/glm-5.1
default
-7%$1.12/M
Cache read$0.208/M
-20%$3.52/M
97.8 t/s
17.08 s
L1
2%
glm-5.1
default
$75.00/M
$75.00/M
57.4 t/s
5.14 s
L1
100%
glm-5.1
default
-8%$1.10/M
Cache read$0.268/M
-7%$4.08/M
41.9 t/s
15.34 s
L1
100%
glm-5.1
sale
$3.00/M
$12.00/M
L1
100%
glm-5.1
default
-66%$0.411/M
Cache read$0.090/M
-63%$1.64/M
L1
100%
L2
0%
z-ai/glm-5.1
default
$1.40/M
Cache read$0.260/M
$4.40/M
L1
100%
glm-5.1
default
-66%$0.411/M
Cache read$0.090/M
-63%$1.64/M
L1
100%
glm-5.1
default
-36%$0.767/M
-69%$1.38/M
L1
100%
glm-5.1
default
-9%$1.10/M
-13%$3.84/M
L1
100%
Pro/zai-org/GLM-5.1
default
$10.27/M
$10.27/M
L1
100%
glm-5.1
default
-92%$0.096/M
Cache read$0.018/M
-93%$0.301/M
L1
100%
glm-5.1
default
-89%$0.137/M
-81%$0.822/M
L1
100%
glm-5.1
Self-Deployed-2
-83%$0.208/M
Cache read$0.046/M
-85%$0.652/M
L1
100%
glm-5.1
国产模型
$1.25/M
Cache read$1.25/MCache write$0.232/MCache write 1h$0.371/M
-11%$3.92/M
L1
100%
glm-5.1
default
$6.00/M
Cache read$1.20/MCache write$7.50/MCache write 1h$12.00/M
$24.00/M
L1
100%
[mt]q|次/glm-5.1
default
$9.00/request
-
L1
99%
L2
100%
glm-5.1
default
Free
Free
L1
100%
L2
100%
glm-5.1
OpenModels
-33%$0.800/M
Cache read$0.200/MCache write$0.800/MCache write 1h$1.28/M
-36%$2.80/M
Showing 20 model IDs of 90.

Alternatives & Similar Models

Frequently Asked Questions

What benchmark data does GLM-5.1 include?
LMSpeed shows GLM-5.1 benchmark context, API price, output speed, first-token latency, and provider data across 188 providers when those signals are available.
What is the GLM-5.1 API price?
GLM-5.1 has pricing from undefined provider} other undefined providers}}, ranging from $0.0008/M to $97.50/M. S3AI API has the lowest listed price.
What does the GLM-5.1 API pricing table include?
The GLM-5.1 API pricing table compares 188 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
Which provider has the cheapest GLM-5.1 API pricing?
S3AI API currently has the lowest listed GLM-5.1 price at $0.0008/M across undefined provider} other undefined providers}}.
Can I compare GLM-5.1 API price and speed together?
Yes. LMSpeed shows GLM-5.1 API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
Is GLM-5.1 API free?
Yes, GLM-5.1 free API options are available through 5 providerundefined other undefined} on LMSpeed, including DeadlySignal API, Moyanjdc API, 兔子API, Zero API, 兔子API. These providers offer free API credits or a free tier with no per-token charges.
Where can I get GLM-5.1 free API access?
LMSpeed currently lists 5 free API providerundefined other undefined} for GLM-5.1: DeadlySignal API, Moyanjdc API, 兔子API, Zero API, 兔子API. Check each provider row before using it because free tier limits can change.

Also known as

GLM-5-1GLM-5.1GLM-5.1(nsfw)GLM-5.1-免费Pro/zai-org/GLM-5.1

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation