GLM-5.1 API Benchmarks, Pricing & Provider Data
Compare GLM-5.1 with another model
Choose a model to open its comparison page.
GLM-5.1 benchmark, API pricing, and provider data cover undefined API provider} other undefined API providers}}, with prices starting at $0.0008/M. GLM-5.1 free API options are available from undefined provider} other undefined providers}}. The page also shows measured API speed and first-token latency.
Zhipu GLM-5.1 is a next-generation GLM model aimed at frontier reasoning, coding, and bilingual agent applications.
Category Performance
Observed-capability estimates with uncertainty reported separately from independent benchmark families.
- Coverage
- 6 / 8
- Methodology
- V3.0
#1Instruction following54.9Provisional1/4 Measured dimensions
#2Coding54.3RatedGlobal rank #183/4 Measured dimensions
#3Reasoning54Estimated2/4 Measured dimensions
#4Knowledge52.9Provisional1/4 Measured dimensions
#5Agents52.5RatedGlobal rank #314/4 Measured dimensions
#6Math51.5Estimated3/4 Measured dimensions
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Technical Details
- Input
- Output
- Released
- Apr 2026
- Tokenizer
- Other
- Architecture
- text->text
- Moderated
- No
- Supported parameters
- frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
Rankings
Excels at
It's decent at
Falls behind in
Detailed scores
Updated: Sep 13, 2026Overall
undefined metric} other undefined metrics}}
Overall score60.0#48 / 112
Overall
undefined metric} other undefined metrics}}
Pricing
undefined metric} other undefined metrics}}
Input price$1.20/M#120 / 186Output price$4.40/M#119 / 186
Pricing
undefined metric} other undefined metrics}}
Agents
V3.0undefined metric} other undefined metrics}} · Rated
Score52.5#3180% interval46.8–58.24/4 Measured dimensionsAgentic score46.9#60 / 77Terminal-Bench 2.063.5#28 / 53
Agents
V3.0undefined metric} other undefined metrics}} · Rated
Coding
V3.0undefined metric} other undefined metrics}} · Rated
Score54.3#1880% interval45.6–63.13/4 Measured dimensionsSciCode44.8%#46 / 89Coding score52.1#49 / 87
Coding
V3.0undefined metric} other undefined metrics}} · Rated
Reasoning
V3.0undefined metric} other undefined metrics}} · Estimated
Score5480% interval43.3–64.82/4 Measured dimensionsGPQA86.8%#55 / 218HLE30.1%#56 / 216
Reasoning
V3.0undefined metric} other undefined metrics}} · Estimated
Knowledge
V3.0undefined metric} other undefined metrics}} · Provisional
Score52.980% interval39.0–66.91/4 Measured dimensionsKnowledge score71.2#37 / 83GPQA-D86.2#29 / 32
Knowledge
V3.0undefined metric} other undefined metrics}} · Provisional
Math
V3.0undefined metric} other undefined metrics}} · Estimated
Score51.580% interval42.4–60.53/4 Measured dimensionsMath score64.1#25 / 63AIME2695.3#7 / 14
Math
V3.0undefined metric} other undefined metrics}} · Estimated
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
Multimodal
V3.0undefined metric} other undefined metrics}} · No data
80% interval30.8–69.20/4 Measured dimensionsDesign Arena Website1290.0#18 / 78
Multimodal
V3.0undefined metric} other undefined metrics}} · No data
Instruction following
V3.0undefined metric} other undefined metrics}} · Provisional
Score54.980% interval38.9–70.91/4 Measured dimensionsAA-IFBench76.3#10 / 84Instruction Following score93.7#2 / 52
Instruction following
V3.0undefined metric} other undefined metrics}} · Provisional
OpenRouter endpoints
15 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
Alibaba alibaba/fp8 | $1.33/M | $4.18/M | 100.0% | — | — | undefined tokens / undefined tokens |
SiliconFlow siliconflow/fp8 | $1.19/M | $3.74/M | 100.0% | — | — | undefined tokens / undefined tokens |
Friendli friendli | $1.40/M | $4.40/M | 100.0% | — | — | undefined tokens / undefined tokens |
DeepInfra deepinfra/fp4 | $1.05/M | $3.50/M | 100.0% | — | — | undefined tokens / undefined tokens |
StreamLake streamlake/fp8 | $0.966/M | $3.04/M | 99.8% | — | — | undefined tokens / undefined tokens |
Z.AI z-ai/fp8 | $1.40/M | $4.40/M | 99.8% | — | — | undefined tokens / undefined tokens |
GMICloud gmicloud/fp8 | $1.40/M | $4.40/M | 99.7% | — | — | undefined tokens / undefined tokens |
Novita novita/fp8 | $1.38/M | $4.40/M | 99.6% | — | — | undefined tokens / undefined tokens |
Baidu baidu/fp8 | $0.965/M | $3.03/M | 99.3% | — | — | undefined tokens / undefined tokens |
AtlasCloud atlas-cloud/fp8 | $1.26/M | $3.96/M | 99.2% | — | — | undefined tokens / undefined tokens |
Venice venice/fp8 | $1.40/M | $4.40/M | 99.2% | — | — | undefined tokens / undefined tokens |
Crusoe crusoe/fp8 | $1.20/M | $4.40/M | 98.5% | — | — | undefined tokens / undefined tokens |
Nebius nebius/fp8 | $1.40/M | $4.40/M | 94.0% | — | — | undefined tokens / undefined tokens |
Phala phala | $1.21/M | $4.20/M | 90.2% | — | — | undefined tokens / undefined tokens |
Chutes chutes/fp8 | $0.980/M | $3.08/M | 89.8% | — | — | undefined tokens / undefined tokens |
Pricing Comparison
Compare GLM-5.1 API pricing across 183 providers. Prices range from $0.0008/M to $97.50/M. S3AI API offers the lowest rate at $0.0008/M. 5 providers offer free API credits or a free tier.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 99% | glm-5.1 | 临时渠道 | -93%$0.082/M Cache read$0.018/M | -93%$0.329/M | 148.1 t/s | 5.30 s | — | |
L1 99% | z-ai/glm-5.1 | 临时渠道 | -99%$0.017/M Cache read$0.0037/M | -99%$0.054/M | — | — | ||
L1 99% L2 8% | z-ai/glm-5.1 | default | -7%$1.12/M Cache read$0.208/M | -20%$3.52/M | 97.8 t/s | 17.08 s | ||
L1 2% | glm-5.1 | default | $75.00/M | $75.00/M | 57.4 t/s | 5.14 s | — | |
L1 100% | glm-5.1 | default | -8%$1.10/M Cache read$0.268/M | -7%$4.08/M | 41.9 t/s | 15.34 s | — | |
L1 100% | glm-5.1 | sale | $3.00/M | $12.00/M | — | — | — | |
L1 100% | glm-5.1 | default | -66%$0.411/M Cache read$0.090/M | -63%$1.64/M | — | — | — | |
L1 100% L2 0% | z-ai/glm-5.1 | default | $1.40/M Cache read$0.260/M | $4.40/M | — | — | — | |
L1 100% | glm-5.1 | default | -66%$0.411/M Cache read$0.090/M | -63%$1.64/M | — | — | — | |
L1 100% | glm-5.1 | default | -36%$0.767/M | -69%$1.38/M | — | — | — | |
L1 100% | glm-5.1 | default | -9%$1.10/M | -13%$3.84/M | — | — | — | |
L1 100% | Pro/zai-org/GLM-5.1 | default | $10.27/M | $10.27/M | — | — | — | |
L1 100% | glm-5.1 | default | -92%$0.096/M Cache read$0.018/M | -93%$0.301/M | — | — | — | |
L1 100% | glm-5.1 | default | -89%$0.137/M | -81%$0.822/M | — | — | — | |
L1 100% | glm-5.1 | Self-Deployed-2 | -83%$0.208/M Cache read$0.046/M | -85%$0.652/M | — | — | — | |
L1 100% | glm-5.1 | 国产模型 | $1.25/M Cache read$1.25/MCache write$0.232/MCache write 1h$0.371/M | -11%$3.92/M | — | — | — | |
L1 100% | glm-5.1 | default | $6.00/M Cache read$1.20/MCache write$7.50/MCache write 1h$12.00/M | $24.00/M | — | — | — | |
L1 100% | [mt]q|次/glm-5.1 | default | $9.00/request | - | — | — | — | |
DeadlySignal API Free | L1 99% L2 100% | glm-5.1 | default | Free | Free | — | — | — |
L1 100% L2 100% | glm-5.1 | OpenModels | -33%$0.800/M Cache read$0.200/MCache write$0.800/MCache write 1h$1.28/M | -36%$2.80/M | — | — | — |
Alternatives & Similar Models
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
DeepSeek V4 Pro
deepseek-v4-pro
DeepSeek V4 Pro is the professional-tier DeepSeek V4 model, targeting frontier reasoning, coding, and agent workflows with maximum capability.
Kimi K2.5
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.
GLM-5
glm-5
Zhipu GLM-5 is Zhipu flagship GLM series model with enhanced reasoning, agent capabilities, and strong performance on Chinese enterprise and coding scenarios.
MiniMax M2.7
minimax-m2-7
MiniMax M2.7 is a high-tier M2-series model tuned for complex reasoning, long-context dialogue, and production-grade API workloads.
GPT-5.4
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
Frequently Asked Questions
- What benchmark data does GLM-5.1 include?
- LMSpeed shows GLM-5.1 benchmark context, API price, output speed, first-token latency, and provider data across 188 providers when those signals are available.
- What is the GLM-5.1 API price?
- GLM-5.1 has pricing from undefined provider} other undefined providers}}, ranging from $0.0008/M to $97.50/M. S3AI API has the lowest listed price.
- What does the GLM-5.1 API pricing table include?
- The GLM-5.1 API pricing table compares 188 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest GLM-5.1 API pricing?
- S3AI API currently has the lowest listed GLM-5.1 price at $0.0008/M across undefined provider} other undefined providers}}.
- Can I compare GLM-5.1 API price and speed together?
- Yes. LMSpeed shows GLM-5.1 API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
- Is GLM-5.1 API free?
- Yes, GLM-5.1 free API options are available through 5 providerundefined other undefined} on LMSpeed, including DeadlySignal API, Moyanjdc API, 兔子API, Zero API, 兔子API. These providers offer free API credits or a free tier with no per-token charges.
- Where can I get GLM-5.1 free API access?
- LMSpeed currently lists 5 free API providerundefined other undefined} for GLM-5.1: DeadlySignal API, Moyanjdc API, 兔子API, Zero API, 兔子API. Check each provider row before using it because free tier limits can change.
