GLM-5 API Benchmarks, Pricing & Provider Data
Compare GLM-5 with another model
Choose a model to open its comparison page.
GLM-5 benchmark, API pricing, and provider data cover undefined API provider} other undefined API providers}}, with prices starting at $0.010/request. GLM-5 free API options are available from undefined provider} other undefined providers}}. The page also shows measured API speed and first-token latency.
Zhipu GLM-5 is Zhipu flagship GLM series model with enhanced reasoning, agent capabilities, and strong performance on Chinese enterprise and coding scenarios.
Category Performance
Observed-capability estimates with uncertainty reported separately from independent benchmark families.
- Coverage
- 7 / 8
- Methodology
- V3.0
#1Instruction following52.4Estimated2/4 Measured dimensions
#2Coding51.7RatedGlobal rank #223/4 Measured dimensions
#3Math51Estimated3/4 Measured dimensions
#4Reasoning48.8RatedGlobal rank #453/4 Measured dimensions
#5Multilingual43.8Estimated2/4 Measured dimensions
#6Knowledge43.2Provisional1/4 Measured dimensions
#7Agents41.4RatedGlobal rank #584/4 Measured dimensions
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Technical Details
- Input
- Output
- Released
- Feb 2026
- Tokenizer
- Other
- Architecture
- text->text
- Moderated
- No
- Supported parameters
- frequency_penaltyinclude_reasoninglogprobsmax_tokenspresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
Rankings
Excels at
It's decent at
Falls behind in
Detailed scores
Updated: Sep 13, 2026Overall
undefined metric} other undefined metrics}}
Overall score54.0#72 / 112
Overall
undefined metric} other undefined metrics}}
Pricing
undefined metric} other undefined metrics}}
Input price$1.00/M#113 / 186Output price$3.20/M#105 / 186
Pricing
undefined metric} other undefined metrics}}
Agents
V3.0undefined metric} other undefined metrics}} · Rated
Score41.4#5880% interval35.7–47.04/4 Measured dimensionsAgentic score42.8#62 / 77Terminal-Bench 2.056.2#39 / 53
Agents
V3.0undefined metric} other undefined metrics}} · Rated
Coding
V3.0undefined metric} other undefined metrics}} · Rated
Score51.7#2280% interval43.5–59.83/4 Measured dimensionsCoding score49.5#57 / 87SWE-bench Verified77.8#21 / 49
Coding
V3.0undefined metric} other undefined metrics}} · Rated
Reasoning
V3.0undefined metric} other undefined metrics}} · Rated
Score48.8#4580% interval40.7–57.03/4 Measured dimensionsMMLU-Pro85.7%#26 / 129GPQA82.0%#90 / 218
Reasoning
V3.0undefined metric} other undefined metrics}} · Rated
Knowledge
V3.0undefined metric} other undefined metrics}} · Provisional
Score43.280% interval26.9–59.51/4 Measured dimensionsKnowledge score72.8#32 / 83GPQA-D86.0#30 / 32
Knowledge
V3.0undefined metric} other undefined metrics}} · Provisional
Math
V3.0undefined metric} other undefined metrics}} · Estimated
Score5180% interval41.9–60.03/4 Measured dimensionsMath score56.9#32 / 63AIME2695.8#4 / 14
Math
V3.0undefined metric} other undefined metrics}} · Estimated
Multilingual
V3.0undefined metric} other undefined metrics}} · Estimated
Score43.880% interval31.6–56.02/4 Measured dimensionsMultilingual score48.7#6 / 11MMLU-ProX83.1#6 / 11
Multilingual
V3.0undefined metric} other undefined metrics}} · Estimated
Multimodal
V3.0undefined metric} other undefined metrics}} · No data
80% interval30.8–69.20/4 Measured dimensionsDesign Arena Website1260.0#34 / 78
Multimodal
V3.0undefined metric} other undefined metrics}} · No data
Instruction following
V3.0undefined metric} other undefined metrics}} · Estimated
Score52.480% interval40.5–64.42/4 Measured dimensionsInstruction Following score88.5#23 / 52IFEval92.6#8 / 16
Instruction following
V3.0undefined metric} other undefined metrics}} · Estimated
OpenRouter endpoints
8 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
Novita novita/fp8 | $1/M | $3.20/M | 100% | — | — | undefined tokens / undefined tokens |
Z.AI z-ai/fp8 | $1/M | $3.20/M | 100.0% | — | — | undefined tokens / undefined tokens |
SiliconFlow siliconflow/fp8 | $0.950/M | $2.55/M | 99.9% | — | — | undefined tokens / undefined tokens |
StreamLake streamlake/fp8 | $0.600/M | $1.92/M | 99.7% | — | — | undefined tokens / undefined tokens |
GMICloud gmicloud/fp8 | $0.600/M | $1.92/M | 99.6% | — | — | undefined tokens / undefined tokens |
Baidu baidu/fp8 | $0.700/M | $2.24/M | 99.5% | — | — | undefined tokens / undefined tokens |
Venice venice/fp8 | $1/M | $3.20/M | 98.9% | — | — | undefined tokens / undefined tokens |
Amazon Bedrock amazon-bedrock | $1/M | $3.20/M | 97.4% | — | — | undefined tokens / undefined tokens |
Pricing Comparison
Compare GLM-5 API pricing across 150 providers. Prices range from $0.010/request to $150.00/M. 素墨API offers the lowest rate at $0.010/request. 3 providers offer free API credits or a free tier.
Alternatives & Similar Models
GLM-5.1
glm-5-1
Zhipu GLM-5.1 is a next-generation GLM model aimed at frontier reasoning, coding, and bilingual agent applications.
Kimi K2.5
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.
MiniMax M2.5
minimax-m2-5
MiniMax M2.5 is MiniMax's flagship text model for coding and agents, with SOTA-level programming and agentic performance, improved token efficiency, and fast high-TPS API deployment.
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
DeepSeek V4 Pro
deepseek-v4-pro
DeepSeek V4 Pro is the professional-tier DeepSeek V4 model, targeting frontier reasoning, coding, and agent workflows with maximum capability.
GPT-5.4
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
Frequently Asked Questions
- What benchmark data does GLM-5 include?
- LMSpeed shows GLM-5 benchmark context, API price, output speed, first-token latency, and provider data across 153 providers when those signals are available.
- What is the GLM-5 API price?
- GLM-5 has pricing from undefined provider} other undefined providers}}, ranging from $0.010/request to $150.00/M. 素墨API has the lowest listed price.
- What does the GLM-5 API pricing table include?
- The GLM-5 API pricing table compares 153 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest GLM-5 API pricing?
- 素墨API currently has the lowest listed GLM-5 price at $0.010/request across undefined provider} other undefined providers}}.
- Can I compare GLM-5 API price and speed together?
- Yes. LMSpeed shows GLM-5 API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
- Is GLM-5 API free?
- Yes, GLM-5 free API options are available through 3 providerundefined other undefined} on LMSpeed, including 兔子API, 兔子API, Zero API. These providers offer free API credits or a free tier with no per-token charges.
- Where can I get GLM-5 free API access?
- LMSpeed currently lists 3 free API providerundefined other undefined} for GLM-5: 兔子API, 兔子API, Zero API. Check each provider row before using it because free tier limits can change.
