GLM-4.6 API Benchmarks, Pricing & Provider Data
Compare GLM-4.6 with another model
Choose a model to open its comparison page.
GLM-4.6 benchmark, API pricing, and provider data cover 159 API providers, with prices starting at $0.010/request. GLM-4.6 free API options are available from 2 providers. The page also shows measured API speed and first-token latency.
Zhipu AI GLM-4.6 builds on GLM-4.5 with a 200K context window, stronger real-world coding, advanced reasoning with tool use, and improved agentic performance for complex multi-step tasks.
Category Performance
Observed-capability estimates with uncertainty reported separately from independent benchmark families.
- Coverage
- 6 / 8
- Methodology
- V3.0
#1Agents48.4Provisional1/4 Measured dimensions
#2Math44.8Provisional1/4 Measured dimensions
#3Knowledge43.6Provisional1/4 Measured dimensions
#4Reasoning42.8RatedGlobal rank #583/4 Measured dimensions
#5Instruction following42.4Provisional1/4 Measured dimensions
#6Coding41.9Estimated2/4 Measured dimensions
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Technical Details
- Input
- Output
- Released
- Sep 2025
- Knowledge cutoff
- 2025-03-31
- Tokenizer
- Other
- Architecture
- text->text
- Moderated
- No
- Supported parameters
- frequency_penaltyinclude_reasoninglogit_biasmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_p
Rankings
Excels at
Falls behind in
Detailed scores
Updated: Aug 30, 2026Overall
undefined metric} other undefined metrics}}
Overall score44.0#87 / 100
Overall
undefined metric} other undefined metrics}}
Pricing
undefined metric} other undefined metrics}}
Input price$0.550/M#89 / 180Output price$2.20/M#77 / 180
Pricing
undefined metric} other undefined metrics}}
Agents
V3.0undefined metric} other undefined metrics}} · Provisional
Score48.480% interval32.3–64.41/4 Measured dimensionsΤ²-bench results76.9#57 / 82
Agents
V3.0undefined metric} other undefined metrics}} · Provisional
Coding
V3.0undefined metric} other undefined metrics}} · Estimated
Score41.980% interval30.7–53.12/4 Measured dimensionsLiveCodeBench69.5%#30 / 115SciCode38.4%#103 / 205
Coding
V3.0undefined metric} other undefined metrics}} · Estimated
Reasoning
V3.0undefined metric} other undefined metrics}} · Rated
Score42.8#5880% interval34.2–51.53/4 Measured dimensionsMMLU-Pro82.9%#44 / 129GPQA78.0%#102 / 212
Reasoning
V3.0undefined metric} other undefined metrics}} · Rated
Knowledge
V3.0undefined metric} other undefined metrics}} · Provisional
Score43.680% interval27.6–59.61/4 Measured dimensionsArtificial Analysis Intelligence Index23.4#81 / 111AA-GPQA Diamond63.2#94 / 108
Knowledge
V3.0undefined metric} other undefined metrics}} · Provisional
Math
V3.0undefined metric} other undefined metrics}} · Provisional
Score44.880% interval28.7–60.81/4 Measured dimensionsMath score27.5#54 / 62FrontierMath v2 (Tiers 1-3)3.8#40 / 47
Math
V3.0undefined metric} other undefined metrics}} · Provisional
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
Multimodal
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Multimodal
V3.0undefined metric} other undefined metrics}} · No data
Instruction following
V3.0undefined metric} other undefined metrics}} · Provisional
Score42.480% interval26.4–58.41/4 Measured dimensionsAA-IFBench36.7#76 / 84
Instruction following
V3.0undefined metric} other undefined metrics}} · Provisional
OpenRouter endpoints
5 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
AtlasCloud atlas-cloud/fp8 | $0.600/M | $2.20/M | 100.0% | — | — | undefined tokens / undefined tokens |
DeepInfra deepinfra/fp4 | $0.500/M | $2/M | 99.8% | — | — | undefined tokens / undefined tokens |
Venice venice/fp4 | $0.430/M | $1.75/M | 99.8% | — | — | undefined tokens / undefined tokens |
Z.AI z-ai/fp4 | $0.600/M | $2.20/M | 98.1% | — | — | undefined tokens / undefined tokens |
Novita novita/bf16 | $0.550/M | $2.20/M | 98.0% | — | — | undefined tokens / undefined tokens |
Pricing Comparison
Compare GLM-4.6 API pricing across 157 providers. Prices range from $0.010/request to $10.50/M. 素墨API offers the lowest rate at $0.010/request. 2 providers offer free API credits or a free tier.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 100% | glm-4.6 | default | -50%$0.274/M | -50%$1.10/M | 56.0 t/s | 1.36 s | — | |
L1 100% | glm-4.6 | default | -75%$0.137/M | -75%$0.548/M | — | — | — | |
L1 100% | GLM-4.6 | default | $0.040/request | - | — | — | — | |
L1 100% | glm-4.6 | sale | $1.00/M | $4.00/M | — | — | — | |
L1 100% | glm-4.6 | default | -13%$0.479/M | -13%$1.92/M | — | — | — | |
L1 100% | zai-org/GLM-4.6 | default | $10.50/M | $42.00/M | — | — | — | |
L1 100% | glm-4.6 | default | -74%$0.144/M Cache read$0.026/M | -76%$0.527/M | — | — | — | |
L1 100% | glm-4.6 | default | -37%$0.345/M | -37%$1.38/M | — | — | — | |
L1 100% | glm-4.6 | deepseek | -30%$0.384/M | -30%$1.53/M | — | — | — | |
L1 100% | glm-4.6 | default | -75%$0.137/M | -75%$0.548/M | — | — | — | |
L1 99% | glm-4.6 | other | -55%$0.250/M Cache read$0.046/M | -58%$0.917/M | — | — | — | |
L1 100% | glm-4.6 | default | $0.548/M Cache read$0.110/M | $2.19/M | — | — | — | |
L1 100% | glm-4.6 | default | -10%$0.495/M | -4%$2.11/M | — | — | — | |
Moyanjdc API Free | L1 99% | glm-4.6 | default | Free | Free | — | — | — |
L1 99% | glm-4.6 | default | -60%$0.219/M | -60%$0.877/M | — | — | — | |
L1 99% | glm-4.6 | glm | -20%$0.438/M | -20%$1.75/M | — | — | — | |
L1 100% | glm-4.6 | default | -93%$0.041/M Cache read$0.0068/M | -93%$0.151/M | — | — | — | |
L1 99% L2 0% | glm-4.6 | default | $3.00/M Cache read$0.600/M | $14.00/M | — | — | — | |
L1 100% | glm-4.6 | default | $5.25/M | $21.00/M | — | — | — | |
L1 99% | glm-4.6 | default | $10.00/request | - | — | — | — |
Alternatives & Similar Models
GLM-4.7
glm-4-7
Zhipu GLM-4.7 is a flagship GLM release from Zhipu AI with advanced Chinese-English reasoning, coding, and agent features.
GLM-5
glm-5
Zhipu GLM-5 is Zhipu flagship GLM series model with enhanced reasoning, agent capabilities, and strong performance on Chinese enterprise and coding scenarios.
DeepSeek V3.2
deepseek-v3-2
DeepSeek V3.2 is an upgraded V3-series MoE model with stronger reasoning, coding, and math performance, widely available through OpenAI-compatible API relays.
GLM-5.1
glm-5-1
Zhipu GLM-5.1 is a next-generation GLM model aimed at frontier reasoning, coding, and bilingual agent applications.
Kimi K2.5
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.
MiniMax M2.5
minimax-m2-5
MiniMax M2.5 is MiniMax's flagship text model for coding and agents, with SOTA-level programming and agentic performance, improved token efficiency, and fast high-TPS API deployment.
Frequently Asked Questions
- What benchmark data does GLM-4.6 include?
- LMSpeed shows GLM-4.6 benchmark context, API price, output speed, first-token latency, and provider data across 159 providers when those signals are available.
- What is the GLM-4.6 API price?
- GLM-4.6 has pricing from 159 providers, ranging from $0.010/request to $10.50/M. 素墨API has the lowest listed price.
- What does the GLM-4.6 API pricing table include?
- The GLM-4.6 API pricing table compares 159 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest GLM-4.6 API pricing?
- 素墨API currently has the lowest listed GLM-4.6 price at $0.010/request across 159 providers.
- Can I compare GLM-4.6 API price and speed together?
- Yes. LMSpeed shows GLM-4.6 API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
- Is GLM-4.6 API free?
- Yes, GLM-4.6 free API options are available through 2 providerundefined other undefined} on LMSpeed, including Moyanjdc API, Zero API. These providers offer free API credits or a free tier with no per-token charges.
- Where can I get GLM-4.6 free API access?
- LMSpeed currently lists 2 free API providerundefined other undefined} for GLM-4.6: Moyanjdc API, Zero API. Check each provider row before using it because free tier limits can change.
