GLM 5.3 API Benchmarks, Pricing & Provider Data
Compare GLM 5.3 with another model
Choose a model to open its comparison page.
GLM 5.3 benchmark, API pricing, and provider data cover undefined API provider} other undefined API providers}}, with prices starting at $0.0080/M. GLM 5.3 free API options are available from undefined provider} other undefined providers}}. The page also shows measured API speed and first-token latency.
GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves....
Category Performance
Observed-capability estimates with uncertainty reported separately from independent benchmark families.
- Coverage
- 3 / 8
- Methodology
- V3.0
#1Coding62.9Provisional1/4 Measured dimensions
#2Reasoning60.8Provisional1/4 Measured dimensions
#3Agents50.8Provisional1/4 Measured dimensions
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Technical Details
- Input
- Output
- Released
- Aug 2026
- Tokenizer
- Other
- Architecture
- text->text
- Moderated
- No
- Supported parameters
- frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_pparallel_tool_callspresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
Rankings
Excels at
Detailed scores
Updated: Sep 14, 2026Overall
undefined metric} other undefined metrics}}
Overall score72.0#14 / 112
Overall
undefined metric} other undefined metrics}}
Speed & latency
undefined metric} other undefined metrics}}
Output speed68.1 tok/s#62 / 80Time to first token2.70 s#56 / 80
Speed & latency
undefined metric} other undefined metrics}}
Pricing
undefined metric} other undefined metrics}}
Input price$1.40/M#136 / 186Output price$4.40/M#119 / 186
Pricing
undefined metric} other undefined metrics}}
Agents
V3.0undefined metric} other undefined metrics}} · Provisional
Score50.880% interval32.4–69.11/4 Measured dimensionsAgentic score84.8#7 / 77
Agents
V3.0undefined metric} other undefined metrics}} · Provisional
Coding
V3.0undefined metric} other undefined metrics}} · Provisional
Score62.980% interval46.9–78.91/4 Measured dimensionsSciCode59.0%#4 / 89
Coding
V3.0undefined metric} other undefined metrics}} · Provisional
Reasoning
V3.0undefined metric} other undefined metrics}} · Provisional
Score60.880% interval46.9–74.71/4 Measured dimensionsGPQA91.7%#21 / 218HLE42.3%#25 / 216
Reasoning
V3.0undefined metric} other undefined metrics}} · Provisional
Knowledge
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Knowledge
V3.0undefined metric} other undefined metrics}} · No data
Math
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Math
V3.0undefined metric} other undefined metrics}} · No data
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
Multimodal
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Multimodal
V3.0undefined metric} other undefined metrics}} · No data
Instruction following
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Instruction following
V3.0undefined metric} other undefined metrics}} · No data
OpenRouter endpoints
26 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
Friendli friendli | $1.26/M | $3.96/M | 100.0% | — | — | undefined tokens / undefined tokens |
BaseTen baseten/fp8 | $2.10/M | $6.60/M | 100.0% | — | — | undefined tokens / undefined tokens |
Fireworks fireworks | $1.40/M | $4.40/M | 99.9% | — | — | undefined tokens / undefined tokens |
AtlasCloud atlas-cloud/fp8 | $1.40/M | $4.40/M | 99.9% | — | — | undefined tokens / undefined tokens |
Sail Research sail-research/fp8 | $1.26/M | $3.95/M | 99.9% | — | — | undefined tokens / undefined tokens |
Cloudflare cloudflare | $1.40/M | $4.40/M | 99.9% | — | — | undefined tokens / undefined tokens |
Parasail parasail/fp8 | $1.40/M | $4.40/M | 99.6% | — | — | undefined tokens / undefined tokens |
Decart decart/fp4 | $1.19/M | $3.74/M | 99.6% | — | — | undefined tokens / undefined tokens |
Novita novita/fp8 | $1.09/M | $3.43/M | 99.5% | — | — | undefined tokens / undefined tokens |
AkashML akashml/fp8 | $1.17/M | $3.96/M | 99.5% | — | — | undefined tokens / undefined tokens |
Reka reka/fp8 | $0.936/M | $3.17/M | 99.3% | — | — | undefined tokens / undefined tokens |
Inceptron inceptron/fp4 | $1.05/M | $4.10/M | 99.3% | — | — | undefined tokens / undefined tokens |
GMICloud gmicloud/fp8 | $1.12/M | $3.52/M | 99.3% | — | — | undefined tokens / undefined tokens |
Venice venice | $1.40/M | $4.40/M | 99.2% | — | — | undefined tokens / undefined tokens |
Z.AI z-ai/fp8 | $1.40/M | $4.40/M | 99.2% | — | — | undefined tokens / undefined tokens |
SiliconFlow siliconflow/fp8 | $1.40/M | $4.40/M | 99.1% | — | — | undefined tokens / undefined tokens |
Morph morph/fp8 | $0.970/M | $3.31/M | 99.0% | — | — | undefined tokens / undefined tokens |
Wafer wafer | $1.19/M | $4.40/M | 98.9% | — | — | undefined tokens / undefined tokens |
Modal modal | $1.40/M | $4.40/M | 98.9% | — | — | undefined tokens / undefined tokens |
BaseTen baseten/fp4 | $1.40/M | $4.40/M | 98.8% | — | — | undefined tokens / undefined tokens |
DeepInfra deepinfra/fp4 | $1.20/M | $4/M | 98.4% | — | — | undefined tokens / undefined tokens |
Crusoe crusoe/fp4 | $1.40/M | $4.40/M | 97.4% | — | — | undefined tokens / undefined tokens |
Phala phala | $0.980/M | $3.08/M | 97.0% | — | — | undefined tokens / undefined tokens |
Makora makora/fp4 | $1.35/M | $4.40/M | 96.6% | — | — | undefined tokens / undefined tokens |
DigitalOcean digitalocean | $0.950/M | $3.40/M | 96.5% | — | — | undefined tokens / undefined tokens |
Together together | $1.40/M | $4.40/M | 96.2% | — | — | undefined tokens / undefined tokens |
Pricing Comparison
Compare GLM 5.3 API pricing across 147 providers. Prices range from $0.0080/M to $175.20/M. 10dian-API offers the lowest rate at $0.0080/M. 6 providers offer free API credits or a free tier.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 100% | glm-5.3 | sale | $4.00/M | $14.00/M | — | — | — | |
L1 100% | glm-5.3 | zai-officially | -18%$1.14/M Cache read$0.286/M | -9%$4.00/M | — | — | — | |
L1 99% L2 0% | z-ai/glm-5.3 | default | $1.40/M Cache read$0.260/M | $4.40/M | — | — | — | |
L1 100% | glm-5.3 | default | -22%$1.10/M | -13%$3.84/M | — | — | — | |
L1 100% | glm-5.3 | default | -22%$1.10/M Cache read$0.274/M | -13%$3.84/M | — | — | — | |
L1 100% | glm-5.3 | 国模折扣 | -56%$0.614/M Cache read$0.154/M | -51%$2.15/M | — | — | — | |
L1 100% L2 100% | glm-5.3 | Free | -94%$0.080/M Cache read$0.020/MCache write$0.080/MCache write 1h$0.128/M | -94%$0.280/M | — | — | — | |
L1 100% | [次]glm-5.3 | default | $0.200/request | - | — | — | — | |
L1 100% | glm-5.3 | default | -85%$0.205/M | -38%$2.74/M | — | — | — | |
L1 100% | glm-5.3 | Self-Deployed-2 | -85%$0.208/M Cache read$0.039/M | -85%$0.652/M | — | — | — | |
L1 100% | glm-5.3 | deepseek | -69%$0.438/M | -65%$1.53/M | — | — | — | |
L1 100% | glm-5.3 | opencode | $3.20/M Cache read$0.800/M | $11.20/M | — | — | — | |
L1 100% | glm-5.3 | default | $3.40/M Cache read$0.680/M | $15.30/M | — | — | — | |
L1 100% | glm-5.3 | glm | $5.60/M Cache read$1.40/M | $19.60/M | — | — | — | |
L1 100% | [xj]q|abyss/GLM-5.3 | default | $8.00/request | - | — | — | — | |
L1 100% | [mt]q|次/glm-5.3 | default | $11.25/request | - | — | — | — | |
L1 100% | [hm]q|满血/glm-5.3 | default | $25.00/request | - | — | — | — | |
L1 100% | glm-5.3 | default | -93%$0.096/M Cache read$0.018/M | -93%$0.301/M | — | — | — | |
L1 100% | glm-5.3-free | default | -98%$0.030/M Cache read$0/M | -100%$0/M | — | — | — | |
L1 100% | glm-5.3 | default | -40%$0.840/M Cache read$0.156/M | -40%$2.64/M | — | — | — |
Alternatives & Similar Models
GLM-5.2
glm-5-2
Zhipu GLM-5.2 is Zhipu latest flagship coding and agentic model with a 1M-token context window, enhanced reasoning modes, and long-horizon software engineering capabilities.
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
DeepSeek V4 Pro
deepseek-v4-pro
DeepSeek V4 Pro is the professional-tier DeepSeek V4 model, targeting frontier reasoning, coding, and agent workflows with maximum capability.
GLM-5.1
glm-5-1
Zhipu GLM-5.1 is a next-generation GLM model aimed at frontier reasoning, coding, and bilingual agent applications.
GPT-5.4
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
Claude Opus 4.6
claude-opus-4-6
Anthropic Claude Opus 4.6 is the most capable Claude Opus tier, optimized for complex analysis, long-horizon coding, and high-stakes enterprise reasoning workloads.
Frequently Asked Questions
- What benchmark data does GLM 5.3 include?
- LMSpeed shows GLM 5.3 benchmark context, API price, output speed, first-token latency, and provider data across 153 providers when those signals are available.
- What is the GLM 5.3 API price?
- GLM 5.3 has pricing from undefined provider} other undefined providers}}, ranging from $0.0080/M to $175.20/M. 10dian-API has the lowest listed price.
- What does the GLM 5.3 API pricing table include?
- The GLM 5.3 API pricing table compares 153 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest GLM 5.3 API pricing?
- 10dian-API currently has the lowest listed GLM 5.3 price at $0.0080/M across undefined provider} other undefined providers}}.
- Can I compare GLM 5.3 API price and speed together?
- Yes. LMSpeed shows GLM 5.3 API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
- Is GLM 5.3 API free?
- Yes, GLM 5.3 free API options are available through 6 providerundefined other undefined} on LMSpeed, including Zero API, 初叶🍂Furry API, 兔子API, 兔子API, 兔子API. These providers offer free API credits or a free tier with no per-token charges.
- Where can I get GLM 5.3 free API access?
- LMSpeed currently lists 6 free API providerundefined other undefined} for GLM 5.3: Zero API, 初叶🍂Furry API, 兔子API, 兔子API, 兔子API. Check each provider row before using it because free tier limits can change.
