GLM-4.7 Flash API Benchmarks, Pricing & Provider Data
Compare GLM-4.7 Flash with another model
Choose a model to open its comparison page.
GLM-4.7 Flash benchmark, API pricing, and provider data cover 60 API providers, with prices starting at $0.0001/request. GLM-4.7 Flash free API options are available from 4 providers. The page also shows measured API speed and first-token latency.
Zhipu AI GLM-4.7 Flash is a speed-focused GLM variant for real-time chat, function calling, and bilingual enterprise copilots with competitive token economics.
Category Performance
Observed-capability estimates with uncertainty reported separately from independent benchmark families.
- Coverage
- 2 / 8
- Methodology
- V3.0
#1Coding41.5Provisional1/4 Measured dimensions
#2Reasoning40.4Provisional1/4 Measured dimensions
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Technical Details
- Input
- Output
- Released
- Jan 2026
- Tokenizer
- Other
- Architecture
- text->text
- Moderated
- No
- Supported parameters
- frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
Rankings
Excels at
Detailed scores
Updated: Sep 1, 2026Pricing
undefined metric} other undefined metrics}}
Input price$0.070/M#6 / 181Output price$0.400/M#15 / 181
Pricing
undefined metric} other undefined metrics}}
Agents
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Agents
V3.0undefined metric} other undefined metrics}} · No data
Coding
V3.0undefined metric} other undefined metrics}} · Provisional
Score41.580% interval25.5–57.51/4 Measured dimensionsSciCode33.7%#137 / 206
Coding
V3.0undefined metric} other undefined metrics}} · Provisional
Reasoning
V3.0undefined metric} other undefined metrics}} · Provisional
Score40.480% interval26.5–54.31/4 Measured dimensionsGPQA58.1%#171 / 213HLE7.6%#130 / 210
Reasoning
V3.0undefined metric} other undefined metrics}} · Provisional
Knowledge
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Knowledge
V3.0undefined metric} other undefined metrics}} · No data
Math
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Math
V3.0undefined metric} other undefined metrics}} · No data
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
Multimodal
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Multimodal
V3.0undefined metric} other undefined metrics}} · No data
Instruction following
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Instruction following
V3.0undefined metric} other undefined metrics}} · No data
OpenRouter endpoints
4 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
DeepInfra deepinfra/bf16 | $0.060/M | $0.400/M | 99.7% | — | — | undefined tokens / undefined tokens |
Venice venice/fp8 | $0.060/M | $0.400/M | 99.5% | — | — | undefined tokens / undefined tokens |
Cloudflare cloudflare | $0.061/M | $0.400/M | 95.9% | — | — | undefined tokens / undefined tokens |
Novita novita/bf16 | $0.070/M | $0.400/M | 81.5% | — | — | undefined tokens / undefined tokens |
Pricing Comparison
Compare GLM-4.7 Flash API pricing across 56 providers. Prices range from $0.0001/request to $535.71/M. CM-API 公益站 offers the lowest rate at $0.0001/request. 4 providers offer free API credits or a free tier.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
Moyanjdc API Free | L1 100% | glm-4.7-flash | default | Free | Free | 53.0 t/s | 20.54 s | — |
L1 100% | glm-4.7-flash | default | $1.46/M Audio in$0/M | -27%$0.292/M | 40.6 t/s | 21.30 s | — | |
L1 100% | glm-4.7-flash-free | free | Free | Free | — | — | — | |
L1 100% L2 62% | glm-4.7-flash | default | $0.090/M | $0.530/M | — | — | — | |
L1 100% | glm-4.7-flash | test | $10.27/M | $10.27/M | — | — | — | |
L1 100% | glm-4.7-flash | default | $0.080/M Cache read$0.016/M | $0.600/M | — | — | — | |
L1 100% | z-ai/glm-4.7-flash-free | default | $375.00/M | $375.00/M | — | — | — | |
L1 100% | glm-4.7-flash | default | -91%$0.0062/M Cache read$0.0012/M | -91%$0.036/M | — | — | — | |
L1 100% | glm-4.7-flash | 自动令牌兜底分组 | $0.0001/request | - | — | — | — | |
L1 100% L2 0% | glm-4.7-flash | default | $0.500/M | $3.00/M | — | — | — | |
L1 99% | glm-4.7-flash | default | $0.098/M | -2%$0.392/M | — | — | — | |
L1 100% | glm-4.7-flash | default | $0.010/request | - | — | — | — | |
L1 99% | glm-4.7-flash | 91vip | $2.00/M | $6.00/M | — | — | — | |
L1 99% | GLM-4.7-Flash | default | $535.71/M | $535.71/M | — | — | — | |
L1 98% | zhipu/glm-4.7-flash | 0倍倍率分组 | -80%$0.014/M | -97%$0.014/M | — | — | — | |
L1 12% | glm-4.7-flash | default | $399.60/M Cache read$79.92/M | $2997.00/M | — | — | — | |
L1 100% | glm-4.7-flash | default | $0.010/request | - | — | — | — | |
L1 100% | glm-4.7-flash | glm | $0.050/request | - | — | — | — | |
L1 100% | glm-4.7-flash | 聊天专属喵 | -53%$0.033/M Cache read$0.0066/M | -38%$0.247/M | — | — | — | |
L1 99% | GLM-4.7-Flash | default | $32.00/M | $32.00/M | — | — | — |
Alternatives & Similar Models
GLM-5
glm-5
Zhipu GLM-5 is Zhipu flagship GLM series model with enhanced reasoning, agent capabilities, and strong performance on Chinese enterprise and coding scenarios.
GLM-5.1
glm-5-1
Zhipu GLM-5.1 is a next-generation GLM model aimed at frontier reasoning, coding, and bilingual agent applications.
Kimi K2.5
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.
GLM-4.7
glm-4-7
Zhipu GLM-4.7 is a flagship GLM release from Zhipu AI with advanced Chinese-English reasoning, coding, and agent features.
MiniMax M2.7
minimax-m2-7
MiniMax M2.7 is a high-tier M2-series model tuned for complex reasoning, long-context dialogue, and production-grade API workloads.
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
Frequently Asked Questions
- What benchmark data does GLM-4.7 Flash include?
- LMSpeed shows GLM-4.7 Flash benchmark context, API price, output speed, first-token latency, and provider data across 60 providers when those signals are available.
- What is the GLM-4.7 Flash API price?
- GLM-4.7 Flash has pricing from 60 providers, ranging from $0.0001/request to $535.71/M. CM-API 公益站 has the lowest listed price.
- What does the GLM-4.7 Flash API pricing table include?
- The GLM-4.7 Flash API pricing table compares 60 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest GLM-4.7 Flash API pricing?
- CM-API 公益站 currently has the lowest listed GLM-4.7 Flash price at $0.0001/request across 60 providers.
- Can I compare GLM-4.7 Flash API price and speed together?
- Yes. LMSpeed shows GLM-4.7 Flash API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
- Is GLM-4.7 Flash API free?
- Yes, GLM-4.7 Flash free API options are available through 4 providerundefined other undefined} on LMSpeed, including Moyanjdc API, Zero API, VSLLM, Dext API. These providers offer free API credits or a free tier with no per-token charges.
- Where can I get GLM-4.7 Flash free API access?
- LMSpeed currently lists 4 free API providerundefined other undefined} for GLM-4.7 Flash: Moyanjdc API, Zero API, VSLLM, Dext API. Check each provider row before using it because free tier limits can change.
