Choose a model to open its comparison page.
GLM-Z1 Flash API pricing covers 19 API providers, from $0.0040/M to $75.00/M. GLM-Z1 Flash free API options are available from 4 providers. The page also shows measured API speed and first-token latency.
Zhipu AI GLM-Z1 Flash is a fast reasoning model in the GLM-Z1 family for STEM tutoring, step-by-step logic, and lightweight agent loops.
Compare GLM-Z1 Flash API pricing across 15 providers. Prices range from $0.0040/M to $75.00/M. 小天公益站 offers the lowest rate at $0.0040/M. 4 providers offer free API credits or a free tier.
| Провайдер | Работоспособность | Вариант модели | Группа | Входные данные ($/M) | Выходные данные ($/M) | Скорость (т/с) | Первый токен | Аудит |
|---|---|---|---|---|---|---|---|---|
6655 翻译小站 Бесплатно | L1 100% | glm-z1-flash | default | Бесплатно | Бесплатно | — | — | — |
L1 100% | glm-z1-flash | default | $0.0040/M | $0.0004/M | — | — | — | |
L1 100% | glm-z1-flash | default | $20.55/M | $20.55/M | — | — | — | |
IXIOCCAPI Бесплатно | L1 100% | glm-z1-flash | default | Бесплатно | Бесплатно | — | — | — |
L1 100% | glm-z1-flash | default | $0.098/M | $0.098/M | — | — | — | |
L1 100% | glm-z1-flash | default | $15.41/M | $15.41/M | — | — | — | |
L1 100% | glm-z1-flash | default | $75.00/M | $75.00/M | — | — | — | |
L1 99% | glm-z1-flash | default | $10.27/M | $10.27/M | — | — | — | |
L1 99% | glm-z1-flash | default | $0.150/M | $0.150/M | — | — | — | |
L1 98% | glm-z1-flash | default | $0.010/request | - | — | — | — | |
L1 98% | GLM-Z1-Flash | default | $0.010/request | - | — | — | — | |
L1 95% | glm-z1-flash | default | $0.014/M | $0.014/M | — | — | — | |
L1 100% | glm-z1-flash | 测试 | $75.00/M | $75.00/M | — | — | — | |
L1 0% | glm-z1-flash | default | $10.27/M | $10.27/M | — | — | — |
qwen3
Alibaba Qwen3 is the Qwen family's flagship LLM series with dense and MoE variants, seamless thinking/non-thinking modes, and leading open-source performance in math, code, and agent tasks.
glm-4-7
Zhipu GLM-4.7 is a flagship GLM release from Zhipu AI with advanced Chinese-English reasoning, coding, and agent features.
glm-4-5-air
Zhipu AI GLM-4.5 Air is a lightweight GLM-4.5 variant optimized for low-latency Chinese and English dialogue, retrieval-augmented apps, and edge deployments.
glm-4-5
Zhipu AI GLM-4.5 is a 355B-parameter MoE agent foundation model that unifies reasoning, coding, and tool use with hybrid thinking modes and a 128K context window.
glm-4-5-flash
Zhipu AI GLM-4.5 Flash is a speed-first GLM-4.5 model for interactive chat, function calling, and bilingual customer-service automation.
minimax-m2-5
MiniMax M2.5 is MiniMax's flagship text model for coding and agents, with SOTA-level programming and agentic performance, improved token efficiency, and fast high-TPS API deployment.