Choose a model to open its comparison page.
GLM-4.1v Thinking FlashX API pricing covers 15 API providers, from $0.098/M to $150.00/M.
Zhipu AI GLM-4.1v Thinking FlashX is a reasoning model in the GLM series, designed for complex reasoning, problem-solving, and analytical tasks.
Compare GLM-4.1v Thinking FlashX API pricing across 15 providers. Prices range from $0.098/M to $150.00/M. 钱多多 API offers the lowest rate at $0.098/M.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 100% | glm-4.1v-thinking-flashx | default | $0.098/M | $0.098/M | — | — | — | |
L1 100% | glm-4.1v-thinking-flashx | default | $0.274/M | $0.274/M | — | — | — | |
L1 100% | glm-4.1v-thinking-flashx | default | $0.560/M | $0.560/M | — | — | — | |
L1 100% | glm-4.1v-thinking-flashx | default | $20.55/M | $20.55/M | — | — | — | |
L1 100% | glm-4.1v-thinking-flashx | default | $15.41/M | $15.41/M | — | — | — | |
L1 100% | glm-4.1v-thinking-flashx | default | $10.27/M | $10.27/M | — | — | — | |
L1 99% | GLM-4.1V-Thinking-FlashX | default | $75.00/M | $75.00/M | — | — | — | |
L1 1% | glm-4.1v-thinking-flashx | default | $10.27/M | $10.27/M | — | — | — |
glm-4-7
Zhipu GLM-4.7 is a flagship GLM release from Zhipu AI with advanced Chinese-English reasoning, coding, and agent features.
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
gpt-5-3-codex
OpenAI GPT-5.3 Codex is a code-specialized variant in the GPT-5 series, optimized for code generation, debugging, and software development tasks.
gpt-5-2
OpenAI GPT-5.2 is a GPT-5 series model emphasizing advanced reasoning, multimodal understanding, and high-quality outputs for complex enterprise workloads.
glm-5
Zhipu GLM-5 is Zhipu flagship GLM series model with enhanced reasoning, agent capabilities, and strong performance on Chinese enterprise and coding scenarios.
gemini-2-5-flash
Google Gemini 2.5 Flash is a fast multimodal model balancing speed and intelligence for chat, tool use, and large-context workloads at lower cost than Pro tiers.