Choose a model to open its comparison page.
WizardLM-2 8x22B API pricing covers 3 API providers, from $29.37/M to $76.36/M.
Microsoft WizardLM-2 8x22B is a 141B-parameter MoE instruct model built on Mixtral 8x22B, tuned for complex chat, multilingual reasoning, and agent-style tasks.
Input and output token limits for this model, plus how it ranks on long-context understanding.
Third-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
Novita novita/bf16 | $0.620/M | $0.620/M | 100.0% | — | — | 65.5K tokens / 8K tokens |
Compare WizardLM-2 8x22B API pricing across 3 providers. Prices range from $29.37/M to $76.36/M. 钱多多 API offers the lowest rate at $29.37/M.
gemini-2-5-pro
Google Gemini 2.5 Pro is Google advanced multimodal model with a 1M-token context window, strong STEM reasoning, and native support for images, audio, and video understanding.
gemini-2-5-flash
Google Gemini 2.5 Flash is a fast multimodal model balancing speed and intelligence for chat, tool use, and large-context workloads at lower cost than Pro tiers.
gpt-5
OpenAI GPT-5 is OpenAI frontier general-purpose model with improved reasoning depth, coding reliability, and multimodal understanding for production assistants and agent workflows.
deepseek-r1
DeepSeek R1 is a reasoning-focused language model in the DeepSeek series, designed for complex reasoning, problem-solving, and analytical tasks.
glm-4-7
Zhipu GLM-4.7 is a flagship GLM release from Zhipu AI with advanced Chinese-English reasoning, coding, and agent features.
gpt-oss
GPT-OSS is an open-weight language model family designed for self-hosted inference, research, and cost-efficient alternatives to proprietary GPT-class models.