Sponsored byFusecodeEnterprise coding API for Claude Code, Codex, and model workflows.
LogoLMSpeed
  • Free
  • Models
  • Providers
  • Leaderboard
LogoLMSpeed
  1. Home
  2. Models
LogoLMSpeed

The best API speed test tool

GitHubGitHubTwitterX (Twitter)Email
Product
  • Features
  • Pricing
  • FAQ
Leaderboard
  • Overview
  • Speed Ranking
  • Latency Ranking
  • Health Ranking
  • Model Pricing
  • Model Speed
  • Reasoning
  • Coding
Models
  • All Models
  • GPT
  • Claude
  • Gemini
  • DeepSeek
  • Llama
  • Qwen
Free Models
  • All Free Models
  • Free GPT
  • Free Claude
  • Free Gemini
  • Free DeepSeek
  • Free Llama
  • Free Qwen
Tools
  • Speed Test
  • Provider Audit
Company
  • About
Resources
  • Provider Directory
  • Documentation
  • Public API
  • Botab
  • VidBee
Legal
  • Cookie Policy
  • Privacy Policy
  • Terms of Service
© 2026 LMSpeed All Rights Reserved.Made by Nexmoe with ❤️
All85Text80Embeddings2Speech2Transcription1

AI Model Benchmark Directory

LMSpeed is an AI model directory for comparing API price, output speed, first-token latency, provider coverage, and benchmark data. Use it to narrow your model and provider choice, then test the endpoint that fits your workload.

Price, speed, latency, and availability can change. Treat this table as a current signal and verify your own endpoint before you deploy.

Top for agents
  1. 1MoonshotAIKimi K368.9±7.3
  2. 2OpenAIGPT-5.6 Sol68.2±7.3
  3. 3ClaudeClaude Opus 567.5±8.2
Top for reasoning
Top for coding
Highest throughput
Category leaderboards
Release date
Context1MInput$2.00/M
  • 1
  • 2

How to read category scores

Agents, Coding, Reasoning, and the other capability columns are 0–100 observed-capability estimates relative to the eligible model population in a dated Category Score V3 run.

They are not success rates, IQ scores, or an average across all eight categories. Read them with the 80% interval, measured dimensions, benchmark families, and evidence shown on each model page.

What this directory shows

Each row brings together the information you need to compare an LLM API. Some fields are blank when LMSpeed has no current data for that model or provider.

API price
Compare input and output price per million tokens when it is available.
Speed and latency
Use throughput and first-token latency to compare response behavior.
Provider coverage
Open a model to review its listed providers and their current details.
Capability data
Use the capability columns when current model scores are available.

How to use the model directory

Start with the decision that matters most for your workload, then compare the current rows before you test an endpoint.

  1. Find a model.Search by model name, slug, or description.
  2. Sort the key metric.Sort by price, throughput, latency, provider count, or capability data.
  3. Compare providers.Open a model page to review the provider options and current data.
  4. Test your endpoint.Run a speed test before you use an endpoint in production.

Frequently Asked Questions

How does LMSpeed benchmark AI models?

LMSpeed runs standardized five-round API speed tests on each model, measuring output throughput (tokens per second), first-token latency, and total response time across multiple providers.

Which AI model has the lowest API latency?

Latency varies by provider and model. Use the LMSpeed model directory to sort by latency and find the model with the fastest first-token response time. Check the latency leaderboard for monthly rankings.

How to compare LLM API pricing across providers?

LMSpeed lists input and output token prices per million for each model across available providers. Sort by price to find a lower-cost option, or filter by provider to compare rates side by side.

Output
$6.00/M
Providers
+12
65.8±11.3E
67.3±11.2E
65.9±11.3E
51.6±16.1P
—
—
65.9±6.3
57.8±16.0P
Throughput
—
Latency
—
Release date2026-08-03
  • 65.8
  • 67.3
  • 65.9
  • 51.6
  • —
  • —
  • 65.9
  • 57.8
QwenQwen3.7 FlashQwenContext1MInput—Output—Providers
+1
—
—
—
—
—
—
—
—
Throughput
—
Latency
—
Release date2026-07-27
  • —
  • —
  • —
  • —
  • —
  • —
  • —
  • —
QwenQwen-Audio-3.0-TTS PlusQwenContext0Input—Output—Providers—
—
—
—
—
—
—
—
—
Throughput
—
Latency
—
Release date2026-07-23
  • —
  • —
  • —
  • —
  • —
  • —
  • —
  • —
QwenQwen-Audio-3.0-TTS FlashQwenContext0Input—Output—Providers—
—
—
—
—
—
—
—
—
Throughput
—
Latency
—
Release date2026-07-23
  • —
  • —
  • —
  • —
  • —
  • —
  • —
  • —
Nex n2 ProNex AGIContext262.1KInput$0.500/MOutput$2.50/MProviders
+7
—
54.3±16.0P
59.2±13.9P
—
—
—
—
—
Throughput
83 t/s
Latency
1.16s
Release date2026-06-08
  • —
  • 54.3
  • 59.2
  • —
  • —
  • —
  • —
  • —
QwenQwen3.7 PlusQwenContext1MInput$0.400/MOutput$1.60/MProviders
+48
54±5.7
53.2±6.0
58.1±8.7
48.6±12.2E
57.2±12.3E
51.1±11.7E
55.8±6.3
56.9±11.9E
Throughput
46 t/s
Latency
29.34s
Release date2026-06-03
  • 54
  • 53.2
  • 58.1
  • 48.6
  • 57.2
  • 51.1
  • 55.8
  • 56.9
QwenQwen3.7 MaxQwenContext1MInput$2.50/MOutput$7.50/MProviders
+48
56.9±5.4
57±6.0
60±8.7
55.8±12.2E
63.9±12.3E
58.1±11.7E
—
56.8±11.9E
Throughput
68 t/s
Latency
17.93s
Release date2026-05-21
  • 56.9
  • 57
  • 60
  • 55.8
  • 63.9
  • 58.1
  • —
  • 56.8
QwenQwen3 ASR FlashQwenContext0Input—Output—Providers
—
—
—
—
—
—
—
—
Throughput
—
Latency
—
Release date2026-05-14
  • —
  • —
  • —
  • —
  • —
  • —
  • —
  • —
QwenQwen3.5 Plus 2026-04-20QwenContext1MInput—Output—Providers—
—
—
—
—
—
—
—
—
Throughput
—
Latency
—
Release date2026-04-27
  • —
  • —
  • —
  • —
  • —
  • —
  • —
  • —
QwenQwen3.5 PlusAlibabaContext1MInput—Output—Providers
+112
44.7±16.2P
46.2±16.1P
—
—
48.8±16.1P
—
—
—
Throughput
51 t/s
Latency
13.81s
Release date2026-04-27
  • 44.7
  • 46.2
  • —
  • —
  • 48.8
  • —
  • —
  • —
QwenQwen3.6 FlashQwenContext1MInput—Output—Providers
+16
—
—
—
—
—
—
—
—
Throughput
106 t/s
Latency
8.21s
Release date2026-04-27
  • —
  • —
  • —
  • —
  • —
  • —
  • —
  • —
QwenQwen3.6 35B A3BQwenContext262.1KInput$0.248/MOutput$1.49/MProviders
+5
46.8±5.9
38±6.0
53±8.7
39.1±12.2E
35.1±11.5E
—
43.7±7.0
50.7±16.0P
Throughput
—
Latency
—
Release date2026-04-27
  • 46.8
  • 38
  • 53
  • 39.1
  • 35.1
  • —
  • 43.7
  • 50.7
QwenQwen3.6 Max PreviewQwenContext262.1KInput$1.30/MOutput$7.80/MProviders
+15
54±11.8E
52.2±8.9
56.6±10.8E
55.1±16.3P
50.6±16.1P
—
—
55±16.0P
Throughput
—
Latency
—
Release date2026-04-27
  • 54
  • 52.2
  • 56.6
  • 55.1
  • 50.6
  • —
  • —
  • 55
QwenQwen3.6 27BQwenContext262.1KInput$0.600/MOutput$3.60/MProviders
+3
56.1±6.6
47.2±6.0
54.9±8.7
37.3±11.6E
41.6±11.5E
—
46.1±6.5
51.8±16.0P
Throughput
—
Latency
—
Release date2026-04-27
  • 56.1
  • 47.2
  • 54.9
  • 37.3
  • 41.6
  • —
  • 46.1
  • 51.8
QwenQwen3.6 PlusQwenContext1MInput$0.500/MOutput$3.00/MProviders
+121
50±5.5
52.3±8.2
57.4±8.1
52±11.6E
52.4±9.1E
52.2±12.2E
50.9±6.5
55.7±11.9E
Throughput
47 t/s
Latency
24.53s
Release date2026-04-02
  • 50
  • 52.3
  • 57.4
  • 52
  • 52.4
  • 52.2
  • 50.9
  • 55.7
QwenQwen3.5-9BQwenContext262.1KInput$0.170/MOutput$0.250/MProviders—
—
43±16.0P
52.4±13.9P
—
—
—
—
—
Throughput
—
Latency
—
Release date2026-03-10
  • —
  • 43
  • 52.4
  • —
  • —
  • —
  • —
  • —
QwenQwen3.5Context262.1KInput$0.030/MOutput$0.150/MProviders
+98
45.4±5.3
39.9±9.0
52.6±8.1
53.1±11.6E
42±11.5E
54.6±12.2E
50.4±6.5
53.7±11.9E
Throughput
46 t/s
Latency
10.19s
Release date2026-03-10
  • 45.4
  • 39.9
  • 52.6
  • 53.1
  • 42
  • 54.6
  • 50.4
  • 53.7
QwenQwen3.5-35B-A3BQwenContext262.1KInput$0.250/MOutput$2.00/MProviders
43.1±8.7
42.3±11.2E
51.6±8.3
38±16.3P
—
41.6±16.4P
45.5±12.1E
51.5±11.9E
Throughput
—
Latency
—
Release date2026-02-25
  • 43.1
  • 42.3
  • 51.6
  • 38
  • —
  • 41.6
  • 45.5
  • 51.5
QwenQwen3.5-27BQwenContext262.1KInput$0.300/MOutput$2.40/MProviders
46.2±8.7
49.3±11.2E
54±8.3
41.4±16.3P
—
45.4±16.4P
49.1±12.1E
57.2±11.9E
Throughput
—
Latency
—
Release date2026-02-25
  • 46.2
  • 49.3
  • 54
  • 41.4
  • —
  • 45.4
  • 49.1
  • 57.2
QwenQwen3.5-122B-A10BQwenContext262.1KInput$0.400/MOutput$3.20/MProviders
48.5±10.6E
49.2±11.8E
53.7±8.3
43.7±16.3P
—
45.4±16.4P
47.3±9.4E
54.3±11.9E
Throughput
—
Latency
—
Release date2026-02-25
  • 48.5
  • 49.2
  • 53.7
  • 43.7
  • —
  • 45.4
  • 47.3
  • 54.3
QwenQwen3.5-FlashQwenContext1MInput—Output—Providers—
—
—
—
—
—
—
—
—
Throughput
—
Latency
—
Release date2026-02-25
  • —
  • —
  • —
  • —
  • —
  • —
  • —
  • —
QwenQwen3.5 FlashContext1MInput—Output—Providers
+70
—
—
—
—
42.7±16.1P
—
—
—
Throughput
115 t/s
Latency
7.92s
Release date2026-02-25
  • —
  • —
  • —
  • —
  • 42.7
  • —
  • —
  • —
QwenQwen3.5 Plus 2026-02-15QwenContext1MInput—Output—Providers—
—
—
—
—
—
—
—
—
Throughput
—
Latency
—
Release date2026-02-16
  • —
  • —
  • —
  • —
  • —
  • —
  • —
  • —
QwenQwen3.5 397B A17BQwenContext262.1KInput$0.600/MOutput$3.60/MProviders
—
54.4±16.0P
58.5±13.9P
—
—
—
—
—
Throughput
—
Latency
—
Release date2026-02-16
  • —
  • 54.4
  • 58.5
  • —
  • —
  • —
  • —
  • —
QwenQwen3 Max ThinkingQwenContext262.1KInput$1.20/MOutput$6.00/MProviders
+5
—
52.5±13.9P
55.4±11.1E
—
51.7±16.1P
—
—
—
Throughput
—
Latency
—
Release date2026-02-09
  • —
  • 52.5
  • 55.4
  • —
  • 51.7
  • —
  • —
  • —
QwenQwen3 Coder NextQwenContext262.1KInput$0.350/MOutput$1.20/MProviders
+17
—
47±16.0P
49.1±13.9P
—
—
—
—
—
Throughput
147 t/s
Latency
0.97s
Release date2026-02-04
  • —
  • 47
  • 49.1
  • —
  • —
  • —
  • —
  • —
FreeOpenrouterContext200KInput—Output—Providers
+15
—
—
—
—
—
—
—
—
Throughput
24 t/s
Latency
25.61s
Release date2026-02-01
  • —
  • —
  • —
  • —
  • —
  • —
  • —
  • —
DeepSeekDeepSeek V3.2DeepSeekContext163.8KInput$0.280/MOutput$0.420/MProviders
+176
39.2±6.9
51.3±8.6
51.2±8.7
40.5±16.0P
49±16.1P
—
—
46.1±16.0P
Throughput
46 t/s
Latency
5.36s
Release date2025-12-01
  • 39.2
  • 51.3
  • 51.2
  • 40.5
  • 49
  • —
  • —
  • 46.1
QwenQwen3 Embedding 8BQwenContext32.8KInput—Output—Providers
—
—
—
—
—
—
—
—
Throughput
—
Latency
—
Release date2025-10-28
  • —
  • —
  • —
  • —
  • —
  • —
  • —
  • —
QwenQwen3 Embedding 4BQwenContext32.8KInput—Output—Providers
—
—
—
—
—
—
—
—
Throughput
—
Latency
—
Release date2025-10-28
  • —
  • —
  • —
  • —
  • —
  • —
  • —
  • —
QwenQwen3 VL 32B InstructQwenContext131.1KInput$0.700/MOutput$2.80/MProviders
+2
—
45.5±13.9P
48.5±11.1E
—
49.1±16.1P
—
—
—
Throughput
—
Latency
—
Release date2025-10-23
  • —
  • 45.5
  • 48.5
  • —
  • 49.1
  • —
  • —
  • —
QwenQwen3 VL 8B ThinkingQwenContext131.1KInput—Output—Providers
—
—
—
—
—
—
—
—
Throughput
—
Latency
—
Release date2025-10-14
  • —
  • —
  • —
  • —
  • —
  • —
  • —
  • —
QwenQwen3 VL 8B InstructQwenContext262.1KInput$0.180/MOutput$0.700/MProviders
+2
—
34.7±13.9P
40.6±11.1E
—
41.5±16.1P
—
—
—
Throughput
—
Latency
—
Release date2025-10-14
  • —
  • 34.7
  • 40.6
  • —
  • 41.5
  • —
  • —
  • —
QwenQwen3 VL 30B A3B ThinkingQwenContext262.1KInput—Output—Providers
+2
—
—
—
—
—
—
—
—
Throughput
—
Latency
—
Release date2025-10-06
  • —
  • —
  • —
  • —
  • —
  • —
  • —
  • —
QwenQwen3 VL 30B A3B InstructQwenContext262.1KInput$0.200/MOutput$0.800/MProviders
+1
—
45.4±13.9P
47.5±11.1E
—
49.8±16.1P
—
—
—
Throughput
—
Latency
—
Release date2025-10-06
  • —
  • 45.4
  • 47.5
  • —
  • 49.8
  • —
  • —
  • —
QwenQwen3 VL 235B A22B ThinkingQwenContext131.1KInput—Output—Providers
+5
—
—
—
—
—
—
—
—
Throughput
—
Latency
—
Release date2025-09-23
  • —
  • —
  • —
  • —
  • —
  • —
  • —
  • —
QwenQwen3 VL 235B A22B InstructQwenContext262.1KInput$0.700/MOutput$2.80/MProviders
+5
—
49.7±13.9P
50.3±11.1E
—
49.5±16.1P
—
—
—
Throughput
—
Latency
—
Release date2025-09-23
  • —
  • 49.7
  • 50.3
  • —
  • 49.5
  • —
  • —
  • —
QwenQwen3 MaxQwenContext262.1KInput$1.20/MOutput$6.00/MProviders
+91
47.6±11.8E
45.1±11.2E
49.5±8.7
41.5±16.0P
51.4±16.1P
—
—
44.7±16.0P
Throughput
38 t/s
Latency
7.37s
Release date2025-09-23
  • 47.6
  • 45.1
  • 49.5
  • 41.5
  • 51.4
  • —
  • —
  • 44.7
QwenQwen3 Coder PlusQwenContext1MInput—Output—Providers
+71
—
—
—
—
—
—
—
—
Throughput
45 t/s
Latency
1.35s
Release date2025-09-23
  • —
  • —
  • —
  • —
  • —
  • —
  • —
  • —
QwenQwen3 Coder FlashQwenContext1MInput—Output—Providers
+27
—
—
—
—
—
—
—
—
Throughput
—
Latency
—
Release date2025-09-17
  • —
  • —
  • —
  • —
  • —
  • —
  • —
  • —
QwenQwen3 Next 80B A3B ThinkingQwenContext262.1KInput—Output—Providers
+5
—
—
—
—
—
—
—
—
Throughput
—
Latency
—
Release date2025-09-11
  • —
  • —
  • —
  • —
  • —
  • —
  • —
  • —
QwenQwen3 Next 80B A3B InstructQwenContext262.1KInput$0.500/MOutput$2.00/MProviders
+17
—
48.2±13.9P
50.7±11.1E
—
48.7±16.1P
—
—
—
Throughput
—
Latency
—
Release date2025-09-11
  • —
  • 48.2
  • 50.7
  • —
  • 48.7
  • —
  • —
  • —
QwenQwen Plus 0728QwenContext1MInput—Output—Providers—
—
—
—
—
—
—
—
—
Throughput
—
Latency
—
Release date2025-09-08
  • —
  • —
  • —
  • —
  • —
  • —
  • —
  • —
QwenQwen Plus 0728 (thinking)QwenContext1MInput—Output—Providers
—
—
—
—
—
—
—
—
Throughput
—
Latency
—
Release date2025-09-08
  • —
  • —
  • —
  • —
  • —
  • —
  • —
  • —
QwenQwen PlusQwenContext1MInput—Output—Providers
+32
—
—
—
—
—
—
—
—
Throughput
39 t/s
Latency
0.83s
Release date2025-09-08
  • —
  • —
  • —
  • —
  • —
  • —
  • —
  • —
QwenQwen3 30B A3B Thinking 2507QwenContext81.9KInput—Output—Providers
+4
—
—
—
—
—
—
—
—
Throughput
—
Latency
—
Release date2025-08-28
  • —
  • —
  • —
  • —
  • —
  • —
  • —
  • —
QwenQwen3 Coder 30B A3B InstructQwenContext262.1KInput$0.450/MOutput$2.25/MProviders
—
42.8±13.9P
42.7±11.1E
—
50.5±11.8E
—
—
—
Throughput
—
Latency
—
Release date2025-07-31
  • —
  • 42.8
  • 42.7
  • —
  • 50.5
  • —
  • —
  • —
QwenQwen3 30B A3B Instruct 2507QwenContext262.1KInput—Output—Providers
+3
—
—
—
—
—
—
—
—
Throughput
—
Latency
—
Release date2025-07-29
  • —
  • —
  • —
  • —
  • —
  • —
  • —
  • —
QwenQwen3 235B A22B Thinking 2507QwenContext262.1KInput—Output—Providers
+3
—
—
—
—
—
—
—
—
Throughput
—
Latency
—
Release date2025-07-25
  • —
  • —
  • —
  • —
  • —
  • —
  • —
  • —
QwenQwen3 CoderQwenContext262.1KInput—Output—Providers
+59
—
—
—
—
—
—
—
—
Throughput
243 t/s
Latency
5.28s
Release date2025-07-23
  • —
  • —
  • —
  • —
  • —
  • —
  • —
  • —
QwenQwen3 235B A22B Instruct 2507QwenContext262.1KInput$0.700/MOutput$8.40/MProviders—
—
56±13.9P
53.9±11.1E
—
63±11.8E
—
—
—
Throughput
—
Latency
—
Release date2025-07-21
  • —
  • 56
  • 53.9
  • —
  • 63
  • —
  • —
  • —
QwenQwen3Context262.1KInput$0.200/MOutput$0.800/MProviders
+123
—
45.7±13.9P
47.8±11.1E
36.9±16.3P
58.4±11.8E
36.7±16.4P
—
—
Throughput
100 t/s
Latency
12.55s
Release date2025-07-21
  • —
  • 45.7
  • 47.8
  • 36.9
  • 58.4
  • 36.7
  • —
  • —
Coder LargeArcee AIContext32.8KInput—Output—Providers
—
—
—
—
—
—
—
—
Throughput
—
Latency
—
Release date2025-05-05
  • —
  • —
  • —
  • —
  • —
  • —
  • —
  • —
QwenQwen3 30B A3BQwenContext131.1KInput$0.200/MOutput$2.40/MProviders
—
40.8±13.9P
43.1±11.1E
—
49.3±11.8E
—
—
—
Throughput
—
Latency
—
Release date2025-04-28
  • —
  • 40.8
  • 43.1
  • —
  • 49.3
  • —
  • —
  • —
QwenQwen3 8BQwenContext131.1KInput$0.180/MOutput$2.10/MProviders
—
32±13.9P
42.2±11.1E
—
48.4±11.8E
—
—
—
Throughput
—
Latency
—
Release date2025-04-28
  • —
  • 32
  • 42.2
  • —
  • 48.4
  • —
  • —
  • —
QwenQwen3 14BQwenContext131.1KInput$0.350/MOutput$4.20/MProviders
—
40.2±13.9P
44.9±11.1E
—
49.7±11.8E
—
—
—
Throughput
—
Latency
—
Release date2025-04-28
  • —
  • 40.2
  • 44.9
  • —
  • 49.7
  • —
  • —
  • —
QwenQwen3 32BQwenContext131.1KInput$0.700/MOutput$8.40/MProviders
—
41.2±13.9P
46.5±11.1E
—
49.9±11.8E
—
—
—
Throughput
—
Latency
—
Release date2025-04-28
  • —
  • 41.2
  • 46.5
  • —
  • 49.9
  • —
  • —
  • —
QwenQwen3 235B A22BQwenContext131.1KInput$0.700/MOutput$8.40/MProviders
—
43.1±13.9P
45.8±11.1E
—
51±11.8E
—
—
—
Throughput
—
Latency
—
Release date2025-04-28
  • —
  • 43.1
  • 45.8
  • —
  • 51
  • —
  • —
  • —
QwenQwen2.5 VL 72B InstructQwenContext128KInput—Output—Providers
—
—
—
—
—
—
—
—
Throughput
—
Latency
—
Release date2025-02-01
  • —
  • —
  • —
  • —
  • —
  • —
  • —
  • —
QwenDeepSeek R1 Distill QwenDeepSeekContext128KInput—Output—Providers
+51
—
40±13.9P
39.3±8.9
55.4±16.0P
55.7±11.8E
—
—
37.6±16.0P
Throughput
50 t/s
Latency
19.88s
Release date2025-01-29
  • —
  • 40
  • 39.3
  • 55.4
  • 55.7
  • —
  • —
  • 37.6
What is an AI model directory?

An AI model directory is a searchable list of models and their comparison data. On LMSpeed, it brings together API price, throughput, first-token latency, provider coverage, and capability data when available.

How do I choose a model and provider?

Start with the metric that matters most for your workload. Then open a model page to compare provider data. Run your own speed test before you use an endpoint in production.

Can model price and speed change?

Yes. Price, availability, throughput, and latency can change by model, provider, and time. Use current page values as a comparison signal and confirm with your own endpoint test.

4QwenQwen3.8 Max65.8±11.3
5ClaudeClaude Fable 565.7±8.2
1QwenQwen3.8 Max65.9±11.3
2OpenAIGPT-5.6 Sol61.7±8.7
3SparkMuse Spark 1.161.4±10.2
4ClaudeClaude Opus 4.561.4±8.1
5ClaudeClaude Fable 560.7±10.8
1ClaudeClaude Opus 570.1±6.6
2ClaudeClaude Fable 568.1±6.9
3QwenQwen3.8 Max67.3±11.2
4ClaudeClaude Opus 4.866±6.6
5OpenAIGPT-5.6 Sol63.2±9.3
1MetaAILlama 3.3985 t/s
2Mercury 2490 t/s
3GeminiGemini 3.5 Flash425 t/s
4OpenAIGPT-OSS384 t/s
5DeepSeekDeepSeek OCR284 t/s
Newest
Most tested
Recently tested
A–Z
Agents
Coding
Reasoning
Knowledge
Math
Multilingual
Multimodal
Instruction following
Model
Context
Input
Output
Providers
Agents
Coding
Reasoning
Knowledge
Math
Multilingual
Multimodal
Instruction following
Throughput
Latency
QwenQwen3.8 MaxQwen
Read the complete scoring methodology
Model pricing leaderboard
Model speed leaderboard
Latency leaderboard
Provider directory
LLM API speed test