MiniMax M3 API Benchmarks, Pricing & Provider Data
Compare MiniMax M3 with another model
Choose a model to open its comparison page.
MiniMax M3 benchmark, API pricing, and provider data cover 162 API providers, with prices starting at $0.0020/request. MiniMax M3 free API options are available from 5 providers. The page also shows measured API speed and first-token latency.
MiniMax M3 is MiniMax next-generation large language model, designed for advanced reasoning, long-context understanding, and high-quality multilingual dialogue.
Category Performance
Observed-capability estimates with uncertainty reported separately from independent benchmark families.
- Coverage
- 6 / 8
- Methodology
- V3.0
#1Reasoning59.2Estimated2/4 Measured dimensions
#2Instruction following57.8Provisional1/4 Measured dimensions
#3Agents53.7RatedGlobal rank #243/4 Measured dimensions
#4Coding52.7RatedGlobal rank #224/4 Measured dimensions
#5Knowledge51.7Provisional1/4 Measured dimensions
#6Multimodal46.9Estimated2/4 Measured dimensions
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Technical Details
- Input
- Output
- Released
- May 2026
- Tokenizer
- Other
- Architecture
- text+image+video->text
- Moderated
- No
- Supported parameters
- frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
Rankings
Excels at
It's decent at
Falls behind in
Detailed scores
Updated: Sep 1, 2026Overall
undefined metric} other undefined metrics}}
Overall score57.0#47 / 100
Overall
undefined metric} other undefined metrics}}
Speed & latency
undefined metric} other undefined metrics}}
Output speed125.9 tok/s#29 / 77Time to first token0.98 s#28 / 77
Speed & latency
undefined metric} other undefined metrics}}
Pricing
undefined metric} other undefined metrics}}
Input price$0.300/M#55 / 181Output price$1.20/M#50 / 181
Pricing
undefined metric} other undefined metrics}}
Agents
V3.0undefined metric} other undefined metrics}} · Rated
Score53.7#2480% interval46.3–61.03/4 Measured dimensionsAgentic score63.3#34 / 69Terminal-Bench 2.066.0#22 / 53
Agents
V3.0undefined metric} other undefined metrics}} · Rated
Coding
V3.0undefined metric} other undefined metrics}} · Rated
Score52.7#2280% interval46.1–59.34/4 Measured dimensionsSciCode45.4%#48 / 206Coding score53.8#38 / 81
Coding
V3.0undefined metric} other undefined metrics}} · Rated
Reasoning
V3.0undefined metric} other undefined metrics}} · Estimated
Score59.280% interval48.4–70.02/4 Measured dimensionsGPQA92.9%#10 / 213HLE39.0%#29 / 210
Reasoning
V3.0undefined metric} other undefined metrics}} · Estimated
Knowledge
V3.0undefined metric} other undefined metrics}} · Provisional
Score51.780% interval37.7–65.71/4 Measured dimensionsArtificial Analysis Intelligence Index45.4#26 / 111AA-GPQA Diamond92.9#9 / 108
Knowledge
V3.0undefined metric} other undefined metrics}} · Provisional
Math
V3.0undefined metric} other undefined metrics}} · No data
80% interval30.8–69.20/4 Measured dimensionsUSAMO 202685.7#2 / 2
Math
V3.0undefined metric} other undefined metrics}} · No data
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
Multimodal
V3.0undefined metric} other undefined metrics}} · Estimated
Score46.980% interval34.9–58.92/4 Measured dimensionsMultimodal Grounded score42.2#37 / 44OfficeQA Pro45.1#8 / 9
Multimodal
V3.0undefined metric} other undefined metrics}} · Estimated
Instruction following
V3.0undefined metric} other undefined metrics}} · Provisional
Score57.880% interval41.8–73.81/4 Measured dimensionsAA-IFBench82.9#1 / 84
Instruction following
V3.0undefined metric} other undefined metrics}} · Provisional
OpenRouter endpoints
11 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
Venice venice/fp8 | $0.300/M | $1.20/M | 100.0% | — | — | undefined tokens / undefined tokens |
AtlasCloud atlas-cloud/fp8 | $0.300/M | $1.20/M | 99.8% | — | — | undefined tokens / undefined tokens |
Novita novita/fp8 | $0.300/M | $1.20/M | 99.8% | — | — | undefined tokens / undefined tokens |
Minimax minimax/fp8 | $0.300/M | $1.20/M | 99.8% | — | — | undefined tokens / undefined tokens |
DeepInfra deepinfra/fp8 | $0.280/M | $1.10/M | 99.7% | — | — | undefined tokens / undefined tokens |
Parasail parasail/fp8 | $0.300/M | $1.20/M | 99.6% | — | — | undefined tokens / undefined tokens |
StreamLake streamlake/fp8 | $0.300/M | $1.20/M | 99.5% | — | — | undefined tokens / undefined tokens |
ModelRun modelrun/fp4 | $0.750/M | $3/M | 99.1% | — | — | undefined tokens / undefined tokens |
CoreWeave coreweave/fp4 | $0.230/M | $0.960/M | 98.3% | — | — | undefined tokens / undefined tokens |
SambaNova sambanova | $0.600/M | $2.40/M | 98.0% | — | — | undefined tokens / undefined tokens |
GMICloud gmicloud/fp8 | $0.600/M | $2.40/M | 82.3% | — | — | undefined tokens / undefined tokens |
Pricing Comparison
Compare MiniMax M3 API pricing across 157 providers. Prices range from $0.0020/request to $2997.00/M. 3173721 API offers the lowest rate at $0.0020/request. 5 providers offer free API credits or a free tier.
Alternatives & Similar Models
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
DeepSeek V4 Pro
deepseek-v4-pro
DeepSeek V4 Pro is the professional-tier DeepSeek V4 model, targeting frontier reasoning, coding, and agent workflows with maximum capability.
GLM-5.2
glm-5-2
Zhipu GLM-5.2 is Zhipu latest flagship coding and agentic model with a 1M-token context window, enhanced reasoning modes, and long-horizon software engineering capabilities.
GLM-5.1
glm-5-1
Zhipu GLM-5.1 is a next-generation GLM model aimed at frontier reasoning, coding, and bilingual agent applications.
MiniMax M2.7
minimax-m2-7
MiniMax M2.7 is a high-tier M2-series model tuned for complex reasoning, long-context dialogue, and production-grade API workloads.
Kimi K2.5
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.
Frequently Asked Questions
- What benchmark data does MiniMax M3 include?
- LMSpeed shows MiniMax M3 benchmark context, API price, output speed, first-token latency, and provider data across 162 providers when those signals are available.
- What is the MiniMax M3 API price?
- MiniMax M3 has pricing from 162 providers, ranging from $0.0020/request to $2997.00/M. 3173721 API has the lowest listed price.
- What does the MiniMax M3 API pricing table include?
- The MiniMax M3 API pricing table compares 162 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest MiniMax M3 API pricing?
- 3173721 API currently has the lowest listed MiniMax M3 price at $0.0020/request across 162 providers.
- Can I compare MiniMax M3 API price and speed together?
- Yes. LMSpeed shows MiniMax M3 API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
- Is MiniMax M3 API free?
- Yes, MiniMax M3 free API options are available through 5 providerundefined other undefined} on LMSpeed, including 初叶🍂Furry API, Zero API, Moyanjdc API, Dext API, Dext API. These providers offer free API credits or a free tier with no per-token charges.
- Where can I get MiniMax M3 free API access?
- LMSpeed currently lists 5 free API providerundefined other undefined} for MiniMax M3: 初叶🍂Furry API, Zero API, Moyanjdc API, Dext API, Dext API. Check each provider row before using it because free tier limits can change.
