MiniMax M3 API Benchmarks, Pricing & Provider Data
Compare MiniMax M3 with another model
Choose a model to open its comparison page.
MiniMax M3 benchmark, API pricing, and provider data cover undefined API provider} other undefined API providers}}, with prices starting at $0.0008/request. MiniMax M3 free API options are available from undefined provider} other undefined providers}}. The page also shows measured API speed and first-token latency.
MiniMax M3 is MiniMax next-generation large language model, designed for advanced reasoning, long-context understanding, and high-quality multilingual dialogue.
Category Performance
Observed-capability estimates with uncertainty reported separately from independent benchmark families.
- Coverage
- 6 / 8
- Methodology
- V3.0
#1Reasoning58.7Estimated2/4 Measured dimensions
#2Instruction following57.9Provisional1/4 Measured dimensions
#3Agents52.6RatedGlobal rank #303/4 Measured dimensions
#4Coding51.4RatedGlobal rank #244/4 Measured dimensions
#5Knowledge49.4Provisional1/4 Measured dimensions
#6Multimodal46.7Estimated2/4 Measured dimensions
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Technical Details
- Input
- Output
- Released
- May 2026
- Tokenizer
- Other
- Architecture
- text+image+video->text
- Moderated
- No
- Supported parameters
- frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
Rankings
Excels at
It's decent at
Falls behind in
Detailed scores
Updated: Sep 13, 2026Overall
undefined metric} other undefined metrics}}
Overall score61.0#43 / 112
Overall
undefined metric} other undefined metrics}}
Speed & latency
undefined metric} other undefined metrics}}
Output speed119.1 tok/s#36 / 80Time to first token0.83 s#20 / 80
Speed & latency
undefined metric} other undefined metrics}}
Pricing
undefined metric} other undefined metrics}}
Input price$0.300/M#55 / 186Output price$1.20/M#50 / 186
Pricing
undefined metric} other undefined metrics}}
Agents
V3.0undefined metric} other undefined metrics}} · Rated
Score52.6#3080% interval45.3–59.93/4 Measured dimensionsAgentic score58.4#43 / 77Terminal-Bench 2.066.0#22 / 53
Agents
V3.0undefined metric} other undefined metrics}} · Rated
Coding
V3.0undefined metric} other undefined metrics}} · Rated
Score51.4#2480% interval44.8–58.04/4 Measured dimensionsSciCode47.1%#40 / 89Coding score52.8#45 / 87
Coding
V3.0undefined metric} other undefined metrics}} · Rated
Reasoning
V3.0undefined metric} other undefined metrics}} · Estimated
Score58.780% interval47.9–69.52/4 Measured dimensionsGPQA92.9%#13 / 218HLE39.0%#34 / 216
Reasoning
V3.0undefined metric} other undefined metrics}} · Estimated
Knowledge
V3.0undefined metric} other undefined metrics}} · Provisional
Score49.480% interval35.4–63.31/4 Measured dimensionsArtificial Analysis Intelligence Index29.6#50 / 117AA-GPQA Diamond92.9#13 / 113
Knowledge
V3.0undefined metric} other undefined metrics}} · Provisional
Math
V3.0undefined metric} other undefined metrics}} · No data
80% interval30.8–69.20/4 Measured dimensionsUSAMO 202685.7#2 / 2
Math
V3.0undefined metric} other undefined metrics}} · No data
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
Multimodal
V3.0undefined metric} other undefined metrics}} · Estimated
Score46.780% interval34.7–58.72/4 Measured dimensionsMultimodal Grounded score52.2#45 / 56OfficeQA Pro45.1#9 / 10
Multimodal
V3.0undefined metric} other undefined metrics}} · Estimated
Instruction following
V3.0undefined metric} other undefined metrics}} · Provisional
Score57.980% interval41.9–73.91/4 Measured dimensionsAA-IFBench82.9#1 / 84Instruction Following score93.7#2 / 52
Instruction following
V3.0undefined metric} other undefined metrics}} · Provisional
OpenRouter endpoints
12 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
CoreWeave coreweave/fp4 | $0.230/M | $0.960/M | 99.9% | — | — | undefined tokens / undefined tokens |
Novita novita/fp8 | $0.300/M | $1.20/M | 99.9% | — | — | undefined tokens / undefined tokens |
Together together | $0.300/M | $1.20/M | 99.7% | — | — | undefined tokens / undefined tokens |
GMICloud gmicloud/fp8 | $0.240/M | $0.960/M | 99.7% | — | — | undefined tokens / undefined tokens |
ModelRun modelrun/fp4 | $0.750/M | $3/M | 99.6% | — | — | undefined tokens / undefined tokens |
StreamLake streamlake/fp8 | $0.300/M | $1.20/M | 99.3% | — | — | undefined tokens / undefined tokens |
SambaNova sambanova | $0.600/M | $2.40/M | 99.3% | — | — | undefined tokens / undefined tokens |
Minimax minimax/fp8 | $0.300/M | $1.20/M | 99.0% | — | — | undefined tokens / undefined tokens |
DeepInfra deepinfra/fp8 | $0.280/M | $1.10/M | 97.6% | — | — | undefined tokens / undefined tokens |
Parasail parasail/fp8 | $0.300/M | $1.20/M | 97.3% | — | — | undefined tokens / undefined tokens |
Venice venice/fp8 | $0.300/M | $1.20/M | 92.8% | — | — | undefined tokens / undefined tokens |
AtlasCloud atlas-cloud/fp8 | $0.300/M | $1.20/M | 42.0% | — | — | undefined tokens / undefined tokens |
Pricing Comparison
Compare MiniMax M3 API pricing across 170 providers. Prices range from $0.0008/request to $300.00/M. OpenApi offers the lowest rate at $0.0008/request. 6 providers offer free API credits or a free tier.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 100% L2 100% | minimax-m3 | OpenModels | -30%$0.210/M Cache read$0.042/M | -30%$0.840/M | 6.2 t/s | 2.57 s | ||
L1 100% | MiniMax-M3 | default | -52%$0.144/M Cache read$0.029/M | -52%$0.575/M | — | — | — | |
L1 100% | MiniMax-M3 | minimax-officially | $0.300/M Cache read$0.060/M | $1.20/M | — | — | — | |
L1 100% | minimax-m3 | default | $0.030/request | - | — | — | — | |
L1 100% | MiniMax-M3 | default | -52%$0.144/M Cache read$0.029/M | -52%$0.575/M | — | — | — | |
L1 100% | minimax-m3 | default | -23%$0.230/M | -23%$0.921/M | — | — | — | |
L1 100% | MiniMax-M3 | default | $10.27/M | $10.27/M | — | — | — | |
L1 100% | MiniMax-M3 | default | $10.27/M | $10.27/M | — | — | — | |
L1 100% | minimax-m3 | default | -93%$0.021/M | -93%$0.082/M | — | — | — | |
L1 100% | minimax-m3 | level1 | $0.0010/request | - | — | — | — | |
L1 100% | minimax-m3 | deepseek | -23%$0.230/M | -23%$0.921/M | — | — | — | |
L1 100% | MiniMax-M3 | default | $4.20/M Cache read$0.840/M | $16.80/M | — | — | — | |
L1 100% | [hm]q|nv/minimax-m3 | default | $4.00/request | - | — | — | — | |
L1 100% | [xj]q|minimaxai/minimax-m3 | default | $4.00/request | - | — | — | — | |
L1 100% | MiniMax-M3 | Minimax-officially | $5.14/M | $5.14/M | — | — | — | |
L1 100% | minimax/minimax-m3:free | default | $75.00/M | $75.00/M | — | — | — | |
DeadlySignal API Free | L1 99% L2 100% | minimax-m3 | default | Free | Free | — | — | — |
L1 99% L2 85% | MiniMax/MiniMax-M3 | diamond-glm | $54.75/M | $54.75/M | — | — | — | |
L1 100% | minimax-m3 | default | -40%$0.180/M Cache read$0.036/M | -40%$0.720/M | — | — | ||
L1 100% | minimax-m3 | free | $0.420/M Cache read$0.084/M | $1.68/M | — | — | — |
Alternatives & Similar Models
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
DeepSeek V4 Pro
deepseek-v4-pro
DeepSeek V4 Pro is the professional-tier DeepSeek V4 model, targeting frontier reasoning, coding, and agent workflows with maximum capability.
GLM-5.2
glm-5-2
Zhipu GLM-5.2 is Zhipu latest flagship coding and agentic model with a 1M-token context window, enhanced reasoning modes, and long-horizon software engineering capabilities.
GLM-5.1
glm-5-1
Zhipu GLM-5.1 is a next-generation GLM model aimed at frontier reasoning, coding, and bilingual agent applications.
MiniMax M2.7
minimax-m2-7
MiniMax M2.7 is a high-tier M2-series model tuned for complex reasoning, long-context dialogue, and production-grade API workloads.
Kimi K2.5
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.
Frequently Asked Questions
- What benchmark data does MiniMax M3 include?
- LMSpeed shows MiniMax M3 benchmark context, API price, output speed, first-token latency, and provider data across 176 providers when those signals are available.
- What is the MiniMax M3 API price?
- MiniMax M3 has pricing from undefined provider} other undefined providers}}, ranging from $0.0008/request to $300.00/M. OpenApi has the lowest listed price.
- What does the MiniMax M3 API pricing table include?
- The MiniMax M3 API pricing table compares 176 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest MiniMax M3 API pricing?
- OpenApi currently has the lowest listed MiniMax M3 price at $0.0008/request across undefined provider} other undefined providers}}.
- Can I compare MiniMax M3 API price and speed together?
- Yes. LMSpeed shows MiniMax M3 API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
- Is MiniMax M3 API free?
- Yes, MiniMax M3 free API options are available through 6 providerundefined other undefined} on LMSpeed, including Future Hub, Zero API, DeadlySignal API, 初叶🍂Furry API, Moyanjdc API. These providers offer free API credits or a free tier with no per-token charges.
- Where can I get MiniMax M3 free API access?
- LMSpeed currently lists 6 free API providerundefined other undefined} for MiniMax M3: Future Hub, Zero API, DeadlySignal API, 初叶🍂Furry API, Moyanjdc API. Check each provider row before using it because free tier limits can change.
