MiniMax-M3 API Pricing & Cost Calculator
MiniMax-M3 is a strong value pick for high-volume utility work - 3x cheaper on output than Gemini Flash, with a real 1M context window. The published 50% batch discount makes it even cheaper at scale.
MiniMax-M3 is MiniMax's frontier offering in the ultra-cheap tier, designed for high-volume classification and routing workloads. A real 1M context window distinguishes it from same-tier peers.
By TechCompare · Updated
Calculator
Cost Comparison
Based on 30,000 input tokens (50% cached), 3,000 output tokens, and 1,000 requests.
How this is calculated
MiniMax-M3 is priced at $0.30 per million input tokens and $1.20 per million output tokens, with an 80% prompt caching discount ($0.06 per million cached reads). A 50% batch discount is also published on OpenRouter ($0.15 per million inputs, $0.60 per million outputs) - one of the few labs to expose batch directly via the OR API.
Verdict
MiniMax-M3 is a strong value pick for high-volume utility work - 3x cheaper on output than Gemini Flash, with a real 1M context window. The published 50% batch discount makes it even cheaper at scale.
More API Standalones scenarios
Frequently asked questions
How does MiniMax-M3 stack up against GPT-5.6 Luna?
Related tools
LLM VRAM Calculator
Calculate the VRAM needed to run or fine-tune any LLM at any quantization.
Use tool ➜Power Cost Estimator
Estimate annual electricity costs for your PC, Server, or TV.
Use tool ➜Data Transfer Calculator
Estimate transfer times for files over USB, WiFi, Ethernet, and more.
Use tool ➜JSON Formatter
Validate, format, and minify JSON data with syntax highlighting.
Use tool ➜