MiniMax-M3 API Pricing & Cost Calculator
MiniMax-M3 is a strong value pick for high-volume utility work - 3x cheaper on output than Gemini Flash, with a real 1M context window. The published 50% batch discount makes it even cheaper at scale.
MiniMax-M3 is MiniMax's frontier offering in the ultra-cheap tier, designed for high-volume classification and routing workloads. A real 1M context window distinguishes it from same-tier peers.
By TechCompare · Updated
Calculator
Cost Comparison
Based on 30,000 input tokens (50% cached), 3,000 output tokens, and 1,000 requests.Prices are fetched live from OpenRouter and may include temporary promotional discounts not accounted for in our article and comparison figures.
How this is calculated
MiniMax-M3 is priced at $0.30 per million input tokens and $1.20 per million output tokens, with an 80% prompt caching discount ($0.06 per million cached reads). A 50% batch discount is also published on OpenRouter ($0.15 per million inputs, $0.60 per million outputs) - one of the few labs to expose batch directly via the OR API.
Verdict
MiniMax-M3 sits at $0.30/M input and $1.20/M output, with the 80% cache discount dropping cached input to $0.06/M and a published 50% batch discount (on OpenRouter) dropping both rows to $0.15/M input and $0.60/M output. The math against Gemini 3.6 Flash is roughly 3x cheaper on output, and the 1M context window matches MiniMax-M3's larger-tier rivals. The cache hit rate decides most of the per-call cost: a high-volume utility pipeline with stable prompts (classification, routing, formatting) sees input collapse to $0.06/M, and where OpenRouter's batch discount applies, the per-call total falls by another half. This is one of the few labs that exposes batch directly via the OR API, which is what unlocks the half-price bulk path.
More API Standalones scenarios
Related guides
Frequently asked questions
How does MiniMax-M3 stack up against GPT-5.6 Luna?
How much does MiniMax-M3 cost per million output tokens?
Related tools
LLM VRAM Calculator
Calculate the VRAM needed to run or fine-tune any LLM at any quantization.
Use tool ➜Power Cost Estimator
Estimate annual electricity costs for your PC, Server, or TV.
Use tool ➜Data Transfer Calculator
Estimate transfer times for files over USB, WiFi, Ethernet, and more.
Use tool ➜JSON Formatter
Validate, format, and minify JSON data with readable output and error detection.
Use tool ➜