MiniMax API Pricing Calculator & Cost Guide
Calculate MiniMax text API costs for 9 models. Compare per-token pricing across MiniMax-M2 and the long-context MiniMax-01 family.
Pricing TLDR
- • Pay-per-token with no monthly fees, priced separately for input and output
- • Long context windows and agentic text models are the MiniMax draw
- • 9 MiniMax text models with live rates from the OpenRouter catalog
MiniMax API Cost Calculator: Monthly Pricing
Calculate by
Input Tokens
Output Tokens
API Calls / Month
Quick Examples:
Sort:
(minimax/minimax-m3)
Context
Quality
Popularity
Per 1M Tokens
In: $0.30
Out: $1.20
Monthly Cost
(minimax/minimax-m3:batch)
Context
Quality
Popularity
Per 1M Tokens
In: $0.30
Out: $1.20
Monthly Cost
(minimax/minimax-m2.7)
Context
Quality
Popularity
Per 1M Tokens
In: $0.30
Out: $1.20
Monthly Cost
(minimax/minimax-m2.5)
Context
Quality
Popularity
Per 1M Tokens
In: $0.27
Out: $1.08
Monthly Cost
(minimax/minimax-m2.1)
Context
Quality
Popularity
Per 1M Tokens
In: $0.30
Out: $1.20
Monthly Cost
(minimax/minimax-m2)
Context
Quality
Popularity
Per 1M Tokens
In: $0.26
Out: $1.02
Monthly Cost
(minimax/minimax-m1)
Context
Quality
Popularity
Per 1M Tokens
In: $0.55
Out: $2.20
Monthly Cost
(minimax/minimax-01)
Context
Quality
Popularity
Per 1M Tokens
In: $0.20
Out: $1.10
Monthly Cost
(minimax/minimax-m2-her)
Context
Quality
Popularity
Per 1M Tokens
In: $0.30
Out: $1.20
Monthly Cost
Spending across LLM providers?
Track your AI API costs across all providers in real-time.

About MiniMax
What is MiniMax?
MiniMax is a Chinese AI lab whose text API covers a family of large language models built for long-context and agentic work, led by MiniMax-M2 and the long-context MiniMax-01 family. Pricing is per token, billed separately for input and output. This page covers only the text and LLM endpoints; MiniMax's audio, text-to-speech, and Hailuo video products are billed on their own separate pricing.
- Very Long Context Windows: The MiniMax-01 family is designed around large context windows, so it can take in long documents, codebases, and extended histories in a single call without heavy chunking.
- Agentic Text Models: MiniMax-M2 targets tool use and multi-step agentic workflows, aiming for capable reasoning at a low per-token cost so agents can run many steps affordably.
- Text Only Here: MiniMax also sells audio, TTS, and Hailuo video generation. Those are separate products with separate billing and are out of scope for this text pricing page.
When to Use MiniMax
MiniMax text models are a strong fit when you need long context or many agent steps at a low price and can work with a lab outside the Western frontier providers.
Ideal for
- Long-document and long-context text workloads
- Agentic pipelines that run many low-cost steps
- Cost-sensitive high-volume text generation
- Teams comparing Chinese labs on price and context
Not ideal for
- Audio, TTS, or Hailuo video needs, which are billed separately
- Tasks that need the absolute top of the quality leaderboard
- Workloads requiring Western data-residency guarantees
MiniMax Pricing Breakdown
How MiniMax API Billing Works
Per-Token Pricing
Each text model has separate input (prompt) and output (completion) rates per million tokens. You pay only for tokens processed, with no monthly minimum.
Long Context Drives Cost
The long context windows that make MiniMax attractive also mean big prompts consume many input tokens. Feeding whole documents on every call is the fastest way to run up a bill.
Text and Media Bill Separately
The rates here apply to text and LLM endpoints only. MiniMax audio, TTS, and Hailuo video generation are separate products billed per second, per character, or per generation.
Usage Tracking
Monitor spend per model and per key in the MiniMax platform console. Set alerts so a runaway agent loop or a switch to a pricier model does not surprise you at month end.
MiniMax API Monthly Cost Estimates
Hobby / Testing
$0-15/mo
• Short prompts
• <1K requests/day
• Single project
Light Use
$15-75/mo
• MiniMax-M2
• 1-5K requests/day
• Mixed text tasks
Medium Use
$75-400/mo
• Long context
• 5-20K requests/day
• Production apps
Heavy Use
$400+/mo
• Agentic loops
• 20K+ requests/day
• Long documents
5 MiniMax Cost Optimization Tips
Watch Your Context Size
Long context is MiniMax's strength and its biggest cost driver. Only load the context a call actually needs, and summarize or retrieve instead of pasting whole documents every time.
Match the Model to the Task
Reserve the larger models for genuinely hard reasoning and long-context work. Route simple classification, extraction, and short chat to the cheapest model that clears your quality bar.
Cap Output Length
Output tokens usually cost more than input. Set a max token limit, ask for structured or terse responses, and avoid open-ended generations when a short answer will do.
Budget Media Separately
If you also use MiniMax audio, TTS, or Hailuo video, track that spend on its own. Mixing it into your text budget hides where the money actually goes.
Track Spend with CostGoat
Watch your MiniMax credit balance and usage in real time with CostGoat. Get desktop or email alerts before you run low, and see which models drive your spend.
Start Tracking Your LLM API Spending
Monitor spending across OpenAI, Anthropic, Google, and other LLM providers from one menubar app.

MiniMax API Pricing FAQ
Common questions about MiniMax AI text API costs and billing
