Bring your own key: MiniMax's API rates, no markup. Keys and chats stay in your browser. Run MiniMax in parallel with other models, then compare and merge the answers.
MiniMax M3
Flagship: frontier coding and agentic reasoning, natively multimodal (text, image, video input). 1M context, 131K max output.
1M
$0.3
$1.2
May 2026
MiniMax M2.7
Recursive self-improvement and agentic capabilities. 200K context, 131K max output. ~60 t/s.
205K
$0.3
$1.2
Mar 2026
MiniMax M2.7 (Highspeed)
Faster M2.7 variant at ~100 t/s. 200K context, 131K max output.
205K
$0.6
$2.4
Mar 2026
MiniMax M2.5
Strong coding and reasoning, best value. 200K context, 65K max output.
205K
$0.3
$1.2
Feb 2026
MiniMax M2.5 (Highspeed)
Faster M2.5 variant at ~100 t/s. 200K context, 65K max output.
205K
$0.6
$2.4
Feb 2026
MiniMax M2-her
Dialogue-first model for immersive roleplay, character-driven chat, and expressive multi-turn conversations. 64K context.
66K
$0.3
$1.2
Jan 2026
MiniMax M2.1
230B params (10B active), multilingual coding. 200K context, 65K max output.
205K
$0.3
$1.2
Dec 2025
MiniMax M2.1 (Highspeed)
Faster M2.1 variant. 200K context, 65K max output.
205K
$0.6
$2.4
Dec 2025
MiniMax M2
230B params (10B active), agentic and reasoning. 200K context, 128K max output.
205K
$0.3
$1.2
Oct 2025
MiniMax M1
456B total / 45.9B active MoE with lightning attention. 1M context, 40K max output.
1M
$0.4
$2.2
Jun 2025
MiniMax 01
Legacy flagship. 1M context.
1M
$0.2
$1.1
Jan 2025
MiniMax M3
Flagship: frontier coding and agentic reasoning, natively multimodal (text, image, video input). 1M context, 131K max output.
MiniMax M2.7
Recursive self-improvement and agentic capabilities. 200K context, 131K max output. ~60 t/s.
MiniMax M2.7 (Highspeed)
Faster M2.7 variant at ~100 t/s. 200K context, 131K max output.
MiniMax M2.5
Strong coding and reasoning, best value. 200K context, 65K max output.
MiniMax M2.5 (Highspeed)
Faster M2.5 variant at ~100 t/s. 200K context, 65K max output.
MiniMax M2-her
Dialogue-first model for immersive roleplay, character-driven chat, and expressive multi-turn conversations. 64K context.
MiniMax M2.1
230B params (10B active), multilingual coding. 200K context, 65K max output.
MiniMax M2.1 (Highspeed)
Faster M2.1 variant. 200K context, 65K max output.
MiniMax M2
230B params (10B active), agentic and reasoning. 200K context, 128K max output.
MiniMax M1
456B total / 45.9B active MoE with lightning attention. 1M context, 40K max output.
MiniMax 01
Legacy flagship. 1M context.
1
Create an API key at the MiniMax console.
2
Paste it into Big-AGI's model settings.
3
Start chatting, or Beam it against other models and fuse the answers.
Connect MiniMax's OpenAI-compatible API with your own key, or route the M-series through OpenRouter. Either way, Big-AGI adds no markup and no intermediary: billing runs directly between you and MiniMax or OpenRouter, and your keys stay in your browser.
MiniMax's most visible consumer product is Talkie, a character-chat app, not a workbench for the M-series models. Beam is that workbench: it runs MiniMax next to GPT, Claude, and Gemini on the same prompt, so you can see whether its long-context read or its agentic plan holds up against the rest of the field before you rely on it. You also get parameters no character app exposes (temperature, system prompt, per-turn model swaps), and a key that stays in your browser.
Turn on Direct Connection and the browser calls MiniMax, or OpenRouter, directly, bypassing the Big-AGI server, whenever your key is client-side and the provider allows it. Your keys stay in your browser. Chats are stored locally first and sync only if you turn it on. The AI Inspector opens on any message to show the exact request, the token counts, and a cost estimate for that call.
Run MiniMax in parallel with Claude, GPT, and Gemini on the same prompt, then reach for Fusions: several strategies that combine, cross-check, and synthesize the parallel answers, which beats just picking the single best one. Parallel runs use more tokens than a single chat.
Your key, your data, your choice of model. Big-AGI is open source and self-hostable, so you can check exactly how MiniMax is called.
BIG-AGI
Resources
© 2026 Token Fabrics·Built with passion in San Diego