MiniMax M2.5
minimax-m2-5minimax-m2.5MiniMax-M2.5minimax-m25- 🆕 MiniMax's productivity-focused LLM optimized for coding and agentic workflows.
- 🧠 Trained via large-scale RL across 200,000+ real-world environments.
- 📏 Catalog lists a 198K-token context window for long tasks.
- 🔧 Supports function calling, web search, and multi-step tool use.
- 🎯 Vendor-reported 80.2% on SWE-Bench Verified, 51.3% Multi-SWE-Bench.
- ⚡ Completes SWE-Bench Verified roughly 37% faster than M2.1.
- 📚 "Spec-writing" behavior: plans architecture before writing code.
- 💬 Strong on office workflows like Word, PowerPoint, Excel modeling.
MiniMax is an AI company building generative models across multiple modalities, with a focus that spans both language understanding and audio creation. Their rapid release cadence in early 2026—delivering several new models within just a few months—reflects an ambitious and…
Explore 3 more models by Minimax →MiniMax M2.5, released in February 2026, is a text model from MiniMax built for coding, agentic tool use, and office productivity. According to MiniMax, it was extensively trained with reinforcement learning across more than 200,000 complex real-world environments using the company's Forge agent-native RL framework and CISPO algorithm, with a process-reward mechanism for monitoring generation quality in long-context agent rollouts. The catalog lists a 198K-token context window plus reasoning, code-optimization, function-calling, and web-search capabilities.
Against its same-family predecessor M2.1, MiniMax reports concrete gains. On the provider's reported SWE-Bench Verified, M2.5 scores 80.2% while completing the evaluation about 37% faster than M2.1—end-to-end runtime dropping from 31.3 to 22.8 minutes and tokens per task falling from 3.72M to 3.52M. MiniMax also reports 51.3% on Multi-SWE-Bench and 76.3% on BrowseComp with context management. A notable behavioral change is M2.5's tendency to decompose and plan features, structure, and UI like a software architect before coding.
MiniMax positions M2.5 for the full development lifecycle across Web, Android, iOS, Windows, and Mac, and for workspace tasks such as financial modeling and report generation. A higher-throughput M2.5-highspeed variant is also offered.
M2.5 was later succeeded within the family by MiniMax M2.7, MiniMax M3, and MiniMax M3 Preview, all sharing the same coding-and-agentic focus. For deployment, MiniMax recommends vLLM or SGLang.
This About section is AI-generated from public sources via VeniceStats + Venice inference, with no human editing. It may contain inaccuracies.
| Seller | Reputation↓ | Routing | Input $/M | Cached $/M | Output $/M | Categories | API |
|---|---|---|---|---|---|---|---|
| Venice.ai Proxy 0x1f22…18c9 | 100.00 | #4 | $0.17 | $0.02 | $0.595 | chat,reasoning,coding,web-search | openai-chat-completions |
| Argus AI 0x7adb…c915 | 100.00 | #2 | $0.10 | $0.01 | $0.30 | chat,coding,writing,creative | openai-chat-completions |
| Dark Signal 0x4668…62f2 | 95.76 | #6 | $0.21 | $0.04 | $0.84 | chat,writing,creative | openai-chat-completions |
| ▲ Apex Ant 0x73b4…e736 | 92.13 | #1 | $0.09 | $0.01 | $0.27 | chat,open-source,reasoning,long-context,agents | openai-chat-completions |
| Open Forge 0x1d90…b0aa | 76.54 | #7 | $0.27 | $0.03 | $0.95 | chat,writing,creative | openai-chat-completions |
| Open Ant 0xe4f6…5bc4 | 74.05 | #8 | $0.27 | $0.03 | $0.95 | chat,coding,reasoning,tools,cheap | openai-chat-completions |
| Vito-Minimax 0xddfa…27fe | 73.95 | gated | $0.17 | $0.17 | $0.595 | chat,coding,reasoning | openai-chat-completions |
| Edith AI 0xb269…b1a6 | 70.44 | gated | $0.24 | $0.24 | $0.96 | chat,coding | openai-chat-completions |
| NovaRoute AI 0xc50d…ed7b | 69.23 | #5 | $0.198 | $0.198 | $0.792 | chat,coding,code,writing,creative,tasks,minimax,value,surplus,openai-compatible,low-cost,verified,github,response-auth,base-usdc,monitored | openai-chat-completions |
| Super Seeder 0xd19f…41f3 | 68.85 | #3 | $0.135 | $0.015 | $0.475 | chat,coding,reasoning,tools,cheap | openai-chat-completions |
| BabyCai 0x8509…27b4 | 59.78 | gated | $0.10 | $0.10 | $0.20 | chat,agent,fast | openai-chat-completions |
| Fire Ant 🔥🐜 0xbe05…bc5d | 41.29 | gated | $0.2005 | $0.0223 | $0.7054 | — | — |
| bartly.eth64.de 0x666e…4666 | 36.10 | gated | $0.18 | $0.03 | $0.75 | chat,coding,reasoning,web-search,function-calling | openai-chat-completions |
| D5V1N2 0xd5e7…7be0 | 33.32 | gated | $0.03 | $0.006 | $0.105 | chat,coding,tasks,cheap,minimax,m25,m2-5,minimax-m25 | openai-chat-completions |
| Open Bird 0xc0f1…8183 | 18.32 | gated | $0.105 | $0.021 | $0.42 | chat | openai-chat-completions |
| Meridian AI 0x8c8c…06f5 | 5.64 | gated | $0.0255 | $0.0255 | $0.1021 | chat,reasoning | openai-chat-completions |
| uomi.ai 0x87df…48e3 | 4.80 | gated | $0.294 | $0.294 | $1.32 | chat,math,coding | openai-chat-completions |
| Apex TEE Test 0xe672…7955 | 0.51 | gated | $0.00 | $0.00 | $0.00 | chat,open-source,reasoning,long-context,agents | openai-chat-completions |
| antseed-tidal-falcon-d92f 0x25e1…d92f | 0.45 | gated | $0.005 | $0.00 | $0.05 | chat,coding,value | openai-chat-completions |
| Skeffo Inference 0x1af8…e2b5 | 0.02 | gated | $0.30 | $0.30 | $1.20 | chat,reasoning,agent,function-calling | openai-chat-completions |
| AntFeed 0xddb6…1442 | 0.00 | gated | $0.182 | $0.182 | $1.2841 | chat | openai-chat-completions |
| antseed-neon-puma-944e 0x6650…944e | 0.00 | gated | $0.1159 | $0.04 | $0.4058 | chat,coding,math | openai-chat-completions |
| Leftermute 0x388b…5389 | 0.00 | gated | $0.0309 | $0.0309 | $0.1082 | chat,coding,json,tools | openai-chat-completions |
"Best price" and the seller table are live AntSeed catalog data (advertised $/1M tokens — or $ per generated image for unit-billed image models — not settled amounts). Reputation = on-chain trust (0-100). "Routing" = the SDK's default buyer routing (what the VPR desktop app ships with): a trust ≥ 60 gate on the effective reputation, then cheapest-first among routable sellers; live failover state (per-peer cooldowns) is buyer-side runtime and not included. Model knowledge (TLDR, provider, About) via the VeniceStats enrichment layer. Advertised catalog, not the model used in any specific purchase. "Usage on AntSeed" counts only settlements whose buyers share the per-model split on-chain (metadata v2/v3, opt-in), so every usage figure is a lower bound.