MinimaxMinimax·text

MiniMax M3 Preview

CodeVisionReasoningWeb searchFunction calling
Advertised as m3-previewminimax-m3-preview
Quick reference
MiniMax M3 Preview — TLDR
  • 🧠 Frontier 1.4-trillion-parameter MiniMax model for coding, agents, reasoning.
  • 📏 512K-token context window in this preview, served at fp8.
  • 🆕 Built on new MiniMax Sparse Attention (MSA) for long context.
  • ⚡ MSA enables efficient native ultra-long-context pretraining.
  • 🔧 Function calling, tool use, and structured agentic task execution.
  • 👁️ Sibling M3 is natively multimodal (text, image, video input).
  • 🌐 Web search and long-horizon agentic workflows supported.
  • 🎯 Targets autonomous coding and multi-step agentic reasoning.
💰 Best price on AntSeed
$0.090 / $0.360
per 1M · cheapest in / out
📏 Context
524K tokens
🐜 Sellers
4
advertising on AntSeed
Provider

MiniMax is an AI company building generative models across multiple modalities, with a focus that spans both language understanding and audio creation. Their rapid release cadence in early 2026—delivering several new models within just a few months—reflects an ambitious and…

Explore 3 more models by Minimax
About this model

MiniMax M3 Preview is a preview build of MiniMax's flagship M-series language model, described in this catalog as a 1.4-trillion-parameter frontier model for coding, agentic workflows, and complex reasoning, served at fp8 with a 512K-token context window. It is positioned alongside the full [[sibling:minimax-m3|MiniMax M3]] release, which MiniMax presents as a model combining frontier coding, ultra-long context, and native multimodal input in a single architecture.

The central change from earlier M-series models is MSA (MiniMax Sparse Attention), which replaces the quadratic cost of full attention to enable native ultra-long-context pretraining, according to MiniMax. The production M3 supports up to 1M tokens with a guaranteed 512K minimum; this preview exposes the 512K tier.

Compared with prior family members such as [[sibling:minimax-m27|MiniMax M2.7]] and [[sibling:minimax-m25|MiniMax M2.5]], which remain available for existing workflows, MiniMax frames coding and agentic capability as M3's key areas of improvement, with autonomous task decomposition, tool invocation, and multi-step reasoning.

As a function-calling and web-search-capable model, M3 Preview is aimed at long-horizon agentic and computer-use tasks, including code generation and tool-driven workflows. Being a preview, weights, the technical report, and full availability above 512K tokens were still being rolled out around launch, with the model exposed here at the 512K context tier.

View source on GitHub ↗View model card on HuggingFace ↗
Sources
minimax.ioMiniMax M3: Frontier Coding, 1M Context, Native Multimodality — All in One Model - MiniMax Research | MiniMax· minimax.ioplatform.minimax.ioModel Invocation - MiniMax API Docs· platform.minimax.io

This About section is AI-generated from public sources via VeniceStats + Venice inference, with no human editing. It may contain inaccuracies.

Usage on AntSeed
Tokens served
17.90k
input + output
Requests
9
settled calls
Buyers
1
distinct, on this model
Sellers used
1
of 4 advertising
Settled
$0.00
gross USDC, this model
Sellers serving MiniMax M3 Preview (4)compare on the network explorer →
SellerReputationRoutingInput $/MCached $/MOutput $/MCategoriesAPI

"Best price" and the seller table are live AntSeed catalog data (advertised $/1M tokens — or $ per generated image for unit-billed image models — not settled amounts). Reputation = on-chain trust (0-100). "Routing" = the SDK's default buyer routing (what the VPR desktop app ships with): a trust ≥ 60 gate on the effective reputation, then cheapest-first among routable sellers; live failover state (per-peer cooldowns) is buyer-side runtime and not included. Model knowledge (TLDR, provider, About) via the VeniceStats enrichment layer. Advertised catalog, not the model used in any specific purchase. "Usage on AntSeed" counts only settlements whose buyers share the per-model split on-chain (metadata v2/v3, opt-in), so every usage figure is a lower bound.