
MiniMax consistently has some of the most underrated models. It has now announced M3, which is a new smart model with 1M context window. It performs admirably according to SWE-Bench Pro,Terminal Bench 2.1, SWE-fficiency, KernelBench Hard, and MCP Atlas.
Introducing MiniMax M3: The First Open-Weights Model to Combine Three Frontier Capabilities
– Coding & Agentic Frontier: 59.0% SWE-Bench Pro, 66.0% Terminal Bench 2.1, 34.8% SWE-fficiency, 28.8% KernelBench Hard, 74.2% MCP Atlas
– MiniMax Sparse Attention scales context to 1M
-… pic.twitter.com/TF891iJukF— MiniMax (official) (@MiniMax_AI) June 1, 2026
The model costs $1.2 for 1m input tokens and $4.80 for 1m output tokens. As we have covered in the past, you can also use this model with Claude Code’s ultracode mode.

