Models·3 min read·The Decoder

MiniMax M3 Lands as an Open-Weight, Million-Token Coding Model That Claims to Edge Out GPT-5.5

The Chinese lab says its new open-weight model pairs frontier coding, a 1M-token context window and native multimodality on a sparse-attention architecture that cuts long-context compute 20x — but the weights and technical report are still days away, so every number is company-reported.

OPEN MODELS · CODING · 1M CONTEXTMINIMAX M3 · JUNE 1MiniMax M3Open-weight · native multimodal · MSA attention~1/20 compute at 1M ctx · 9× faster prefill · 15× decodeBENCHMARK SCORECARD · MINIMAX-REPORTEDSWE-Bench Pro59%Terminal-Bench 2.166%SWE-fficiency34.8%BrowseComp83.5BROWSECOMPBeats Opus 4.783.5 vs 79.3Weights + technical report promised within 10 days · scores not yet independently verifiedBITSMINDS.COMSource: The Decoder · MiniMax
Share:

Chinese AI lab MiniMax unveiled M3 on June 1, 2026, calling it the first open-weight model to combine top-tier coding performance, a one-million-token context window and native multimodality in a single system — a bundle of capabilities the company says had until now been the exclusive domain of proprietary frontier models such as Anthropic's Claude Opus 4.7, OpenAI's GPT-5.5 and Google's Gemini 3.1 Pro.

The headline engineering claim is a new attention mechanism called MiniMax Sparse Attention, or MSA. Rather than comparing every token against every other token — the quadratic cost that makes long contexts expensive — MSA pre-filters down to the relevant key-value blocks and then processes them sequentially, batching the queries that need each block into a single contiguous memory read. MiniMax says the result is roughly one-twentieth the per-token compute at a million-token context versus its previous generation, more than 9x faster prefill, more than 15x faster decoding, and an implementation that runs over four times faster than competing open-source alternatives.

On benchmarks, MiniMax reports M3 scoring 59% on SWE-Bench Pro — ahead of GPT-5.5 and Gemini 3.1 Pro and just behind Opus 4.7 — and 83.5 on the BrowseComp web-search test, edging past Opus 4.7's 79.3. The company also published three long-horizon autonomy experiments: M3 reproduced an ICLR 2025 fine-tuning paper over about 12 hours, producing 18 commits and 23 figures for a reproduction score of 0.650; it optimized an FP8 GEMM kernel on Nvidia Hopper GPUs from a broken 7.6% hardware utilization up to 71.3% across roughly 24 hours and 147 attempts; and on PostTrainBench it trained four base models end to end, landing just behind Opus 4.7 and GPT-5.5.

The crucial caveat is that none of this can yet be checked. At launch MiniMax had released neither the weights nor a technical report, promising both within ten days on Hugging Face and GitHub, along with open-sourcing its in-house MiniMax Code agent. Token plans run from about $20 a month for roughly 1.7 billion tokens up to $120 for around 9.8 billion, with a toggleable thinking mode. Until independent engineers can reproduce the architecture and rerun the benchmarks, M3's frontier and open-weight claims remain a company commitment rather than a verified fact.

Want AI news before everyone else?

The morning's most important AI stories, straight to your inbox. No fluff.

Related Articles

GOOGLE · GEMINI Three Flash models ship… …while Gemini 3.5 Pro is still nowhere to be seen 3.6 Flash −17% output tokens 3.5 Flash-Lite cheap · beats Gemini 3 3.5 Flash Cyber finds & fixes bugs Gemini 3.5 Pro promised June · still not shipped → Meanwhile, Google says it has begun its biggest pretraining run yet — Gemini 4 Shipped July 21 via the Gemini API & Gemini Enterprise · the price war grinds on
Models

Google Ships Three New Gemini Flash Models — but 3.5 Pro Is Still Missing, and It's Already Pretraining Gemini 4

RUMOR WATCH · UNCONFIRMED Claude Opus 5 — the evidence board Cursor · Jul 9 “Claude Honeycomb EAP” — then gone Vertex AI · ×2 listing spotted twice, vanished both times 5? the next flagship? 1M context? leaked, unverified “xhigh” mode? launch “this week”?? No official API ID · no system card · every thread on this board is still a rumor
Models

Claude Opus 5 Rumor Watch: The 'Honeycomb' Leak, a Reported 1M Context, and Talk of a Launch This Week

MOONSHOT AI · OPEN WEIGHTS Kimi K3 The largest open-weight model ever — now #1 at frontend code, ahead of Fable 5 #1 Frontend Code Arena · 1,679 pts · ahead of Fable 5 2.8T · MoE 1M context 16/896 active Weights Jul 27 Open weights on Hugging Face · native multimodal · $0.30/$3 in · $15 out per million tokens
Models

Moonshot's Kimi K3 Is the Largest Open-Weight Model Ever — and It Just Beat Fable 5 at Frontend Code