3 entries with this tag
One new research article published: comprehensive deep-dive on MiniMax M2.7 — the first model to participate in its own evolution through self-improving agent harnesses, achieving 56.2% SWE-Pro at $0.30/M pricing with open weights.
MiniMax launches M2.7 on July 16, 2026 — the first model to participate in its own evolution through self-improving agent harnesses. Achieves 56.22% on SWE-Pro, 55.6% on VIBE-Pro, and 66.6% medal rate on MLE Bench Lite, all at $0.30/$1.20 per million tokens with open weights available on Hugging Face.
MiniMax M3 launched June 1, 2026 as the first open-weight model combining frontier coding (59% SWE-Bench Pro), 1M context, and native multimodality. Built on a new MiniMax Sparse Attention (MSA) architecture, it beats GPT-5.5 on SWE-Bench Pro at 12× lower cost. But vendor-run benchmarks, unreleased weights, China's National Intelligence Law, and restrictive licensing create serious caveats. M3 is the most compelling open-weight challenger yet — but the gap to Opus 4.8 remains real, and the geopolitical risks are structural.