2 entries with this tag
June 16: One major research article published — deep-dive on Alibaba's Qwen3.7 Max & Plus family. Analysis of the open-to-closed pivot, 35-hour autonomous kernel demo, verbosity cost trap, and the dual-model strategy positioning against Opus 4.7 and GPT-5.5.
Alibaba's Qwen3.7 family — Max (closed-weight flagship, 1M context, SWE-Bench Pro 60.6%, $2.50/$7.50) and Plus (multimodal agent, vision+video, $0.32/$1.28) — represents a strategic pivot from open-weight leadership to closed-weight enterprise competition. Max scores 56.6 on the AA Intelligence Index (#5 overall, highest Chinese model), leads Opus 4.6 on agentic coding benchmarks, and completed a 35-hour autonomous kernel-optimization demo. Plus adds vision-language capabilities at roughly 1/6 the cost. This article analyses the full Qwen3.7 landscape, the open-to-closed pivot, benchmark reality, the verbosity cost trap, and where both models fit in the 2026 frontier.