Models Are Getting Dumber on Purpose
🚀 Check out this insightful post from Hacker News 📖
📂 **Category**:
📌 **What You’ll Learn**:
Reasoning scores keep climbing while per-token compute keeps dropping. GLM-5.2 scores 99.2% on AIME 2026 with about 40 billion parameters active per token. Qwen3.5 scores 91.3% with 17 billion ...
