2026 Gemini 3.7 Flash (Aug 13) β $0.75/$3.75 intro, DeepSWE 65.3%. Grok 4.6 (1753 ELO, $2/$6). GLM-5.3 (743B open-weight coder). DeepSeek V4-Pro GA + peak/off-peak pricing. Open-weight and mid-tier frontier compress on price +84%
2026 Qwen 3.8-Max GA at $2/$6 flat across 1M context; MiniMax M2 ($0.30/$1.20) and Tencent Hy3 ($0.13/$0.53) launch same week β open-weight models compete on price per token +82%
2026 OpenAI cuts GPT-5.6 prices up to 80% three weeks post-launch β Luna to $0.20/$1.20, Terra to $2/$12. Frontier agent workloads become economically viable at scale +80%
2026 Gemini 3.6 Flash ships β 17% fewer output tokens, better coding & computer use. 3.5 Flash-Lite at 350 TPS. Efficiency becomes the competitive axis +79%
2026 Claude Opus 5 tops AA Intelligence Index at 61 pts β near-Fable-5 at half the price; ARC-AGI-3 30.2%. Top 3 models separated by 2 points +78%
2026 Gemini 3.5 Pro ships with 2M-token context + Deep Think; Kimi K3 (2.8T) and Qwen 3.8 Max (2.4T) launch β open-weight frontier catches up +75%
2026 GPT-5.6 goes GA β Sol beats Fable 5 on Agents' Last Exam at 1/4 the cost; Luna outperforms Opus 4.8 at ~1/16 the cost +72%
2026 Claude Fable 5 restored globally after US government suspension β government export controls now part of frontier model releases +68%
2026 OpenAI solves 80-year ErdΕs conjecture β AI's first breakthrough open math proof +65%
2026 GPT-5.5 launches; Opus 4.7 hits 16h+ task horizon with self-verification +60%
2025 AI agents hit multi-hour time horizons +12%
2024 Multimodal: vision, voice, reasoning unified +18%
2022 ChatGPT: 100M users in 2 months +10%
2020 GPT-3: Few-shot learning +8%
2016 AlphaGo defeats Lee Sedol +6%