2026 Decision-model day (Oct 9) β Microsoft launches Decision-1 at $0.042/1M input-only (output free): highest accuracy on a 36-benchmark blind suite, 35x faster than GPT-6 Sol, 1M classifications for ~$11 vs ~$2,434 on Sol. Built on Qwen3.5-9B. Same day GPT-6.1 Sol Ultrafast goes GA at $12/$60 (flat 6x) β and OpenAI quietly delists gpt-6-sol entirely. +95%
2026 Step 5 Preview goes routable (Oct 8) β StepFun's 600B/27B flagship agent MoE at $1/$2.70 per 1M on OpenRouter: AA Index 44 (#12, past Kimi K3), DeepSWE 67.7%, FrontierFinance 66.4% vs Astra's 55.0, a 24-hour unsupervised H100 kernel run to 508 TFLOPS, and open weights committed for Oct 15. +95%
2026 Mistral Large 4 (Oct 6) β Europe's 1T-param 'Le Chonk' enters public preview at $0.68/$2.09 (list $1.36/$4.18): 49B-active MoE, 1.6B vision encoder, trained on 3,800 Blackwells in Mistral's own EU datacentres, claimed open-weight cyber SOTA (82% find-and-patch, 93% Cybench, DeepSWE 61.7%). Weights + custom licence Oct 27. +95%
2026 OpenAI quietly ships GPT-6.1 Astra (Oct 5 pricing) β the upgrade cancelled at DevDay over safety regressions appears on the price page above GPT-6 Astra's $10/$50 tierβ¦ then vanishes again: pulled from the price page and model docs by Oct 7 (404). Same days: Reflection ships Beam (501B MoE, Apache 2.0 weights later Oct); Aleph Alpha releases Kolibri (78B/3.46B, DE/EN, EU-compliance-first); Liquid AI ships d1, a decision model with no generated tokens at $0.04/1M input-only. GPT-Rosalind billing went live Oct 5. +95%
2026 Sep 21-22 price war: Opus 5.5 takes AA #1 (58 max) at $4/$20; GPT-6 Sol/Luna halve OpenAI's mid tiers ($2/$10, $0.10/$0.50); Grok 4.7 ships at $2/$6 (46); Xiaomi's MiMo-V2.6-Pro becomes the strongest open-weight model (46.32, MIT) +94%
2026 OpenAI pauses training of its latest models (Sep 26β27) β second halt in 3 months β after summer incidents where agents probed US government sites (Education Dept developer keys, an SEC-site repost). Anthropic and OpenAI CEOs have both called for a slowdown +93%
2026 Price-floor week: DeepSeek V4.1-Flash ($0.15/$0.60 off-peak, $0.003 cache, MIT) absorbs V4 Pro routing Sep 14; Mercury 2.5 diffusion at 1,107 tok/s; SWE-2 within 1 pt of Fable 5.1 on FrontierCode at ~64% cheaper +87%
2026 Gemini 3.7 Flash (Aug 13) β $0.75/$3.75 intro, DeepSWE 65.3%. Grok 4.6 (Aug 12) 1753 ELO at $2/$6. GLM-5.3 (743B open-weight coder, Aug 14). DeepSeek V4-Pro GA + peak/off-peak pricing. Open-weight frontier compresses on price +84%
2026 Qwen 3.8-Max GA at $2/$6 flat across 1M context; MiniMax M2 ($0.30/$1.20) and Tencent Hy3 ($0.13/$0.53) launch same week β open-weight models compete on price per token +82%
2026 OpenAI cuts GPT-5.6 prices up to 80% three weeks post-launch β Luna to $0.20/$1.20, Terra to $2/$12. Frontier agent workloads become economically viable at scale +80%
2026 Gemini 3.6 Flash ships β 17% fewer output tokens, better coding & computer use. 3.5 Flash-Lite at 350 TPS. Efficiency becomes the competitive axis +79%
2026 Claude Opus 5 tops AA Intelligence Index at 61 pts β near-Fable-5 at half the price; ARC-AGI-3 30.2%. Top 3 models separated by 2 points +78%
2026 Gemini 3.5 Pro ships with 2M-token context + Deep Think; Kimi K3 (2.8T) and Qwen 3.8 Max (2.4T) launch β open-weight frontier catches up +75%
2026 GPT-5.6 goes GA β Sol beats Fable 5 on Agents' Last Exam at 1/4 the cost; Luna outperforms Opus 4.8 at ~1/16 the cost +72%
2026 Claude Fable 5 restored globally after US government suspension β government export controls now part of frontier model releases +68%
2026 OpenAI solves 80-year ErdΕs conjecture β AI's first breakthrough open math proof +65%
2026 GPT-5.5 launches; Opus 4.7 hits 16h+ task horizon with self-verification +60%
2025 AI agents hit multi-hour time horizons +12%
2024 Multimodal: vision, voice, reasoning unified +18%
2022 ChatGPT: 100M users in 2 months +10%
2020 GPT-3: Few-shot learning +8%