AI Model Timeline
Key model releases from Transformer (2017) to today — filterable by category (LLMs, DLLMs, VLMs, Agents) and organization. ★ = PaperTrace deep-dive available. 🔓 = open weights.
Extended thinking, improved agentic capabilities — powers Claude Code. Released Feb 5, 2026.
Frontier performance across coding, agents, and professional work. Released Feb 17, 2026.
Native multimodal reasoning + agentic tool use — Google's strongest model at launch
Efficient MoE with novel load-balancing and multi-token prediction — strong coding and math at low training cost
o1-level reasoning via RL, fully open-weights MIT license — shocked markets
Best coding model at launch — outperforms GPT-4o on most benchmarks
Group Relative Policy Optimization — simpler PPO without critic network
Direct Preference Optimization — eliminates reward model and PPO from RLHF
RLHF alignment via PPO on human preferences — precursor to ChatGPT
Bidirectional Transformer pre-training with MLM — GLUE SOTA across all tasks
Attention Is All You Need — self-attention replaces RNNs entirely
Dates are approximate. Parameters are estimates where not officially confirmed.