Jul 16, 2026
Kimi K3
2.8T params · 1M ctx · open weights Jul 27
Largest open-source model ever released at 2.8 trillion parameters with 1M-token context. API launched July 16; full weights dropped July 27. Rivals top US frontier systems on coding and reasoning benchmarks.
Jul 2026
GPT-5.6 (Sol, Luna, Terra)
ChatGPT + API · three-variant line
OpenAI's latest GPT-5 generation released as three named variants — Sol (speed), Luna (balanced), Terra (deep reasoning) — continuing the unified-router architecture.
Jul 8, 2026
Grok 4.5
xAI · coding-focused
Major Grok upgrade with tripled parameter count, targeting coding leadership. Strong SWE-bench performance and expanded agentic capabilities.
Jun 30, 2026
Claude Sonnet 5
Anthropic API + Claude.ai
Broadly available Claude 5 model combining the next-generation architecture with the speed and cost profile developers expect from the Sonnet tier.
Jun 9, 2026
Claude Fable 5 & Mythos 5
Anthropic · Claude 5 generation launch
First models in the new Claude 5 Mythos-class tier. Mythos 5 is limited to approved organizations. Fable 5 marks a new generation of long-context reasoning architecture.
Jun 16, 2026
GLM-5.2
MIT license · 1M context · coding-focused
Open-weight coding model under MIT license with 1M-token context. Part of Z.ai's GLM Coding Plan — NIST conducted a CAISI safety assessment at release, signaling growing international recognition of Chinese open-weight labs.
Jun 2026
Kimi K2.7 Code
Open weights · coding specialist
Coding-specialized iteration of the K2 line, competing in the rapidly crowding open-weight coding model space against Qwen3-Coder and DeepSeek V4-Flash.
May 2026
Claude Opus 4.8
Anthropic API
Final major iteration in the Opus 4.x line before the Claude 5 generation. Delivered Anthropic's highest benchmark scores to that point.
Apr 29, 2026
Mistral Medium 3.5
Frontier multimodal · agentic
Frontier-class multimodal model optimized for agentic and coding use cases. Significant step up from Medium 3.1 in both capability and vision understanding.
Apr 24, 2026
DeepSeek V4 (V4-Pro & V4-Flash)
1M context · open weights · aggressive pricing
Two open-weight models launched simultaneously — Pro and Flash — on Hugging Face, API, and chat. 1M context, full tool use, and API pricing well below US competitors.
Apr 23, 2026
GPT-5.5
ChatGPT + API · computer use
Added native computer-use capabilities — browser control, form-filling, and workflow execution. Marks GPT-5's evolution into an agentic operating model.
Apr 2026
Claude Opus 4.7
Anthropic API
Continued Anthropic's fast iteration on the Opus tier with enhanced performance on complex agentic and multi-turn tasks.
Apr 8, 2026
GLM-5.1
Open-sourced · Z.ai
Open-sourced two months after initial release, part of Z.ai's strategy of staged openness. Continued the GLM-5 generation's improvements in multilingual instruction following and long-context reasoning.
Apr 2026
Kimi K2.6
Open weights · agentic
Iterative upgrade in the K2 line with strengthened multi-step agentic reasoning. Moonshot shipped incremental versions rapidly in early 2026 ahead of the K3 generation.
Mar 3, 2026
Mistral Small 4
Apache 2.0
Next-generation small model continuing Mistral's shift to permissive open licensing. Efficient on modest hardware with strong multilingual and instruction-following performance.
Mar 2026
GPT-5.4
ChatGPT + API
Significant capability upgrade continuing OpenAI's rapid iteration cadence on the GPT-5 line, with improvements to instruction following and tool use reliability.
Jan 2026
Kimi K2.5
1T param MoE · multimodal · agentic
1 trillion parameter MoE with 32B active parameters, adding multimodal vision-language understanding and advanced agentic capabilities over K2. Maintained the 1M-token context window from its predecessor.
Feb 16, 2026
Qwen3.5
397B-A17B MoE · open weights
397B-A17B MoE open-weight model building on Qwen3's hybrid-reasoning architecture. Also released a closed proprietary Qwen3.5-Plus variant for enterprise deployments.
Feb 2026
GLM-5
Z.ai · new generation launch
Next-generation launch from Z.ai (formerly Zhipu AI), introducing a new capability tier with improved multilingual reasoning, coding, and long-context performance across Chinese and English.
Feb 2026
Claude Opus 4.6 & Sonnet 4.6
Anthropic API + Claude.ai default
Sonnet 4.6 became the default free model on Claude.ai. Competitive benchmark results across coding, analysis, and long-context tasks.
Dec 2, 2025
Mistral Large 3
41B active / 675B total · Apache 2.0
Sparse MoE model with 675B total parameters under Apache 2.0 — a major licensing shift from Mistral's earlier restrictive terms. 256K context, frontier-class performance on coding and reasoning.
Dec 2025
GPT-5.2
ChatGPT + API
Further refinement of the GPT-5 line with stronger reasoning consistency and expanded multimodal capabilities.
Dec 2025
GLM-4.7
Z.ai · pre-5 generation
Final major release in the GLM-4.x line before the GLM-5 generation. Refined multilingual instruction-following and reasoning, maintaining Z.ai's rapid quarterly release cadence.
Nov 2025
Kimi K2 Thinking
Open weights · extended reasoning
Reasoning variant of K2 adding extended chain-of-thought via a $4.6M training run. Delivered top-tier math and coding results in the open-weight class.
Nov 19, 2025
Grok 4.1
xAI · all grok.com users
Grok 4 iteration released broadly to all grok.com users with expanded tool use, improved coding, and tighter X platform integration.
Nov 2025
Claude Opus 4.5
Anthropic API
Enhanced Opus with stronger multi-step reasoning and improved agent capabilities for complex, long-horizon tasks.
Nov 2025
GPT-5.1
ChatGPT + API
Point release to GPT-5 with improved instruction-following, reduced hallucination rate, and faster inference under the unified model router.
Oct 17, 2025
Phi-4 Mini Instruct
Open weights · Azure + Hugging Face
Compact instruction-tuned SLM for on-device deployment. Optimized for latency-sensitive and privacy-first enterprise scenarios.
Oct 2025
Claude Haiku 4.5
Anthropic API
Fastest and most cost-efficient Claude model, targeting high-throughput production use cases and real-time applications.
Oct 2025
Kimi Linear
48B total · 3B active · linear attention
48B MoE with only 3B active parameters using Kimi Delta Attention (KDA) — a novel linear attention mechanism that significantly reduces memory usage and improves generation throughput versus standard attention.
Sep 2025
DeepSeek V3.2 Experimental
Open weights
Experimental checkpoint showing significant gains in coding and math. Later V3.2-Speciale won gold at IMO, IOI, and ICPC 2026.
Sep 2025
GLM-4.6
Z.ai · quarterly release
Iterative improvement in the GLM-4 family with stronger Chinese-English bilingual instruction following. Z.ai maintained a brisk sub-quarterly release cadence throughout 2025.
Aug 2025
DeepSeek V3.1
Hybrid thinking mode · open weights
First DeepSeek model with dual-mode functionality — thinking and non-thinking — within a single model endpoint, matching a pattern set by Qwen3.
Aug 2025
Claude Opus 4.1
Anthropic API
Incremental improvement to the Opus line with enhanced multi-step agentic task handling and better instruction-following across long documents.
Aug 7, 2025
GPT-5
ChatGPT + API
OpenAI unified its model lines into a single GPT-5 endpoint with an intelligent router that decides how much to think per request. Paired a fast response mode with a deep reasoning mode seamlessly.
Jul 10, 2025
Grok 4
xAI / X platform
xAI's fourth-generation flagship with enhanced reasoning and multimodal understanding. Trained on expanded Colossus cluster capacity.
Jul 2025
Kimi K2
1T params MoE · 32B active · 1M ctx
Massive trillion-parameter MoE open-weight model trained on 15.5T tokens. Released under a modified MIT license with 1M context, strong agentic and coding capability.
Jul 22, 2025
Qwen3-Coder (480B-A35B)
480B total · 35B active · 256K→1M ctx
Alibaba's flagship agentic coding model — 480B-parameter MoE with 35B active, native 256K context extrapolating to 1M. Open-source release set a new bar for open coding models, challenging DeepSeek Coder V3 and Kimi K2 on SWE-bench.
Jul 2025
GLM-4.5
Z.ai (Zhipu AI)
First 2025 GLM release from Zhipu AI (later rebranded Z.ai), continuing their strong bilingual Chinese-English instruction model line. Competitive with Western mid-tier models on multilingual and reasoning tasks.
Jun 17, 2025
Gemini 2.5 Pro, Flash & Flash-Lite (GA)
General availability · 1M context
Full production rollout of the 2.5 family. Flash became the default Gemini model at Google I/O; Flash-Lite optimized for cost-sensitive high-volume workloads.
May 28, 2025
DeepSeek R1-0528
Updated reasoning · open weights
Significant reasoning upgrade to the R1 family, achieving 97.3% on MATH-500. Continued DeepSeek's strategy of rapid open-weights iteration.
May 7, 2025
Mistral Medium 3
API + on-premise
Positioned between Small and Large tiers, offering GPT-4-class performance with significantly reduced inference cost for enterprise deployments.
May 2025
Claude Sonnet 4 & Opus 4
Anthropic API + Claude.ai
Claude 4 generation launch. Sonnet 4 replaced 3.7 as the default API model; Opus 4 targeted complex long-context and multi-step agent workflows.
Apr 30, 2025
Phi-4-reasoning & Phi-4-mini-reasoning
Open weights · GitHub Models
Small reasoning models outperforming models 5–10× larger on STEM tasks. Proved that distilled reasoning capability scales to sub-10B parameter models.
Apr 29, 2025
Qwen3 Family
0.6B → 235B MoE · full open source
Six dense + two MoE models, all fully open-sourced at launch. Hybrid reasoning mode (thinking/non-thinking) within a single model. The 235B-A22B MoE set new open-weight benchmarks.
Apr 16, 2025
OpenAI o3 & o4-mini
API + ChatGPT
o3 delivers frontier reasoning; o4-mini brings reasoning at speed and lower cost. Both support tool use natively within the reasoning chain.
Apr 5, 2025
Llama 4 Scout & Maverick
Scout: 10M ctx · Maverick: 1M ctx
Meta's natively multimodal MoE flagship. Scout (17B active / 109B total) ships with a record 10M-token context; Maverick (17B active / 400B total) targets instruction and coding tasks.
Mar 25, 2025
Gemini 2.5 Pro (Experimental)
Google AI Studio
Google's most capable reasoning model to date at launch. Topped multiple coding and math benchmarks with native thinking mode and 1M-token context.
Mar 12, 2025
Gemma 3
Open weights · 1B / 4B / 12B / 27B
Four-size family of open-weight models with multimodal capabilities. Strong on-device performance; the 27B model rivaled proprietary models from 2024.
Mar 5, 2025
QwQ-32B
Apache 2.0 · 32B · reasoning
Alibaba's reasoning-focused open-weight model that matched DeepSeek-R1 at 1/21 the parameter count. Apache 2.0, fits on a single 24GB GPU with quantization — made frontier-class reasoning accessible to consumer hardware.
Feb 27, 2025
Phi-4-mini & Phi-4-multimodal
Open weights · Hugging Face
Small-but-mighty SLMs punching above their weight class. Phi-4-multimodal integrates text, vision, and speech into a single efficient model.
Feb 2025
Grok 3
xAI / X platform
xAI's flagship model trained on the Colossus supercluster. Included DeepSearch web-grounding and Think mode for step-by-step reasoning.
Feb 2025
Claude 3.7 Sonnet
Anthropic API + Claude.ai
Anthropic's first hybrid-reasoning model with extended-thinking capability. Customers can toggle between standard and deep-reasoning modes on a single endpoint.
Jan 20, 2025
Kimi k1.5
Moonshot AI · long-context reasoning
Moonshot AI's multimodal long-thinking model released the same day as DeepSeek R1. Used reinforcement learning to achieve strong reasoning performance in a frontier API offering with 128K context.
Jan 20, 2025
DeepSeek R1
MIT license · open weights
Shocked the industry — GPT-o1-level reasoning at a fraction of the cost. MIT license. Sparked the open-weight reasoning race and triggered a global AI policy reaction.
Jan 31, 2025
OpenAI o3-mini
API + ChatGPT
Compact reasoning model offering o1-class performance at lower cost and latency. Available across API tiers and ChatGPT Plus/Pro.
Jan 2025
Mistral Small 3
Apache 2.0
Small, efficient instruction model released under Apache 2.0. Strong multilingual performance optimized for edge deployments.
Labs Covered
11 Labs · Two Model Classes
GPT-5 line · o-series
Frontier reasoning
Claude 4.x + 5
Sonnet / Opus / Haiku
Gemini 2.5 · Gemma 3
Frontier + open
Llama 4 Scout / Maverick
Open weights MoE
Grok 3 → 4.5
X platform
R1 · V3.x · V4
MIT open weights
Small / Medium / Large
Apache 2.0 shift
Qwen3 · QwQ · Qwen3-Coder
Qwen3.5 · full open-source
Phi-4 SLM line
On-device focus
Kimi k1.5 · K2 · K2.5
K2.6 · K2.7 · K3 (2.8T)
GLM-4.5 → GLM-5.2
MIT open coding models
Stay Ahead
We Track AI So You Can Ship
Our AI Software, Media, and Investing teams track the frontier and the open-weight ecosystem — so you deploy the right model for the right workload, not just the loudest launch.