| Qwen |
Qwen3.7-Max |
2026-05-20 |
1M |
The next Max-series flagship. Thinking is on by default, with 1 million context. Later snapshots added visual understanding. |
| Gemini |
Gemini 3.5 Flash |
2026-05-01 |
1048576 |
The 3.5 high-speed tier, for low-latency multimodal apps. Context is about 1049K. |
| DeepSeek |
DeepSeek-V4 Preview |
2026-04-24 |
1M |
A V4 preview with 1 million context by default. Split into Pro and Flash MoE tiers, with open weights. |
| OpenAI |
GPT-5.5 |
2026-04-23 |
1M |
The next default flagship for ChatGPT and Codex. Instant became the full default on May 5, with fewer hallucinations and better personalization. |
| Kimi |
Kimi K2.6 |
2026-04-21 |
256K |
A general multimodal K2 iteration with stronger long-horizon coding and agent workflows. It was the recommended model before K3. |
| Anthropic |
Claude Opus 4.7 |
2026-04-16 |
1M |
A broadly released flagship for complex reasoning and long-horizon coding. Higher visual resolution, with Claude Design launched at the same time. |
| Qwen |
Qwen3.6 |
2026-04-16 |
262144 |
Built on 3.5 with stronger stability and real-world coding. Open-sourced sizes such as 35B-A3B and 27B. |
| Zhipu GLM |
GLM-5.1 |
2026-04-07 |
200K |
Aimed at extra-long tasks, strategy correction, and bug fixing. Officially, agents can run for hours. |
| OpenAI |
GPT-5.4 mini |
2026-03-17 |
1M |
Brings GPT-5.4-class capability down to a mid-small model. Fits high-concurrency agents and tool calling. |
| OpenAI |
GPT-5.4 nano |
2026-03-17 |
1M |
The smallest GPT-5.4 size. Keeps million-token context and fits classification, routing, and light tool calling. |
| OpenAI |
GPT-5.4 |
2026-03-05 |
1M |
The first mainline native computer-use model. API/Codex offer million-token context and tool search, with better token efficiency than GPT-5.2. |
| xAI |
Grok-4.20 |
2026-02-17 |
256K |
A multi-agent preview generation. Default 256K, expandable to 2M. Later Beta 2 strengthened instruction following and reduced hallucinations. |
| Anthropic |
Claude Sonnet 4.6 |
2026-02-17 |
1M |
A full upgrade in coding, computer use, long-context reasoning, and agent planning. 1M context first shipped as beta. |
| Qwen |
Qwen3.5 |
2026-02-16 |
262144 |
A new generation that combines multimodal capability, architecture efficiency, and large-scale RL. It first shipped MoE sizes such as 397B-A17B. |
| Zhipu GLM |
GLM-5 |
2026-02-12 |
200K |
About 744B total parameters and 40B active. Sparse attention improves long-context efficiency. |
| Anthropic |
Claude Opus 4.6 |
2026-02-05 |
1M |
An Opus with stronger coding. 1M context became a standard capability, and it became the default in office add-ins such as Excel and PowerPoint. |
| Gemini |
Gemini 3.1 Pro |
2026-02-01 |
1048576 |
A mid-3.x Pro with stronger custom tools and coding. Context is about 1049K. |
| Kimi |
Kimi K2.5 |
2026-01-27 |
256K |
Native vision and Agent Swarm, bringing multimodal and multi-agent collaboration into a general model. |
| Zhipu GLM |
GLM-4.7 |
2025-12-22 |
200K |
The closing 4.x generation, with another step up in end-to-end agents and coding reasoning. |
| OpenAI |
GPT-5.2 |
2025-12-11 |
400K |
A GPT-5 flagship for professional knowledge work. Stronger long-horizon tool flows, and Thinking supports compact mode to go beyond a single context limit. |