| OpenAI |
GPT-6 Astra |
2026-09-03 |
1050K |
The GPT-6 flagship (Astra), built for the hardest end-to-end work: complex reasoning, coding, computer use, research, and long-horizon agents. 1.05M context. Advanced cyber capabilities first shipped through the Daybreak Blue program. |
| Anthropic |
Claude Fable 5.1 |
2026-09-01 |
1M |
The latest public Mythos-class tier, for demanding reasoning and ultra-long-horizon agents. Use it when high-effort Opus 5 is still not enough. |
| Zhipu GLM |
GLM-5.3 |
2026-08-14 |
1M |
Post-training redone on the 5.2 base. Migration requires thinking to be on. Rolled out in stages for Coding Plan and the API. |
| DeepSeek |
DeepSeek-V4-Pro |
2026-08-13 |
1M |
The official V4 flagship. About 1.6T / 49B active, MIT-licensed, aimed at the strongest open-source general capability. |
| xAI |
Grok-4.6 |
2026-08-12 |
500K |
The latest public Grok generation as of August 2026. About 500K context, for general chat, search, and reasoning. |
| Qwen |
Qwen3.8-Max |
2026-08-03 |
1M |
A native-vision MoE flagship of about 2.4T. Hybrid thinking is on by default, with 1 million context. It was Qwen's strongest model at the time. |
| DeepSeek |
DeepSeek-V4-Flash |
2026-07-31 |
1M |
The official V4 efficient tier. About 284B / 13B active, aimed at low latency and large-scale calls. |
| Anthropic |
Claude Opus 5 |
2026-07-24 |
1M |
The Claude 5 flagship. Near Fable intelligence at about half the price, for complex enterprise coding and long-horizon agents. |
| Gemini |
Gemini 3.6 Flash |
2026-07-21 |
1048576 |
The latest public Flash as of July 2026. About 1049K context, with coding and speed clearly stronger than 2.5 Flash. |
| Kimi |
Kimi K3 |
2026-07-16 |
1M |
A 2.8T open-source flagship with native vision and 1 million context, for long-horizon coding, knowledge work, and reasoning. Full weights opened on July 27. |
| OpenAI |
GPT-5.6 Sol |
2026-07-09 |
1050K |
The GPT-5.6 frontier tier (Sol). After public release it was the strongest GPT at the time, with programmatic tool calling and multi-agent orchestration. |
| OpenAI |
GPT-5.6 Terra |
2026-07-09 |
1050K |
The GPT-5.6 balanced tier (Terra). A tradeoff between intelligence and cost, suited to most production workloads. |
| OpenAI |
GPT-5.6 Luna |
2026-07-09 |
1050K |
The fastest and cheapest GPT-5.6 tier (Luna), for high-throughput chat and lightweight tasks. |
| xAI |
Grok-4.5 |
2026-07-08 |
500K |
A mid-to-late 4.x flagship with strong coding and text Arena results. Context and reasoning efficiency kept improving. |
| Anthropic |
Claude Sonnet 5 |
2026-06-30 |
1M |
The Claude 5 balanced tier, balancing speed and intelligence. Adaptive thinking is on by default, for most agents and knowledge work. |
| Zhipu GLM |
GLM-5.2 |
2026-06-16 |
1M |
Context expanded to 1 million. Adds IndexShare, MTP, and reasoning-effort control. MIT-licensed. |
| Kimi |
Kimi K2.7 Code |
2026-06-12 |
256K |
A thinking-mode coding model for software engineering. Thinking only, optimized for repo-scale long tasks. |
| Anthropic |
Claude Fable 5 |
2026-06-09 |
1M |
The first Mythos-class high-end model opened to regular users. Aimed at ultra-long-horizon agents and the strongest reasoning, with 1M context. |
| Qwen |
Qwen3.7-Plus |
2026-06-01 |
1M |
The 3.7 balanced hosted tier. 1 million context, covering office production and long-horizon agents. |
| Anthropic |
Claude Opus 4.8 |
2026-05-28 |
1M |
The closing 4.x flagship, continuing to beat 4.7 on coding, agent skills, reasoning, and knowledge work. |