Release dates and context windows of representative models from major providers.
  • Provider:
  • Model:
Provider Model Released Context window Description
xAI Grok-3 2025-02-17 1M A reasoning-era flagship trained at Colossus scale. It offers Think mode, with context up to 1 million tokens.
xAI Grok-3 mini 2025-02-17 1M A lighter reasoning version of Grok-3 that keeps Think mode, for lower cost and faster responses.
Gemini Gemini 2.0 Pro 2025-02-05 2M The 2.0 experimental flagship, focused on coding and complex prompts, with context up to 2 million.
OpenAI o3-mini 2025-01-31 200K The first reasoning model opened to free ChatGPT users. Near o1 on math and coding, with faster speed and lower price.
Qwen Qwen2.5-Max 2025-01-28 131072 Alibaba Cloud's hosted flagship MoE. It targeted then-current first-tier closed models, for general chat and complex tasks.
Kimi Kimi K1.5 2025-01-20 128K The start of the numbered model line. Stronger multimodal reasoning and RL, and the start of Kimi's thinking path.
DeepSeek DeepSeek-R1 2025-01-20 128K A reasoning model trained with pure RL. Open-sourcing it drew global attention. R1-Zero and several distilled sizes shipped with it.
DeepSeek DeepSeek-V3 2024-12-26 128K A 671B MoE (about 37B active) with FP8 training and MTP. It became the base for later R1 and the app.
Gemini Gemini 2.0 Flash 2024-12-11 1M The default high-speed model of the 2.0 era. Stronger native tools, multimodal output, and realtime capability. Veo 2 launched with it.
OpenAI o1 2024-12-05 200K The full o1 reasoning model, launched with ChatGPT Pro. The API opened on December 17, and context rose to 200K.
Anthropic Claude 3.5 Haiku 2024-10-22 200K The 3.5-generation fast model. Speed near Haiku and intelligence near early Sonnet, for low-latency production traffic.
Qwen Qwen2.5 2024-09-19 131072 The most widely used open-source generation. Full sizes from 0.5B to 72B, with broad gains in coding, math, and instruction following.
OpenAI o1-preview 2024-09-12 128K Preview of the o-series reasoning models. It runs an internal chain of thought before answering and was clearly stronger than then-current GPT-4o on math, coding, and science.
OpenAI o1-mini 2024-09-12 128K A lighter reasoning version of o1, focused on STEM and coding. Lower cost and faster, for cases that need some reasoning but not the strongest results.
DeepSeek DeepSeek-V2.5 2024-09-05 128K Merged V2-0628 and Coder-V2-0724 so general chat and coding share one model.
xAI Grok-2 2024-08-13 128K A full-generation Grok upgrade with stronger coding and writing, plus realtime X search and image generation.
OpenAI GPT-4o mini 2024-07-18 128K A cost-efficient small multimodal version of GPT-4o. It replaced GPT-3.5 Turbo in ChatGPT and fits high-volume lightweight tasks.
Anthropic Claude 3.5 Sonnet 2024-06-20 200K A breakout Sonnet generation. Coding, vision, and tool use far surpassed Claude 3 Opus, with excellent value.
DeepSeek DeepSeek-Coder-V2 2024-06-17 128K A code-specialized model based on V2. Covers more programming languages and is strong at long-context repo Q&A.
Qwen Qwen2 2024-06-06 131072 The third Qwen generation, stronger in multilingual and long-context work. 7B and above generally support 128K.

Description

Release dates and context windows are compiled from public sources. Limits can differ by platform, plan, or preview API. Please refer to each provider's official documentation.
The table above is sorted by release date in descending order.