Language model timeline

155 releases21 organizations

Entries by year

Release history

Organizations
By AA Intelligence Index

155 entries

2026

Multimodality and agents67 entries

A new multimodal model for complex coding, knowledge work, and cyber defense, initially available to invited teams.

AnnouncementClosed sourceMultimodal
Source

Improves coding, computer use, and professional work, with task delegation available within a request.

ReleaseClosed sourceMultimodal
Source

The second Claude 5.5 model supports coding, document processing, and office tasks.

ReleaseClosed sourceMultimodal
Source

Opus targets complex knowledge work, coding, and judgment-intensive tasks.

ReleaseClosed sourceMultimodal
Source

Two GPT-6 models address different speed and cost requirements.

ReleaseClosed sourceMultimodal
Source

MiMo combines multimodal understanding with large-scale reinforcement learning.

ReleaseOpen sourceMultimodal
Source

Grok improves coding, self-verification, and context management.

ReleaseClosed sourceMultimodal
Source

A flagship preview targets long-running engineering, knowledge work, and research.

PreviewClosed sourceMultimodal
Source

Text, images, audio, and video share one multimodal model.

ReleaseClosed sourceMultimodal
Source

Kimi Code gains a stronger coding and agent preview.

PreviewClosed sourceMultimodal
Source

A new architecture adds native vision and more efficient agent reasoning.

ReleaseOpen sourceMultimodal
Source

GPT-6 Astra

OpenAIMilestone

GPT combines reasoning, computer use, and multi-step professional work.

ReleaseClosed sourceMultimodal
Source

Flash improves sustained coding and professional reasoning.

ReleaseClosed sourceMultimodal
Source

Muse improves long-running agents, coding, and continued collaboration.

ReleaseClosed sourceMultimodal
Source

Improves engineering projects, multi-tool orchestration, chart reasoning, and document parsing.

ReleaseClosed sourceMultimodal
Source

Tencent opens a flagship preview for long-context and complex work.

PreviewOpen sourceText
Source

Native vision and hybrid linear/sparse attention support coding, browser, and GUI tasks.

ReleaseOpen sourceMultimodal
Source

An efficient multimodal model supports coding and high-volume agent use.

ReleaseOpen sourceMultimodal
Source

GLM strengthens complex coding and extended task execution.

ReleaseOpen sourceText
Source

V4 Pro reaches general availability with stronger agents and Responses API support.

ReleaseOpen sourceText
Source

Flash improves multi-step coding, knowledge work, and web development.

ReleaseClosed sourceMultimodal
Source

Grok improves sustained work across agents, interactive apps, and vision.

AnnouncementClosed sourceMultimodal
Source

A native vision-language flagship supports a one-million-token context.

ReleaseOpen sourceMultimodal
Source

Claude Opus 5

Anthropic

Opus upgrades engineering, knowledge work, and computer use.

ReleaseClosed sourceMultimodal
Source

Grok advances coding, agent tasks, and professional knowledge work.

AnnouncementClosed sourceMultimodal
Source

Kimi K3

Moonshot AI

Kimi combines native vision, long context, and complex coding.

ReleaseOpen sourceMultimodal
Source

GPT-5.6

OpenAI

Sol, Terra, and Luna address different knowledge-work and coding needs.

ReleaseClosed sourceMultimodal
Source

Muse improves coding, computer use, and multimodal understanding.

AnnouncementClosed sourceMultimodal
Source

Tencent improves long-running agents and product integration.

ReleaseOpen sourceText
Source

Sonnet expands planning, browser use, terminal work, and sustained execution.

ReleaseClosed sourceMultimodal
Source

GLM extends long-running task and engineering capabilities.

ReleaseOpen sourceText
Source

Kimi K2.7 Code

Moonshot AI

Kimi improves long-running coding, tool calls, and efficient thinking.

ReleaseOpen sourceMultimodal
Source

MiniMax-M3

MiniMax

Sparse attention adds native image and video understanding.

Open weightsOpen sourceMultimodal
Source

Step improves efficient reasoning and practical agent execution.

ReleaseOpen sourceMultimodal
Source

Opus improves tool use, complex judgments, and collaboration.

ReleaseClosed sourceMultimodal
Source

A public preview targets sustained coding and remote agent tasks.

PreviewOpen sourceMultimodal
Source

Qwen improves agentic coding, office work, and long-running execution.

AnnouncementClosed sourceText
Source

Google I/O introduces a Flash update focused on speed and agent tasks.

ReleaseClosed sourceMultimodal
Source

DeepSeek-V4 Preview

DeepSeekMilestone

Pro and Flash previews introduce million-token context and open weights.

PreviewOpen sourceText
Source

GPT-5.5

OpenAI

GPT improves tasks spanning multiple tools and applications.

ReleaseClosed sourceMultimodal
Source

MiMo upgrades complex tasks, multimodal work, and inference efficiency.

ReleaseOpen sourceMultimodal
Source

A flagship preview strengthens coding and complex task execution.

PreviewClosed sourceText
Source

Opus improves difficult engineering and visual understanding.

ReleaseClosed sourceMultimodal
Source

Muse Spark

MetaMilestone

Meta introduces a new family combining multimodality, reasoning, and tools.

ReleaseClosed sourceMultimodal
Source

GLM improves long-running engineering, planning, and execution.

ReleaseOpen sourceText
Source

Gemma 4

Google

Gemma advances open-model reasoning and agent capabilities.

Open weightsOpen sourceMultimodal
Source

MiMo expands flagship reasoning and multimodal understanding.

ReleaseClosed sourceMultimodal
Source

Software engineering and professional work gain stronger agent collaboration.

ReleaseOpen sourceText
Source

One model combines reasoning, vision, and coding with adjustable thinking.

Open weightsOpen sourceMultimodal
Source

Hybrid experts and long context support multi-agent work.

Open weightsOpen sourceText
Source

GPT-5.4

OpenAI

Reasoning, coding, and native computer use target professional work.

ReleaseClosed sourceMultimodal
Source

Gemini strengthens reasoning for scientific, engineering, and complex tasks.

PreviewClosed sourceMultimodal
Source

Sonnet improves coding, visual interaction, planning, and knowledge work.

ReleaseClosed sourceMultimodal
Source

Qwen3.5

AlibabaMilestone

Native multimodality and agent capabilities come together in a large expert model.

AnnouncementOpen sourceMultimodal
Source

Seed 2.0

ByteDanceMilestone

ByteDance introduces multimodal models for real-world agent tasks.

ReleaseClosed sourceMultimodal
Source

GLM-5

Z.aiMilestone

GLM moves from individual coding tasks toward broader software engineering.

AnnouncementOpen sourceText
Source

Reinforcement learning targets practical coding, search, office work, and tools.

ReleaseOpen sourceText
Source

Opus expands complex codebase analysis and long-running agent tasks.

ReleaseClosed sourceMultimodal
Source

Coding and general reasoning combine in longer engineering workflows.

ReleaseClosed sourceMultimodal
Source

Sparse mixture-of-experts modeling balances speed with agent capabilities.

Open weightsOpen sourceText
Source

Kimi K2.5

Moonshot AIMilestone

Kimi combines visual understanding, coding, and agent collaboration.

Open weightsOpen sourceMultimodal
Source

A lightweight model extends the GLM-4.7 family.

ReleaseOpen sourceText
Source

2025

Reasoning models and tool use42 entries

MiniMax improves programming across languages and practical office tasks.

ReleaseOpen sourceText
Source

GLM improves coding and long-context task execution.

ReleaseOpen sourceText
Source

Gemini 3 reasoning reaches a faster, more efficient Flash tier.

ReleaseClosed sourceMultimodal
Source

Xiaomi opens an efficient model for reasoning, coding, and agent tasks.

Open weightsOpen sourceText
Source

NVIDIA announces Nemotron 3 and releases Nano first.

Open weightsOpen sourceText
Source

GPT-5.2

OpenAI

GPT advances professional knowledge work, coding, and long-context tasks.

ReleaseClosed sourceMultimodal
Source

Mistral 3

Mistral

Mistral expands its open-weight multimodal model family.

Open weightsOpen sourceMultimodal
Source

Configurable extended thinking supports reasoning, coding, and agent workflows.

ReleaseClosed sourceMultimodal
Source

DeepSeek advances efficient attention, reasoning, and agentic tool use.

ReleaseOpen sourceText
Source

Opus improves coding, computer use, and long-running work.

ReleaseClosed sourceMultimodal
Source

Gemini 3 Pro

GoogleMilestone

Google introduces a new generation of multimodal reasoning models.

PreviewClosed sourceMultimodal
Source

GPT-5.1

OpenAI

Instant and Thinking improve instructions, style, and adaptive reasoning.

ReleaseClosed sourceMultimodal
Source

Kimi K2 Thinking

Moonshot AI

Kimi extends the K2 family with deeper reasoning and tool-assisted work.

Open weightsOpen sourceText
Source

MiniMax-M2

MiniMax

MiniMax opens a model designed around coding and agent workflows.

Open weightsOpen sourceText
Source

A smaller Claude model improves coding and computer use at lower latency.

ReleaseClosed sourceMultimodal
Source

GLM improves code, reasoning, and tool calls.

ReleaseOpen sourceText
Source

Sonnet advances coding, computer use, and sustained agentic tasks.

ReleaseClosed sourceMultimodal
Source

Qwen3-Next

Alibaba

Hybrid attention and highly sparse activation target efficient long-context processing.

AnnouncementOpen sourceText
Source

Thinking and non-thinking modes are unified with stronger agent capabilities.

ReleaseOpen sourceText
Source

GPT-5

OpenAIMilestone

A new GPT generation advances reasoning, coding, and everyday work.

ReleaseClosed sourceMultimodal
Source

Opus improves coding, practical reasoning, and agent performance.

ReleaseClosed sourceMultimodal
Source

gpt-oss-120b / 20b

OpenAIMilestone

OpenAI releases two open-weight reasoning models.

Open weightsOpen sourceText
Source

GLM-4.5

Z.aiMilestone

GLM combines reasoning, coding, and agentic workflows.

Open weightsOpen sourceText
Source

An open mixture-of-experts model targets agentic coding and long tasks.

Open weightsOpen sourceText
Source

Kimi K2

Moonshot AIMilestone

Kimi introduces open weights focused on coding and agentic tool use.

Open weightsOpen sourceText
Source

xAI advances reasoning and tool-assisted problem solving.

ReleaseClosed sourceMultimodal
Source

MiniMax-M1

MiniMaxMilestone

Hybrid attention explores efficient reasoning and long-context work.

Research paperOpen sourceText
Source

A major R1 update improves reasoning, coding, structured output, and tool calls.

ReleaseOpen sourceText
Source

An enterprise model balances multimodal capability, coding, and deployment cost.

ReleaseClosed sourceMultimodal
Source

Qwen3

AlibabaMilestone

Thinking and non-thinking modes share one model family.

Open weightsOpen sourceText
Source

OpenAI o3 / o4-mini

OpenAIMilestone

Reasoning models learn to combine search, code, and visual tools.

ReleaseClosed sourceMultimodal
Source

GPT-4.1

OpenAI

An API model family improves coding, instructions, and long-context processing.

ReleaseClosed sourceMultimodal
Source

Llama 4

MetaMilestone

Scout and Maverick introduce multimodal mixture-of-experts models.

Open weightsOpen sourceMultimodal
Source

Gemini 2.5 Pro

GoogleMilestone

Thinking becomes central to complex reasoning, coding, and multimodal tasks.

PreviewClosed sourceMultimodal
Source

Hunyuan T1

Tencent

Tencent introduces a deep-reasoning model trained with reinforcement learning.

ReleaseClosed sourceText
Source

Gemma 3

Google

The open-weight family expands image understanding and long-context processing.

Open weightsOpen sourceMultimodal
Source

GPT-4.5

OpenAI

A research preview explores further gains from scaling pretraining.

PreviewClosed sourceMultimodal
Source

Claude 3.7 Sonnet

AnthropicMilestone

One model combines fast responses with extended thinking.

ReleaseClosed sourceMultimodal
Source

xAI describes reasoning capabilities and the DeepSearch experience.

AnnouncementClosed sourceMultimodal
Source

Efficient reasoning targets mathematics, science, and programming.

ReleaseClosed sourceText
Source

DeepSeek-R1

DeepSeekMilestone

DeepSeek releases a reasoning model and distilled variants with downloadable weights.

Open weightsOpen sourceText
Source

2024

Multimodality and reasoning22 entries

DeepSeek-V3

DeepSeekMilestone

Open weights expand large-scale mixture-of-experts modeling.

Open weightsOpen sourceText
Source

Phi-4

Microsoft

Synthetic data and training improvements strengthen small-model reasoning.

ReleaseOpen sourceText
Source

Gemini 2.0 Flash

GoogleMilestone

An experimental Flash model combines multimodality with native tool use.

PreviewClosed sourceMultimodal
Source

A revised 70B instruction model improves multilingual dialogue and deployment efficiency.

Open weightsOpen sourceText
Source

OpenAI o1

OpenAI

The full o1 model launches with ChatGPT Pro and adds image understanding.

ReleaseClosed sourceMultimodal
Source

Amazon introduces text and multimodal models across several cost tiers.

ReleaseClosed sourceMultimodal
Source

Vision models and small on-device text models join the Llama family.

Open weightsOpen sourceMultimodal
Source

Qwen2.5

Alibaba

General, coding, and math models improve instruction following and structured output.

Open weightsOpen sourceText
Source

OpenAI o1-preview

OpenAIMilestone

Reasoning-time computation becomes a new route to stronger task performance.

PreviewClosed sourceText
Source

A 123B model expands coding, multilingual, and long-context capabilities.

Open weightsOpen sourceText
Source

Llama 3.1

MetaMilestone

A 405B model extends the family with longer context, multilingual support, and tool use.

Open weightsOpen sourceText
Source

A lower-cost text and vision model expands lightweight applications.

ReleaseClosed sourceMultimodal
Source

Gemma 2

Google

A redesigned architecture introduces efficient 9B and 27B open-weight models.

Open weightsOpen sourceText
Source

Claude 3.5 Sonnet

AnthropicMilestone

Sonnet improves coding and visual understanding alongside the introduction of Artifacts.

ReleaseClosed sourceMultimodal
Source

Qwen2

Alibaba

A range of model sizes improves multilingual, coding, math, and long-context performance.

Open weightsOpen sourceText
Source

GPT-4o

OpenAIMilestone

Text, vision, and audio move toward a unified, more natural interface.

ReleaseClosed sourceMultimodal
Source

DeepSeek-V2

DeepSeekMilestone

Mixture-of-experts and latent attention improve training and inference efficiency.

Open weightsOpen sourceText
Source

Phi-3

Microsoft

Microsoft introduces small language models suited to local deployment.

ReleaseOpen sourceText
Source

Llama 3

MetaMilestone

8B and 70B weights improve general-purpose open-model capabilities.

Open weightsOpen sourceText
Source

Claude 3

AnthropicMilestone

Opus, Sonnet, and Haiku form a new vision-capable model family.

ReleaseClosed sourceMultimodal
Source

Gemma

GoogleMilestone

Google introduces lightweight open-weight models for local development and research.

Open weightsOpen sourceText
Source

Gemini 1.5

GoogleMilestone

Gemini 1.5 Pro demonstrates million-token context processing.

PreviewClosed sourceMultimodal
Source

2023

Major model families13 entries

A sparse mixture-of-experts model expands efficient open-weight inference.

AnnouncementOpen sourceText
Source

Gemini 1.0

GoogleMilestone

Google introduces the multimodal Gemini family: Ultra, Pro, and Nano.

ReleaseClosed sourceMultimodal
Source

A developer preview expands context to 128K and advances vision and tool use.

PreviewClosed sourceMultimodal
Source

Grok

xAI

xAI announces Grok with early access and information from X.

PreviewOpen sourceText
Source

Mistral 7B

MistralMilestone

An efficient 7B model is released under Apache 2.0.

Open weightsOpen sourceText
Source

Qwen-7B

AlibabaMilestone

Qwen releases 7B base and chat model weights.

Open weightsOpen sourceText
Source

Llama 2

MetaMilestone

Pretrained and chat-tuned weights become available for research and commercial use.

Open weightsOpen sourceText
Source

Claude 2

Anthropic

Improved coding, reasoning, and long-document processing arrive with a public chat beta.

ReleaseClosed sourceText
Source

LLaVA

LLaVA TeamMilestone

Visual instruction tuning connects an image encoder and a language model for multimodal conversation.

Research paperOpen sourceMultimodal
Source

Claude

AnthropicMilestone

Anthropic introduces Claude and Claude Instant.

ReleaseClosed sourceText
Source

GPT-4

OpenAIMilestone

Stronger language capabilities are joined by demonstrated image understanding.

ReleaseClosed sourceMultimodal
Source

Stanford Alpaca

StanfordMilestone

Synthetic instruction data adapts LLaMA into a small instruction-following model.

AnnouncementOpen sourceText
Source

LLaMA

MetaMilestone

Meta introduces a family of foundation models for research access.

AnnouncementOpen sourceText
Source

2022

Instruction alignment and chat products4 entries

ChatGPT

OpenAIMilestone

A GPT-3.5-based research preview supports multi-turn conversation.

Product launchClosed sourceText
Source

BLOOM

BigScienceMilestone

BigScience releases a 176B multilingual model with weights and training artifacts.

Open weightsOpen sourceText
Source

PaLM

Google

Google explores reasoning and language capabilities at larger scale.

AnnouncementClosed sourceText
Source

InstructGPT

OpenAIMilestone

Reinforcement learning from human feedback improves instruction following and alignment with user intent.

AnnouncementClosed sourceText
Source

2021

Code models1 entry

Codex

OpenAIMilestone

Natural-language instructions become executable code.

PreviewClosed sourceText
Source

2020

Few-shot learning1 entry

GPT-3

OpenAIMilestone

A 175B-parameter model demonstrates broad few-shot learning.

Research paperClosed sourceText
Source

2019

Scale and transfer learning2 entries

T5

GoogleMilestone

A text-to-text formulation unifies language tasks within one framework.

Research paperOpen sourceText
Source

GPT-2

OpenAIMilestone

Scaling improves coherent text generation and zero-shot task performance.

ReleaseOpen sourceText
Source

2018

Pretraining research2 entries

BERT

GoogleMilestone

Bidirectional context advances language understanding and reading comprehension.

Research paperOpen sourceText
Source

GPT-1

OpenAIMilestone

Unsupervised language pretraining and task-specific fine-tuning support transfer across tasks.

Research paperOpen sourceText
Source

2017

The Transformer architecture1 entry

Transformer

GoogleMilestone

Google researchers introduce the Transformer, an attention-based sequence model, in Attention Is All You Need.

Research paperOpen sourceText
Source

All 155 entries shown

Search names, organizations, or keywords across all timelines.