A new multimodal model for complex coding, knowledge work, and cyber defense, initially available to invited teams.
Language model timeline
Entries by year
Release history
2026
Multimodality and agents67 entriesImproves coding, computer use, and professional work, with task delegation available within a request.
The second Claude 5.5 model supports coding, document processing, and office tasks.
MiMo combines multimodal understanding with large-scale reinforcement learning.
A flagship preview targets long-running engineering, knowledge work, and research.
Improves engineering projects, multi-tool orchestration, chart reasoning, and document parsing.
Native vision and hybrid linear/sparse attention support coding, browser, and GUI tasks.
V4 Pro reaches general availability with stronger agents and Responses API support.
Flash improves reasoning efficiency alongside a lighter Flash-Lite model.
2025
Reasoning models and tool use42 entriesConfigurable extended thinking supports reasoning, coding, and agent workflows.
Hybrid attention and highly sparse activation target efficient long-context processing.
A major R1 update improves reasoning, coding, structured output, and tool calls.
An enterprise model balances multimodal capability, coding, and deployment cost.
DeepSeek releases a reasoning model and distilled variants with downloadable weights.
2024
Multimodality and reasoning22 entriesReasoning-time computation becomes a new route to stronger task performance.
Sonnet improves coding and visual understanding alongside the introduction of Artifacts.
Mixture-of-experts and latent attention improve training and inference efficiency.
2023
Major model families13 entriesSynthetic instruction data adapts LLaMA into a small instruction-following model.
2022
Instruction alignment and chat products4 entriesReinforcement learning from human feedback improves instruction following and alignment with user intent.
2021
Code models1 entry2020
Few-shot learning1 entry2019
Scale and transfer learning2 entries2018
Pretraining research2 entries2017
The Transformer architecture1 entryGoogle researchers introduce the Transformer, an attention-based sequence model, in Attention Is All You Need.
All 155 entries shown