Major large-language-model releases plotted by public release date (x) and capability (y). Capability is the Artificial Analysis Intelligence Index, a 0–100 aggregate of hard reasoning/coding/agentic benchmarks. Each point is the releasing company's logo.
⚠︎ Read the y-axis carefully. Scores are the current Intelligence Index v4.1 (July 2026 snapshot). v4.1 is far harder than the index versions in use when older models launched, so a model's number here is well below its launch-day headline (e.g. Grok 4 was ~68 at launch → 33 on v4.1). All solid markers are measured on v4.1. Pre-2025 models predate v4.1 and are shown at estimated positions (dashed ring) so the timeline still starts in 2022. Microsoft's MAI models have no published AA score and are shown at estimated positions too — flagged below and in tooltips.
| Model | Company | Released | Index | Basis | Notes / source |
|---|