Product Timeline
A chronological timeline of significant AI product releases and milestones — including the race between closed, API-only frontier models and the open-weight models chasing (and sometimes matching) them.
Early Foundations (2010-2015)
IBM Watson wins Jeopardy!
Apple releases Siri
Google's deep learning algorithm learns to recognize cats in YouTube videos
Amazon releases Alexa
DeepMind's AlphaGo begins development
Microsoft releases Cortana
Rise of Modern AI (2016-2019)
Google releases TensorFlow to the public
AlphaGo defeats Lee Sedol at Go
Google introduces Transformer architecture
Google releases BERT
OpenAI releases GPT-1
Google launches ALBERT
OpenAI releases GPT-2
Language Model Revolution (2020-2022)
OpenAI releases GPT-3
Google releases T5
Google introduces LaMDA
DeepMind releases AlphaFold 2
OpenAI releases DALL-E 2
Google releases PaLM
Stability AI releases Stable Diffusion
Meta releases OPT-175B
AI Goes Mainstream (2023)
Microsoft invests heavily in OpenAI
Google announces Bard
OpenAI releases GPT-4
ClosedAnthropic releases Claude
ClosedMeta releases Llama 2 — first widely-licensed open-weight model to seriously rival a closed frontier model
OpenStability AI releases SDXL
OpenOpenAI releases GPT-4 Turbo
ClosedGoogle releases Gemini
ClosedOpen Models Enter the Frontier Race (2024)
Google releases Gemini Ultra
ClosedGoogle releases Gemma (2B, 7B) — lightweight open models distilled from Gemini research
OpenxAI open-sources the Grok-1 weights
OpenMistral releases Mixtral 8x22B and the closed Mistral Large
OpenClosedOpenAI releases Sora
ClosedAnthropic releases the Claude 3 family (Haiku, Sonnet, Opus)
ClosedCognition Labs launches Devin, marketed as the "first AI software engineer" — an early sign coding agents, not just chat models, were becoming the next battleground
ClosedMeta releases Llama 3 (8B, 70B)
OpenOpenAI releases GPT-4o
ClosedAnthropic releases Claude 3.5 Sonnet
ClosedMeta releases Llama 3.1 405B — the first open-weight model to benchmark competitively against GPT-4o and Claude 3.5 Sonnet
OpenxAI releases Grok-2
ClosedAlibaba releases Qwen2.5
OpenOpenAI releases o1-preview / o1-mini — the first mainstream "reasoning" model, trading extra inference-time compute for accuracy
ClosedAnthropic releases an upgraded Claude 3.5 Sonnet and Claude 3.5 Haiku, with computer-use capability
ClosedDeepSeek releases DeepSeek-V3 — a frontier-class model trained for a small fraction of the cost of comparable Western models
OpenOpenAI releases the full o1; Google releases Gemini 2.0 Flash
ClosedMeta releases Llama 3.3 70B
OpenThe Open-Weight Shock & Reasoning Race (2025)
DeepSeek releases R1 — matches OpenAI's o1 on reasoning benchmarks at a small fraction of the training/inference cost, triggering a major AI/chip stock selloff (the "DeepSeek moment") and proving open models could match closed reasoning models
OpenOpenAI releases o3-mini
ClosedAnthropic launches **Claude Code** as a research preview — an agentic command-line tool that lets Claude read, edit, run, and commit code directly in a developer's terminal
ClosedAnthropic releases Claude 3.7 Sonnet — a hybrid model with switchable "extended thinking"
ClosedxAI releases Grok-3
ClosedGoogle releases Gemma 3 and Gemini 2.5 Pro; OpenAI ships Codex CLI/agent tooling to compete with Claude Code
OpenClosedMeta releases Llama 4 (Scout, Maverick); OpenAI releases o3 and o4-mini
OpenClosedAnthropic releases the Claude 4 family (Opus 4, Sonnet 4) alongside the **general availability of Claude Code** — Claude Code goes on to grow more than 10x in usage within three months
ClosedxAI releases Grok 4
ClosedOpenAI releases GPT-5 — unifies the reasoning and non-reasoning model lines into a single system for the first time
ClosedDeepSeek releases V3.1 Terminus and V3.2
OpenAlibaba releases the Qwen3 family, alongside continued open releases from Mistral and Z.ai's GLM; Cursor's AI-native IDE becomes the fastest-growing SaaS product ever recorded, reaching roughly $2B ARR in about 28 months
OpenClosedGoogle releases Gemini 3; xAI releases Grok 4.1
ClosedAn open-source personal AI agent launches under the name Clawdbot (soon renamed Moltbot, then **OpenClaw**) — an MIT-licensed, local-first agent that lives in chat apps like WhatsApp, Slack, and Telegram and can browse the web, manage files, and run commands on its own schedule
OpenOpenAI ships fast follow-ups GPT-5.1 and GPT-5.2
ClosedFrontier Consolidation (2026)
OpenClaw crosses 100,000 GitHub stars within its first week under its new name — one of the fastest-growing open-source projects ever, outpacing the early growth of Docker, Kubernetes, and React
OpenClaw surpasses 214,000 GitHub stars; Anthropic reports Claude Code's run-rate revenue has climbed past $2.5 billion
ClosedNew Qwen generations (Qwen 3.5, 3.6) ship under the permissive Apache 2.0 license; Google shifts Gemma 4 to Apache 2.0 as well
OpenAnthropic publicly releases **Claude Fable 5** — the first public model from its more powerful "Mythos" family, state-of-the-art on most capability benchmarks, shipped with safety classifiers that block high-risk cyber/bio/chem requests
ClosedDays after Fable 5's release, the US government briefly restricts its export over cybersecurity jailbreak concerns; Anthropic redeploys it globally days later with strengthened safeguards
DeepSeek releases V4 — MIT-licensed, with ultra-long context as a default
OpenAnthropic releases Claude Opus 5, topping both the Intelligence and Agentic indexes on Artificial Analysis
ClosedOpenAI's GPT-5.6 reaches general availability
ClosedGoogle releases Gemini 3.7 Flash; Z.ai releases GLM-5.3
OpenClosedxAI ships Grok 4.6
ClosedIndependent surveys put Claude Code at 39% adoption among professional developers worldwide (47% in the US), up from 18% in January — the fastest-growing of the major coding agents, ahead of OpenAI's Codex
Language Model Revolution (2020-2022)
Midjourney opens its Discord beta, putting diffusion image generation in front of a mass audience
ClosedStability AI releases Stable Diffusion, the first open text-to-image model good enough to run on a consumer GPU
OpenOpenAI releases Whisper, an open speech-recognition model trained on 680,000 hours of weakly labelled audio
OpenAI Goes Mainstream (2023)
ElevenLabs ships a text-to-speech service whose zero-shot voice cloning sets the commercial bar
ClosedSuno brings full-song music generation to a browser prompt
ClosedOpen Models Enter the Frontier Race (2024)
Black Forest Labs, founded by the original Stable Diffusion authors, releases FLUX.1, a rectified-flow transformer that leads open image generation
OpenThe Open-Weight Shock & Reasoning Race (2025)
Google's Veo 3 generates native synchronized dialogue and sound effects in the same pass as the video
ClosedAlibaba releases Qwen-Image, a large MMDiT text-to-image model with a Qwen2.5-VL text encoder and strong text rendering
OpenFrontier Consolidation (2026)
OpenAI ships Sora 2, a video generator with synchronized audio framed as a world model, wrapped in a social app
Closed(2026 entries are drawn from AI-news aggregators rather than primary announcements, since they fall after most models' training cutoffs — treat exact dates/version numbers as approximate.)
Anthropic’s Growth & the Rise of Agents
- Anthropic’s revenue exploded. From roughly $1 billion in annualized revenue in December 2024, Anthropic grew to $9B by end of 2025, $14B in February 2026, $30B in April 2026, $47B in May 2026, and $65B by July 2026 — CEO Dario Amodei described it as 80x annualized growth in a single quarter.
- Its valuation followed. $4.1B (Mar 2023) → $61.5B (Mar 2025) → $380B (Feb 2026) → $965B after a $65B Series H (May 2026), briefly making Anthropic the world’s most valuable private AI company, ahead of OpenAI.
- Claude Code became a business on its own. Launched as a research preview in February 2025 and reaching general availability in May 2025, Claude Code grew more than 10x in usage within three months, crossed a $2.5B revenue run-rate by February 2026, and helped double Anthropic’s $1M+ enterprise customers to over 1,000.
- Agentic coding tools became the main battleground. Devin (2024), Claude Code, OpenAI’s Codex CLI, Cursor, and GitHub Copilot all raced to move beyond autocomplete into autonomously writing features, fixing bugs, running tests, and opening pull requests — by mid-2026 roughly 90% of developers used at least one AI coding tool at work.
- Personal, always-on agents went mainstream outside coding too. OpenClaw (born as Clawdbot/Moltbot in late 2025) showed the same agentic pattern — an LLM given tools, memory, and a persistent loop — could run entirely on personal hardware and act inside everyday chat apps, not just in a dev terminal.
- Anthropic pushed a more capable, more restricted tier. The “Mythos” family and its public release, Claude Fable 5 (June 2026), showed frontier labs beginning to ship models powerful enough to trigger their own safety classifiers and, briefly, government export scrutiny — a preview of how the next capability jump may be gated differently than prior ones.
How the Coding Interface Itself Evolved
The model race gets the headlines, but the bigger day-to-day change for developers has been how they talk to AI while coding — from predicting the next few keystrokes to handing off an entire goal and walking away.
Era 1 — Inline autocomplete (2021–2022): AI finishes your line
- 2021 Jun: GitHub launches Copilot, trained with OpenAI — ghost-text suggestions that complete a line or block as you type, with no chat and no memory beyond the open file
- 2022 Jun: Copilot reaches general availability in VS Code, establishing “autocomplete-as-you-type” as the default mental model for AI coding help
Era 2 — Chat enters the editor (2022–2023): AI answers questions about your code
- GitHub adds Copilot Chat, a sidebar you can ask questions in, but changes still land as suggestions you accept line by line
- 2023 Mar: Anysphere launches Cursor, forking VS Code to put an AI chat panel and its own “Tab” autocomplete directly into the editor with whole-repo context — you’re now talking about the codebase, not just accepting predicted text
Era 3 — The agent loop is invented, outside the IDE (2023): AI is given a goal, not a prompt
- 2023 Mar: AutoGPT and BabyAGI popularize the reason → act → observe loop — give the model a goal, let it choose its own steps, call tools, and keep going until done. Neither was a coding tool, but every later coding agent copies this pattern
- AutoGPT passes 100,000 GitHub stars within weeks, showing developers how much appetite there was for AI that acts instead of just suggesting
Era 4 — Multi-file “Composer” and agent mode arrive in coding tools (2024–2025): AI edits, not suggestions
- Cursor ships Composer, letting the model plan and edit across multiple files in one pass instead of one accepted suggestion at a time
- 2024 (late): Cursor’s “YOLO mode” lets the agent run terminal commands and refactor across a codebase without approving every single step — the first mainstream “let it just work” toggle
- 2024 Mar: Devin markets itself as an autonomous “AI software engineer” that takes a goal and works largely unsupervised
- 2025 Feb: Anthropic launches Claude Code — no IDE at all, just a terminal and a natural-language goal, with the agent reading, editing, running, and committing code in its own loop
- 2025 Feb: GitHub formalizes Agent Mode inside VS Code Copilot, letting it execute multi-step tasks rather than only suggest edits
- 2025 Feb: Andrej Karpathy coins “vibe coding” — describing handing a goal to an agent loop (e.g., Cursor Composer) and “forgetting the code even exists” — later naming the more deliberate version of this practice “agentic engineering”
Era 5 — Background and cloud agents (2025–2026): AI runs the loop without you watching
- 2025 Oct: Cursor 2.0 ships Composer, a coding model purpose-built for low-latency agent loops rather than one-shot completions
- 2026 Feb: Cursor 2.5 adds async subagents — multiple agent loops running in parallel on different parts of one task
- 2026 Apr–May: Cursor 3.0/3.5 move agents into a dedicated Agents Window with Cloud Agents that run in an isolated VM — you hand off a goal, close your laptop, and come back to a finished, tested pull request
The throughline: the unit of AI collaboration kept growing — keystroke → line → file → whole repo → a goal handed to a loop → a goal handed to a cloud agent you check on later. Each era didn’t just add a feature, it changed what a developer’s attention was actually needed for.
Open vs. Closed: The State of the Race
- The gap has shrunk. Open-weight models trailed closed frontier models by roughly two years in the GPT-3 era; by 2025–2026 that gap had narrowed to a matter of months, with DeepSeek-R1 and Llama 4 landing within striking distance of the latest o-series and Claude/Gemini releases on many benchmarks.
- Training costs collapsed. DeepSeek’s efficiency-focused training (V3, R1) showed frontier-class results were achievable for a fraction of the compute cost previously assumed necessary, reshaping assumptions across the industry.
- Licensing is trending more open. Newer open releases (DeepSeek V4, Gemma 4, the open Qwen line) have moved toward permissive licenses like Apache 2.0 and MIT, rather than the more restrictive custom licenses used by early Llama releases.
- It’s increasingly a geographic race too. US labs (OpenAI, Anthropic, Google, xAI) still lead most closed frontier benchmarks, while Chinese labs (DeepSeek, Alibaba’s Qwen, Z.ai’s GLM) have become the dominant force in high-quality open-weight releases — making “open vs. closed” partly a proxy for a broader US–China AI competition.
Last updated August 2026