The LLM Era
2020s–Present · Large language models and agentic systems
Since 2020, large language models such as GPT and Claude have developed rapidly, alongside technologies including AI agents and MCP. These advances have broadened the use of AI in content generation, software development, and tool use.
Technologies such as Claude Code, agents, and MCP allow models to edit files, run commands, and connect to external tools within granted permissions. Their capabilities and risks depend on the model, host environment, permission settings, and product version.
Representative characteristics of this period include larger models, multimodal input and output, agentic workflows, and standard protocols that connect AI to external systems. These capabilities continue to change and should be assessed against verifiable product documentation and research.
Key Milestones
GPT-3 Release
OpenAI released GPT-3, a large language model with 175 billion parameters. GPT-3 demonstrated astonishing language generation capabilities, sparking widespread discussion about AGI.
OpenAI
DALL-E & CLIP
OpenAI released DALL-E (image generation) and CLIP (image-text contrastive learning). These models demonstrated AI’s ability to understand and generate multimodal content.
OpenAI
Codex & GitHub Copilot
OpenAI released Codex, an AI model that understands and generates code. GitHub Copilot became the first widely-used AI programming assistant.
OpenAI Codex
ChatGPT Release
In November 2022, OpenAI released ChatGPT, the first mass-market generative-AI conversational assistant. ChatGPT reached 100 million users in two months, becoming the fastest-growing consumer product in history.
OpenAI ChatGPT
Stable Diffusion Open Source
Stability AI released Stable Diffusion, an open-source image generation model. Open-source community activity rapidly popularized image generation technology.
Stability AI
GPT-4 Release
OpenAI released GPT-4, a multimodal model capable of processing images and text. GPT-4 showed near-human performance on professional and academic benchmarks.
OpenAI
Claude Release
Anthropic released Claude, an AI assistant focused on safety and helpfulness. Claude’s design reflects emphasis on AI safety—AI should be helpful, harmless, and honest.
Anthropic Claude
LLaMA Open Source
Meta released LLaMA, a large language model under a non-commercial research license. Although LLaMA has relatively small parameter size, its excellent performance promoted open-source LLM development.
Meta
AI Agent Concept Rises
The AI Agent concept began to receive widespread attention. Agents are AI systems capable of autonomously planning and executing complex tasks.
AI Industry
Gemini 1.5 and Long Context
In February 2024, Google introduced Gemini 1.5 with a mixture-of-experts architecture and an experimental context window of up to one million tokens, demonstrating retrieval and reasoning across long documents, codebases, and video.
Google DeepMind
Claude 3 Series
Anthropic released the Claude 3 series, including Opus, Sonnet, and Haiku. Claude 3 Opus performed competitively with GPT-4 on several benchmarks including MMLU and GPQA, while trailing GPT-4 on others.
Anthropic
Llama 3 Open Source
Meta released the Llama 3 family with open weights, including 8B and 70B parameter models. The later Llama 3.1 405B was the first open-weights model to match frontier closed-source competitors on common benchmarks.
Meta
EU AI Act Enacted
The European Union enacted the AI Act, the first comprehensive horizontal regulation of artificial intelligence. The Act established a risk-based classification and strict requirements for high-risk AI systems.
European Union
OpenAI o1 Model
OpenAI released the o1 model, a large model with reasoning capabilities. o1 performs excellently on scientific and programming tasks, demonstrating "thinking" ability.
OpenAI
Claude Computer Use Public Beta
In October 2024, Anthropic introduced a public beta of computer use for an upgraded Claude 3.5 Sonnet. Developers could direct Claude to inspect a screen, move a cursor, click, and type, making it one of the first frontier models publicly offered with general graphical-interface control.
Anthropic Claude
MCP Protocol Release
Anthropic released MCP (Model Context Protocol), an open standard protocol for connecting AI models with external tools and data sources. MCP is becoming the "USB-C" of AI applications.
Anthropic
Agentic Ecosystem Matures
By 2025, AI agents matured from research demos to production tools. Coding agents, browser-use agents, and multi-step task agents became standard offerings, with OpenAI, Anthropic, Google, and open-source projects all shipping capable systems.
AI Industry
DeepSeek R1 Release
Chinese AI lab DeepSeek released R1, an open-source reasoning model whose performance rivaled OpenAI’s o1. The release demonstrated that frontier capabilities could be achieved outside the US AI establishment and triggered a global re-evaluation of AI competition.
DeepSeek
Stargate Project Launched
OpenAI, Oracle, SoftBank, and the UAE’s MGX jointly announced a commitment of $500 billion to build AI infrastructure in the United States. The project plans to construct multiple large-scale data centers over four years, the largest private-sector investment in AI history.
OpenAI / Oracle / SoftBank
Claude 3.7 Sonnet Release
Anthropic released Claude 3.7 Sonnet, the first model to support "hybrid reasoning"—users can switch between instant answers and extended thinking in the same model. Claude 3.7 marked the beginning of consolidation in AI reasoning product forms.
Anthropic
Anthropic $3.5B Funding Round
In March 2025, Anthropic announced a $3.5 billion funding round at a $61.5 billion valuation, making it one of the highest-valued AI startups in the world and reflecting further capital concentration in the AI industry.
Anthropic
Llama 4 Open-Source Release
Meta released the natively multimodal open-weight Llama 4 Scout and Maverick models and previewed Behemoth, a teacher model that was still training and was not released. Scout and Maverick use mixture-of-experts architectures.
Meta
Claude 4 Release
Anthropic released the Claude 4 family (Opus 4 and Sonnet 4), introducing Extended Thinking capability with significant improvements in programming, reasoning, and long-horizon tasks. Claude 4 consolidated Anthropic’s position in the enterprise AI market.
Anthropic
EU AI Act GPAI Obligations Begin to Apply
On 2 August 2025, the EU AI Act’s governance rules and obligations for general-purpose AI models became applicable. Other provisions follow separate transition dates, so this milestone is not a blanket start of full enforcement.
EU AI Office
GPT-5 Release
OpenAI released GPT-5, a unified flagship model that integrates the reasoning capabilities of the o-series with the conversational abilities of the GPT series. GPT-5 set a new frontier in programming, mathematics, and reasoning while supporting longer contexts and tool use.
OpenAI
OpenAI Completes Recapitalization
OpenAI completed the recapitalization of its corporate structure. The for-profit business became a public benefit corporation named OpenAI Group PBC, and the OpenAI Foundation — the nonprofit — continues to control it under the same mission. OpenAI says the recapitalization followed nearly a year of discussions with the California and Delaware Attorneys General.
OpenAI
Nature Publishes the Co-Scientist Multi-Agent System
Nature published Co-Scientist, a multi-agent system built on Gemini that formulates novel research hypotheses for experimental verification. The paper reports validation across three biomedical applications: drug repurposing, discovery of novel targets, and explaining mechanisms of antimicrobial resistance. The authors present it as an assistant to researchers rather than a replacement, which makes it peer-reviewed evidence for AI participating in hypothesis generation rather than only in analysis.
Scientific Community
Meta Begins Unwinding the Manus Acquisition
Reporting on June 11, 2026 said Meta had begun unwinding its acquisition integration with AI agent company Manus, halting data sharing and separating operations. The report linked the move to Chinese regulatory requirements; later legal and operational status should be checked against updated disclosures.
Meta / Manus
Claude Fable 5 and Mythos 5 Suspended, Then Redeployed
Anthropic launched safeguarded Claude Fable 5 for general use and the less-restricted Mythos 5 for vetted Project Glasswing partners on June 9, 2026. On June 12, U.S. export controls required restrictions on foreign nationals; unable to verify nationality in real time, Anthropic suspended both models globally. The controls were lifted on June 30. Fable 5 returned globally on July 1, while Mythos 5 initially returned for approved U.S. organizations.
Anthropic
US State Attorneys General Open Probe into OpenAI
On June 13, 2026, a coalition of US state attorneys general led by New York opened an investigation into OpenAI and issued a subpoena. The subpoena demands documents relating to advertising placement, user engagement and retention, model sycophancy, consumer and health data handling, and protections for minors and the elderly. OpenAI said it would cooperate with the attorneys general and noted that ChatGPT has built stronger safeguards for minors. This is the first coordinated multistate coalition investigation of a leading AI company by US state attorneys general.
Coalition of US State Attorneys General
GPT-5.6 Begins with a Government-Requested Limited Preview
OpenAI began a limited preview of GPT-5.6 Sol, Terra, and Luna on June 26, 2026 with a small group of trusted partners at the U.S. government’s request, while saying this process should not become the long-term default. OpenAI made the GPT-5.6 family generally available on July 9.
OpenAI
Samsung and SK Hynix Announce Major Multi-Year Memory Investment Plans
On 29 June 2026, Samsung and SK Hynix announced multi-year semiconductor investment plans totaling 800 trillion won and involving four memory fabs in southwestern South Korea. Dollar conversions vary with exchange rates, while scope, approvals, and execution remain subject to review.
Samsung / SK Hynix
Gemini Spark Lands on Mac as a Desktop Agent
On July 1, 2026 Google released Gemini Spark — its 24/7 agentic AI assistant — for macOS, with real-time tracking and broader app integration. Spark can take persistent action across a user’s digital life, marking Google’s push to bring agentic AI onto the consumer desktop.
Google DeepMind
Moonshot Releases Kimi K3 Open Weights
In July 2026 Moonshot AI published open weights for Kimi K3 on Hugging Face. The model card records 2.8 trillion total parameters with 104 billion activated, a 1,048,576-token context window, and the Kimi K3 License. It is the largest openly downloadable language model published to date, and it places an Asian laboratory at the frontier of open-weight releases.
Moonshot AI
Cloudflare Classifies AI Crawlers by Purpose
On 1 July 2026, Cloudflare announced separate controls for Search, Agent, and Training traffic. From 15 September, new Cloudflare domains will block Training and Agent crawlers by default on ad-bearing pages while allowing Search by default; site owners can change the settings, and multi-purpose crawlers follow the strictest applicable rule.
Cloudflare
Microsoft Launches $2.5B AI Deployment Company
On July 2, 2026 Microsoft announced Microsoft Frontier Company, a new operating business focused on enterprise AI deployments using Microsoft’s existing AI tools. The project is backed by a $2.5 billion Microsoft investment and 6,000 industry and engineering experts — joining Amazon and OpenAI in launching a dedicated AI deployment organization.
Microsoft
EU AI Act Becomes Generally Applicable
On August 2, 2026 the EU AI Act became generally applicable, and the AI Office together with national authorities took up implementation, supervision, and enforcement. The same revision deferred the heaviest high-risk duties: stand-alone Annex III systems apply from December 2, 2027, and AI embedded in regulated products from August 2, 2028. Prohibited practices had already applied since February 2, 2025, and general-purpose model obligations since August 2, 2025.
OECD / G7
Agents Move Into Shipped Enterprise Products
Through the first half of 2026, agent capabilities moved out of demonstrations and into shipped enterprise products: coding, support, and office-automation agents became documented features, and tool-calling and multi-agent infrastructure matured around them. How far adoption actually reached is not settled. Published surveys report very different production rates because they do not agree on what counts as an agent, so this entry records the direction of travel rather than a level.
AI Industry
Open-Weight Models Publish at Frontier Scale
Through the first half of 2026, open-weight releases from DeepSeek, Alibaba Qwen, Moonshot, Zhipu, and Meta were published at a scale previously seen only in closed systems, with weights and model cards downloadable rather than described. Whether any of them match the closed frontier depends on which benchmark index is used and how it is weighted, so this entry records what was released and under what licence, not a ranking.
DeepSeek / Meta / Alibaba
Key Figures
Sam Altman
OpenAI CEO
Altman leads OpenAI in advancing large language model development. ChatGPT’s release made him one of the most influential figures in the AI era. He is pushing for AGI while considering AI’s impact on society.
Dario Amodei
Anthropic CEO
Amodei was formerly VP of Research (2016–2021) at OpenAI before founding Anthropic. He emphasizes AI safety and helpfulness—Claude’s design reflects this value: AI should be helpful, harmless, and honest.
Mark Zuckerberg
Meta CEO
Zuckerberg decided to open-source LLaMA, promoting the democratization of AI technology. He believes open source benefits AI safety and development, impacting the AI ecosystem profoundly.
Ilya Sutskever
Co-founder of OpenAI, founder of Safe Superintelligence Inc.
Sutskever co-founded OpenAI and was its Chief Scientist. He led the GPT-4-era model team and was a principal architect of the post-2023 scaling paradigm. In 2024 he founded Safe Superintelligence Inc., a research lab focused solely on safe superintelligent AI.
Demis Hassabis
DeepMind CEO, 2024 Nobel Laureate in Chemistry
Hassabis co-founded DeepMind and led it to breakthroughs on AlphaGo, AlphaFold, and frontier AI research. In 2024 he shared the Nobel Prize in Chemistry for AlphaFold’s contributions to protein structure prediction. He continues to argue that AI’s deepest contribution will be to scientific discovery.
Emily M. Bender
Computational Linguist
Bender is a professor of linguistics at the University of Washington and first author of On the Dangers of Stochastic Parrots, presented at FAccT in 2021. The paper argued that a language model stitches together sequences of linguistic form without reference to meaning, and it named the environmental, data, and accountability costs of building ever larger ones. It remains the most-cited critical account of the era and a standing counterweight to capability narratives.
Liang Wenfeng
Founder of DeepSeek
Liang studied at Zhejiang University and applied machine learning to quantitative finance, founding the fund High-Flyer in 2015. The computing infrastructure built there later supported DeepSeek, which he founded in 2023. Its R1 model, released in 2025 under the MIT licence, was the first open-weight system to show frontier-class reasoning, and it changed what a laboratory outside the largest Western firms was assumed to be able to publish.
Classic Quotes
We are now confident we know how to build AGI as we have traditionally understood it.
I think in five years’ time it may well be able to reason better than us.
I believe that open source is necessary for a positive AI future.
I think it could come as early as 2026, though there are also ways it could take much longer.
There is potential for serious, even catastrophic, harm, either deliberate or unintentional, stemming from the most significant capabilities of these AI models.