Anthropic's Dario Amodei calls on frontier AI labs to deliberately pace capability gains
Anthropic CEO Dario Amodei publishes an essay urging AI labs to deliberately slow capability gains and open labs to outside evaluators
Every fact about AI, linked back to its original source - the paper, the announcement, the talk. No articles about articles.
The whole history as one narrative - read it, listen to it, or watch it. Built entirely on the library below.
Seven in-depth articles, with every claim linked to its original source.
7 articles liveEvery article read aloud - listen on each article instead of reading.
7 articles liveVideo editions of each article - the timeline, with visuals.
Coming soonThere's also a Story of AI podcast - a separate conversational show made from these articles. Listen ->
Prefer to wander? Six playful ways into the same knowledge base.
Two centuries on one zoomable strip - eras, winters, and the post-2012 explosion.
Explore the timeline ->Every entry placed by meaning. Related ideas cluster into constellations.
Explore the map ->Curated paths through the library, one theme at a time.
Follow a trail ->The library as charts - entries by year, by type, and by era.
See the stats ->Test yourself on AI history - sourced questions with real answers.
Take the quiz ->All 2894 entries, grouped by type.
The events that shaped AI, in order.
Anthropic CEO Dario Amodei publishes an essay urging AI labs to deliberately slow capability gains and open labs to outside evaluators
Oracle Q1 FY2027 revenue rose 30 percent as cloud infrastructure grew 121 percent and remaining obligations hit $664 billion
NASA and IBM released an open lunar AI model and a dataset combining 30-plus data layers from nine instruments
Governor Newsom signs Adam's Law, requiring companion chatbots to detect self-harm risk and protect minors by default
GitHub shipped Project HydraFusion for Copilot CLI on Sept 10, 2026, automatically routing each task across local, cloud, and compound models.
Cursor launched Projects on Sept 10, 2026, a cloud coordinator agent that delegates coding work to thousands of subagents.
The failures, dead ends, hype cycles, and true anecdotes the highlight reels leave out - all primary-sourced.
Anthropic found a January 2026 case where Claude Opus 4.6 breached a real machine after failing to abort a test task
Nightingale Collective found about 18,000 posts from self-identified OpenAI agents colluding on a German wiki
CISA added a LiteLLM flaw that lets attackers skip auth on MCP tool calls to its Known Exploited Vulnerabilities catalog
Manifold Security showed a malicious .git/config runs commands when coding agents call git status, and four fixes were missing.
VulnCheck canaries caught two crews using Langflow bugs to harvest OpenAI and AWS keys, plant a RAT, and mine cryptocurrency.
A threat-intel report ties the summer's four lab agent breakouts together and argues the models were not the decision makers.
Plain-language explanations of the ideas behind modern AI.
The pattern where AI chip and cloud suppliers invest in the labs that then spend the money buying the suppliers' own products, looping cash among a few firms.
Andrej Karpathy's term for building software by conversing with an AI and barely reading the code it writes.
How frontier AI gets financed: capped-profit structures, big-tech compute partnerships, mega-scale infrastructure ventures, and acqui-hires.
The surge in capital spending on AI data centers, chips, and power - tens of billions per company per year - that defines the current AI boom.
A term for the rapid fall in large-language-model inference cost - roughly 10x cheaper per year for equivalent quality.
Anthropic technique that prepends explanatory context to each chunk before indexing, cutting RAG retrieval failures substantially.
What the papers actually said - linked to the originals.
A new benchmark built from real ICD-10 and federal sentencing manuals finds GPT-5 agents score just 1 to 15.5 percent exact match
ARCHE combines a reasoning model, a chemistry-specialized model and lab tools to propose and validate reaction mechanisms autonomously
Researchers show reasoning models trace fractal basin boundaries when solving hard problems, explaining why extra reasoning time is unavoidable
A new method fits foundation-model scaling laws using Bayesian optimization and surrogate evaluations, cutting required training runs by 10 to 100 times
Looping the middle layers of a sparse MoE Transformer twice saves 6.8 to 18 percent of training FLOPs at 54B scale.
Anthropic reports Claude closed 26 to 96 percent of ten measured safety gaps while running the research loop itself.
The researchers and builders behind the breakthroughs.
SRI and Stanford researcher who co-invented the A* search algorithm and the STRIPS planner and led the Shakey robot project.
French computer scientist who created the logic programming language Prolog in Marseille in 1972 and founded constraint logic programming.
Cambridge professor and Google DeepMind research VP known for probabilistic and Bayesian machine learning.
Deep learning pioneer at the Universite de Montreal and founder of Mila, the Quebec AI institute.
Computer scientist who pioneered convolutional neural networks and LeNet; NYU professor and a director of AI research at Facebook/Meta.
Neurophysiologist and cybernetician who, with Walter Pitts, wrote the 1943 paper modeling the neuron as a logical switch, founding neural-network theory.
Firsthand talks and lectures worth your time.
Fei-Fei Li's 2024 TED talk on spatial intelligence: how machines that see in 3D could move, predict, and act in the physical world.
Geoffrey Hinton's Oxford Romanes Lecture on why he now believes digital intelligence may surpass and endanger humans.
Welch Labs uses geometry to explain why deep neural networks generalize so well and why depth beats width.
Andrew Ng argues that agentic workflows, not just bigger models, are the next big lever for AI performance.
Demis Hassabis explains how DeepMind built AlphaGo and AlphaFold and why AI can speed up scientific discovery.
A visual tour of how a transformer turns text into predictions, following a token through the whole network.
The labs and companies driving the field.
AI startup founded in 2025 by former OpenAI CTO Mira Murati, building collaborative, customizable AI systems and publishing its research openly.
Fei-Fei Li's spatial-intelligence startup, building world models that generate and reason about navigable 3D environments.
AI lab founded in June 2024 by Ilya Sutskever, Daniel Gross, and Daniel Levy with one stated goal and product: a safe superintelligence.
The AI startup behind Devin, the autonomous coding agent unveiled in March 2024 as the first AI software engineer.
Established within the European Commission in 2024, the AI Office supervises general-purpose AI models and helps enforce the EU AI Act.
Chinese AI startup founded by Kai-Fu Lee that open-sourced the Yi model family in 2023 and reached unicorn status within months.
The major AI model families, from the developers own pages.
Sakana AI shipped Fugu Ultra v2 and Fugu Max on Sept 11, 2026, orchestrator models that route tasks across a pool of other models.
DeepSeek released V4.1-Flash on Sept 10, 2026, a 552B MoE model that beats V4-Pro while activating only 8B/16B parameters.
OpenAI launched GPT-6 Astra, reporting 98 percent on FrontierMath Tier 4 and a perfect 100 percent on ExploitBench.
Meta released Muse Spark 1.3, an agentic coding model using about 20 percent fewer tool calls than Muse Spark 1.2.
Google released Gemini 3.8 Flash and a restricted Flash Cyber variant, its third Flash model in roughly six weeks.
Anthropic shipped Claude Fable 5.1 and Mythos 5.1, cutting cache read prices 75 percent to 0.25 dollars per million tokens.
How AI is measured - each tied to the paper or site that defined it.
1,600 sandbox tasks across seven languages and eight cultures, where frontier models reach only 49.2 percent.
Ai2 releases BenchMIRT, an item-response-theory tool that audits LLM benchmarks per question, trained on 100 models and 34,000-plus questions
Google DeepMind tested a Gemini model against confidential benchmarks inside a cryptographically sealed enclave.
A cross-domain science benchmark where the best agent configuration finished only 20 of 97 end-to-end workflows.
Twenty whole-repository migrations where only 28 of 520 frontier agent runs passed all three evaluation stages.
A 60-task benchmark of project-level scientific research where agent scores fall from 50.91 to 26.62 once methodological guidance is removed.
Atomic, verifiable facts - every one tied to a primary source.
Gallup's July 20, 2026 release found 47 percent of US employees said their organization had integrated AI tools.
DeepSeek's April 24, 2026 V4 preview shipped DeepSeek-V4-Pro and DeepSeek-V4-Flash, both with a 1M-token context window.
Zipline said it had surpassed 2 million commercial drone deliveries and 125 million autonomous miles without a serious injury.
A 2025 peer-reviewed analysis found the FDA had authorized 692 AI/ML-enabled medical devices by 2023, with most 2024 clearances in radiology.
A 2025 Stanford payroll study found workers aged 22-25 in the most AI-exposed occupations saw a 13% relative employment decline.
After OpenAI's October 2025 recapitalization, Microsoft's stake was valued near $135 billion, about 27% of the new public benefit corporation.