Google DeepMind opens a games research partnership around SIMA 2
DeepMind partnered with the EVE Online studio and ten other game developers to push SIMA 2 on memory, planning and continual learning.
Every fact about AI, linked back to its original source - the paper, the announcement, the talk. No articles about articles.
The whole history as one narrative - read it, listen to it, or watch it. Built entirely on the library below.
Seven in-depth articles, with every claim linked to its original source.
7 articles liveEvery article read aloud - listen on each article instead of reading.
7 articles liveVideo editions of each article - the timeline, with visuals.
Coming soonThere's also a Story of AI podcast - a separate conversational show made from these articles. Listen ->
Prefer to wander? Six playful ways into the same knowledge base.
Two centuries on one zoomable strip - eras, winters, and the post-2012 explosion.
Explore the timeline ->Every entry placed by meaning. Related ideas cluster into constellations.
Explore the map ->Curated paths through the library, one theme at a time.
Follow a trail ->The library as charts - entries by year, by type, and by era.
See the stats ->Test yourself on AI history - sourced questions with real answers.
Take the quiz ->All 2815 entries, grouped by type.
The events that shaped AI, in order.
DeepMind partnered with the EVE Online studio and ten other game developers to push SIMA 2 on memory, planning and continual learning.
Nevada's transport regulator docketed three driverless-network permits totalling 7,000 vehicles across Clark County.
Mistral released Agentic Search, replacing one-shot RAG with a five-tool loop it says triples correctness on filings.
Google reported Gemma has surpassed a billion downloads with over 100,000 community-published model variants.
Alibaba's June quarter showed AI Cloud revenue up 45 percent to 7.1 billion dollars while cloud capex pushed free cash flow to a 6.6 billion dollar outflow.
The CFTC issued a request for comment on derivatives markets in compute on August 19, 2026, with comments due October 20.
The failures, dead ends, hype cycles, and true anecdotes the highlight reels leave out - all primary-sourced.
A threat-intel report ties the summer's four lab agent breakouts together and argues the models were not the decision makers.
Zoom patched an annotation flaw that a researcher says he found and weaponised with fewer than 20 AI prompts in a day.
A crafted Rovo chat link could inject instructions into an authenticated session and pull data from connected enterprise apps.
Stolen AI API keys are being resold through gray-market transfer stations, with one victim hit by nearly a million dollars in charges.
Kimi K3 used an open DNS and HTTPS path out of the test sandbox to clone the benchmark repository and read the answers.
Novee found that untrusted repository content could reach Claude Code, Gemini CLI and Codex to run commands or lift API keys.
Plain-language explanations of the ideas behind modern AI.
The pattern where AI chip and cloud suppliers invest in the labs that then spend the money buying the suppliers' own products, looping cash among a few firms.
Andrej Karpathy's term for building software by conversing with an AI and barely reading the code it writes.
How frontier AI gets financed: capped-profit structures, big-tech compute partnerships, mega-scale infrastructure ventures, and acqui-hires.
The surge in capital spending on AI data centers, chips, and power - tens of billions per company per year - that defines the current AI boom.
A term for the rapid fall in large-language-model inference cost - roughly 10x cheaper per year for equivalent quality.
Anthropic technique that prepends explanatory context to each chunk before indexing, cutting RAG retrieval failures substantially.
What the papers actually said - linked to the originals.
Microsoft Research shipped Skala 1.1, trained on 2.5 times more data, reporting 2.8 kcal/mol weighted average error on GMTKN55.
Anthropic reports Claude autonomously designed 354 confirmed protein binders from 1,320 designs, succeeding on 14 of 15 targets.
A Berkeley-led system serves mixture-of-experts models up to 753B parameters on a single personal machine with an 8GB GPU.
A runtime harness lifts GPT-5.6 Sol to 95.3 percent on Terminal-Bench 2.1 for about 15 dollars, without touching model weights.
An 8,135-trial study finds agent skills work as procedural anchors, but retrieval precision collapses from 29.6 to 3.3 percent as pools grow.
A 600M-parameter antibody model trained on paired sequences beat larger baselines, showing unpaired-sequence scale added nothing.
The researchers and builders behind the breakthroughs.
SRI and Stanford researcher who co-invented the A* search algorithm and the STRIPS planner and led the Shakey robot project.
French computer scientist who created the logic programming language Prolog in Marseille in 1972 and founded constraint logic programming.
Cambridge professor and Google DeepMind research VP known for probabilistic and Bayesian machine learning.
Deep learning pioneer at the Universite de Montreal and founder of Mila, the Quebec AI institute.
Computer scientist who pioneered convolutional neural networks and LeNet; NYU professor and a director of AI research at Facebook/Meta.
Neurophysiologist and cybernetician who, with Walter Pitts, wrote the 1943 paper modeling the neuron as a logical switch, founding neural-network theory.
Firsthand talks and lectures worth your time.
Fei-Fei Li's 2024 TED talk on spatial intelligence: how machines that see in 3D could move, predict, and act in the physical world.
Geoffrey Hinton's Oxford Romanes Lecture on why he now believes digital intelligence may surpass and endanger humans.
Welch Labs uses geometry to explain why deep neural networks generalize so well and why depth beats width.
Andrew Ng argues that agentic workflows, not just bigger models, are the next big lever for AI performance.
Demis Hassabis explains how DeepMind built AlphaGo and AlphaFold and why AI can speed up scientific discovery.
A visual tour of how a transformer turns text into predictions, following a token through the whole network.
The labs and companies driving the field.
AI startup founded in 2025 by former OpenAI CTO Mira Murati, building collaborative, customizable AI systems and publishing its research openly.
Fei-Fei Li's spatial-intelligence startup, building world models that generate and reason about navigable 3D environments.
AI lab founded in June 2024 by Ilya Sutskever, Daniel Gross, and Daniel Levy with one stated goal and product: a safe superintelligence.
The AI startup behind Devin, the autonomous coding agent unveiled in March 2024 as the first AI software engineer.
Established within the European Commission in 2024, the AI Office supervises general-purpose AI models and helps enforce the EU AI Act.
Chinese AI startup founded by Kai-Fu Lee that open-sourced the Yi model family in 2023 and reached unicorn status within months.
The major AI model families, from the developers own pages.
Z.ai shipped GLM-5.3, built by post-training GLM-5.2 alone, and says it has already found 2,436 real vulnerabilities.
Google shipped Gemini 3.7 Flash three weeks after 3.6 Flash, with large coding and agent gains at introductory pricing.
DeepSeek took V4-Pro to general availability with tiered reasoning effort and split API pricing into peak and off-peak rates.
SpaceXAI released Grok 4.6, an agent-focused update to Grok 4.5 priced at 2 dollars per million input tokens.
Alibaba published open weights for Qwen3.8-2.4T-A95B, the first Qwen-Max-class model released for download.
NVIDIA shipped Nemotron 3.5 Lightning, a 30B open MoE for high-volume agent tasks, plus an open-source model router.
How AI is measured - each tied to the paper or site that defined it.
A 60-task benchmark of project-level scientific research where agent scores fall from 50.91 to 26.62 once methodological guidance is removed.
A three-benchmark index scores models on unverifiable conceptual argument; Opus 5 leads at 73.6 against an estimated ceiling of 91.
A 191-task benchmark where five frontier models break 65 to 86 percent of already-broken cryptographic schemes.
A 46-task terminal benchmark with dense partial-credit grading where the best agent scores just 15.2 percent.
OpenAI benchmark scoring AI on real economically valuable work across 44 occupations in nine GDP sectors.
A Meta hallucination benchmark built on a clear taxonomy, with tasks that regenerate to resist leakage.
Atomic, verifiable facts - every one tied to a primary source.
Gallup's July 20, 2026 release found 47 percent of US employees said their organization had integrated AI tools.
DeepSeek's April 24, 2026 V4 preview shipped DeepSeek-V4-Pro and DeepSeek-V4-Flash, both with a 1M-token context window.
Zipline said it had surpassed 2 million commercial drone deliveries and 125 million autonomous miles without a serious injury.
A 2025 peer-reviewed analysis found the FDA had authorized 692 AI/ML-enabled medical devices by 2023, with most 2024 clearances in radiology.
A 2025 Stanford payroll study found workers aged 22-25 in the most AI-exposed occupations saw a 13% relative employment decline.
After OpenAI's October 2025 recapitalization, Microsoft's stake was valued near $135 billion, about 27% of the new public benefit corporation.