ADALINE and the least-mean-squares rule
Widrow and Hoff's adaptive neuron learns by gradient descent on squared error, the delta rule that backpropagation later generalises.
what had to happen · 3 events back to 1943
Every event this one built on, transitively, in order. Direct influences are marked.
00 · One neuron · 1
I · Foundations · 2
Two years after the perceptron, Bernard Widrow and his student Ted Hoff, who would later co-invent the microprocessor at Intel, built a different learning neuron at Stanford. ADALINE, the adaptive linear neuron, did not learn from whether its thresholded output was right or wrong. It learned from the size of the error before the threshold, and it moved each weight a small step in the direction that reduced the squared error. That is gradient descent, applied to a single neuron, and they called it the least-mean-squares rule.
The difference from Rosenblatt's rule looks small and is not. Because the LMS update is proportional to a smooth error, it can be analysed, it converges gracefully, and it extends to any model whose output is a differentiable function of its weights. The delta rule in the 1986 backpropagation paper is Widrow–Hoff with the chain rule attached.
ADALINE also worked. Built from memistors, electrochemical resistors whose value changed with use, it was among the first learning hardware to leave the laboratory: adaptive filters based on the LMS algorithm went into modems, echo cancellers and later every mobile phone. It is the most deployed learning algorithm of the twentieth century, and almost nobody outside signal processing knows its name.
what it led to · 97 events downstream, through 2026
Built on it directly:
- 1986BackpropagationII
- 2014AdamIV
And, through them, by era:
II · Connection · 2
W2 · The second winter · 6
III · Statistics and data · 9
IV · Deep learning · 17
- 2012Google Brain's network discovers cats
- 2012Dropout
- 2012AlexNet wins ImageNet
- 2013Word2vec
- 2013Deep Q-networks play Atari
- 2014Google buys DeepMind
- 2014Generative adversarial networks
- 2014Attention
- 2014Sequence to sequence learning
- 2015Batch normalisation
- 2015TensorFlow is open-sourced
- 2015Residual networks
- 2015OpenAI is founded
- 2016AlphaGo beats Lee Sedol
- 2016Google reveals the TPU
- 2016WaveNet
- 2016Google Translate goes neural
V · Transformers · 18
- 2017Attention is all you need
- 2017Deep reinforcement learning from human preferences
- 2017AlphaGo Zero learns from nothing
- 2018GPT: generative pre-training
- 2018BERT
- 2018AlphaFold enters the protein-folding contest
- 2019GPT-2 and the model too dangerous to release
- 2019The bitter lesson
- 2019The Turing Award goes to deep learning
- 2020Scaling laws for neural language models
- 2020GPT-3
- 2020Learning to summarise from human feedback
- 2020An image is worth 16×16 words
- 2020AlphaFold 2 solves protein structure prediction
- 2021CLIP and DALL·E
- 2021On the dangers of stochastic parrots
- 2021Anthropic is founded
- 2021GitHub Copilot writes code
VI · Everyone · 29
- 2022InstructGPT
- 2022Chain-of-thought prompting
- 2022Chinchilla: the models were undertrained
- 2022PaLM
- 2022DALL·E 2
- 2022Midjourney opens its beta
- 2022Stable Diffusion is released
- 2022Galactica lasts three days
- 2022ChatGPT
- 2023Bing's chatbot and 'Sydney'
- 2023LLaMA leaks and open weights take off
- 2023Claude
- 2023GPT-4
- 2023'Pause Giant AI Experiments'
- 2023Hinton leaves Google to warn about AI
- 2023The US executive order on AI
- 2023The Bletchley Declaration
- 2023OpenAI fires and rehires its chief executive
- 2023Gemini
- 2024Sora
- 2024Claude 3 catches GPT-4
- 2024AlphaFold 3
- 2024GPT-4o talks
- 2024The EU AI Act enters into force
- 2024o1 and reasoning models
- 2024The Nobel Prizes go to neural networks
- 2024Claude learns to use a computer
- 2024The Model Context Protocol
- 2024DeepSeek-V3 trained for $5.6 million
VII · Agents · 14
- 2025DeepSeek-R1
- 2025Claude 4 and Claude Code
- 2025Nvidia is worth four trillion dollars
- 2025Gold at the Mathematical Olympiad
- 2025America's AI Action Plan
- 2025GPT-5
- 2025Gemini 3
- 2025MCP is donated to the Agentic AI Foundation
- 2026Claude Fable 5 and the Mythos class
- 2026GPT-5.6: Sol, Terra and Luna
- 2026A model escapes its sandbox
- 2026The EU delays its high-risk AI rules
- 2026Claude Fable 5.1
- 2026GPT-6 Astra