· turning point
Stable Diffusion is released
A text-to-image diffusion model that runs on a gaming GPU is released with its weights under an open licence; anyone can generate anything, and the argument about that begins.
what had to happen · 40 events back to 1943
Every event this one built on, transitively, in order. Direct influences are marked.
00 · One neuron · 1
I · Foundations · 6
W1 · The first winter · 2
II · Connection · 3
W2 · The second winter · 3
III · Statistics and data · 9
IV · Deep learning · 8
V · Transformers · 8
Robin Rombach and Björn Ommer's group at LMU Munich had shown in December 2021 that diffusion could be run not on pixels but in the compressed latent space of an autoencoder, which made it roughly fifty times cheaper. Trained on LAION-5B, a public dataset of five billion image–text pairs scraped from the web, with compute paid for by Stability AI, the resulting model fitted in the memory of a consumer graphics card. On 22 August 2022 it was released with its weights, under a licence that allowed almost any use.
Within days it was running on laptops, in browsers and, crudely, on phones. Within weeks it had been fine-tuned to draw particular people, styles and products, extended with ControlNet to follow sketches and poses, and wrapped in a hundred interfaces. It was also used to make sexual images of real people, to imitate living artists by name, and to flood art sites with output. Getty Images and a group of artists sued in January 2023.
Stable Diffusion is on this timeline at the top level because it settled a question. DALL·E 2 had shown what the technology could do and had kept it behind a waiting list and a filter. Stable Diffusion showed that once a model existed it would be free, and everything since, in images and then in language, has had to assume that.
what it led to · 1 events downstream, through 2024
Built on it directly:
- 2024SoraVI