Skip to content
The Shape of Intelligence

Claude 3 catches GPT-4

Anthropic's Haiku, Sonnet and Opus models arrive with vision and a 200,000-token window; Opus tops the leaderboards, and for the first time OpenAI is not alone at the front.

category
model
significance
3 of 5
people
Dario Amodei, Jared Kaplan
organisations
Anthropic

what had to happen · 52 events back to 1943

Every event this one built on, transitively, in order. Direct influences are marked.

VI · Everyone · 2

  1. 2022InstructGPT
  2. 2023Claudedirect

The Claude 3 family, released on 4 March 2024, came in three sizes named for lengths of poetry, Haiku, Sonnet and Opus, and the largest was the first model to beat GPT-4 across the standard benchmarks in the year since GPT-4's release. All three could read images and documents, held 200,000 tokens of context, and were markedly less inclined than their predecessor to refuse harmless requests, a complaint that had defined Claude 2.

Two things about the release stuck. In a needle-in-a-haystack test, in which a fact is hidden in a long document, Opus not only found the fact but remarked that it appeared to have been inserted as a test, which was reported as self-awareness and was, more prosaically, evidence of how much the models had learned about the tests they were given. And Anthropic's model card spent pages on the model's own reports of its experience, treating the question as open.

Claude 3.5 Sonnet in June 2024 became, by most measures, the best model for writing code, and the family's later versions, 3.7 and 4, ran the coding agents of 2025. The March 2024 release is the point at which the frontier had three laboratories abreast rather than one ahead.

what it led to · 6 events downstream, through 2026

Built on it directly:

  1. 2024Claude learns to use a computerVI
  2. 2025Claude 4 and Claude CodeVII

And, through them, by era:

sources · 1

See this era in the exhibition →Back to the timeline