Dario Amodei
CEO, Anthropic
AI better than almost all humans at almost all tasks within 2–3 years
01 / THE ORIGINAL CLAIM
“At some point we're going to get to AI systems that are better than almost all humans at almost all tasks. The term I've used for it in an essay I recently wrote is, a country of geniuses in a data center. … That's the thing I think we are quite likely to get in the next two or three years”
Dario Amodei ·
Deadline given: in the next two or three years
02 / THE REALITY CHECK
PendingDeadline is January 2028. Progress has been fast and is partly on track, but no model is yet better than almost all humans at almost all tasks.
16 months left
03 / FOLLOW THE EVIDENCE
What actually happened.
-
01
At Davos on 21 January 2025, Amodei told CNBC that AI 'better than almost all humans at almost all tasks' was 'quite likely' within two or three years. At a WSJ event the same week he hedged: 'I don't know if it'll be 2027.'
cnbc.com ↗ -
02
In January 2026 he repeated the forecast ('we may have AI that is more capable than everyone in only 1–2 years'). He cited METR's finding that Claude Opus 4.5 could do about four hours of human work at 50% reliability.
darioamodei.com ↗ -
03
METR's latest measurement (May 2026) puts Claude Mythos Preview at a 50% time horizon of roughly 17 hours (CI about 8.5–55 h) and an 80% horizon of about 3 hours. METR warns that measurements above 16 hours are unreliable.
metr.org ↗ -
04
ARC-AGI-3 launched in March 2026 with humans at 100% and frontier AI at 0.51%. Claude Opus 5 set a record of 30.2% in July 2026, and in September GPT-6 Astra scored 62.7% with the standard harness. ARC Prize said it is 'not claiming that it is AGI'.
arcprize.org ↗ -
05
On OpenAI's GDPval, GPT-5.4 (March 2026) matched or beat industry professionals in 83% of comparisons across 44 occupations. That is strong, but it is a benchmark of well-specified tasks, not 'almost all tasks'.
the-decoder.com ↗ -
06
In September 2026 Amodei wrote that AI has been 'advancing drastically faster' since the summer because of early recursive self-improvement, and he called on the industry to slow down.
darioamodei.com ↗
Where things stand
The deadline hasn't passed, and parts of the trend are on track. METR measured Claude Mythos Preview at a 50% time horizon of about 17 hours, and GPT-6 Astra scores 62.7% on ARC-AGI-3 with the standard harness (99.9% with a provider-specific adapter harness). Still, METR's 80% horizon is only about 3 hours, and METR says its measurements above 16 hours are unreliable. ARC Prize says Astra is not AGI. Amodei's January 2026 essay repeated the forecast: AI 'more capable than everyone in only 1–2 years'.
Inspect the original source capture
Evidence
- CNBC Davos quote, 21 Jan 2025 cnbc.com ↗
- METR: Claude Mythos Preview 50% horizon ~17 h; >16 h unreliable metr.org ↗
- ARC-AGI-3 launch: humans 100%, frontier AI 0.51% (Mar 2026) arcprize.org ↗
- ARC Prize: Claude Opus 5 30.2% on ARC-AGI-3 (July 24, 2026) arcprize.org ↗
- ARC Prize: GPT-6 Astra 62.7% standard / 99.9% adapter; 'not claiming that it is AGI' arcprize.org ↗
- GPT-5.4 GDPval 83% wins or ties vs professionals the-decoder.com ↗