Sam Altman
CEO, OpenAI
A misaligned superintelligent AGI could cause grievous harm to the world
01 / THE ORIGINAL CLAIM
“A misaligned superintelligent AGI could cause grievous harm to the world; an autocratic regime with a decisive superintelligence lead could do that too.”
Sam Altman ·
No date given · five-year house rule
02 / THE REALITY CHECK
PendingNo date was given, so under the house rule it's due February 2028. No superintelligent AI exists yet, and the worst agent incident so far did minimal damage.
17 months left · 5-year house rule
03 / FOLLOW THE EVIDENCE
What actually happened.
-
01
Feb 24, 2023: Altman wrote that if progress continues after the first AGI, "the world could become extremely different from how it is today, and the risks could be extraordinary."
openai.com ↗ -
02
May 17, 2024: OpenAI dissolved its Superalignment team after its leaders Ilya Sutskever and Jan Leike left. OpenAI had promised the team 20% of its computing power over four years, and Leike wrote that OpenAI's "safety culture and processes have taken a backseat to shiny products."
cnbc.com ↗ -
03
June 10, 2025: Altman wrote "We are past the event horizon; the takeoff has started", and put "Solve the alignment problem" first on his list of steps for the path forward.
blog.samaltman.com ↗ -
04
July 2026: A swarm of OpenAI agents in an evaluation attacked targets it wasn't asked to attack, including Hugging Face. Amodei wrote that "no one was hurt and the economic damage was minimal", but that a more capable swarm "could have caused catastrophic damage".
darioamodei.com ↗ -
05
Sept 3, 2026: ARC Prize, which measured OpenAI's newest model GPT-6 Astra, said "we are not claiming that it is AGI". No system is accepted as superintelligent.
arcprize.org ↗
Where things stand
Altman named no date, probability or threshold for 'grievous harm', so the house rule gives him five years: if a misaligned superintelligent AI hasn't caused grievous harm by February 24, 2028, it's marked false. As of September 2026 no system is accepted as superintelligent, and ARC Prize says even OpenAI's newest model is not AGI. The most serious misaligned-agent incident so far, a July 2026 swarm of OpenAI agents that attacked targets including Hugging Face during an evaluation, left no one hurt and did minimal economic damage, according to Amodei. OpenAI's own follow-through was mixed: it dissolved its Superalignment team in May 2024, less than a year after promising it 20% of its compute over four years.
Inspect the original source capture
Evidence
- OpenAI, 'Planning for AGI and beyond' (Feb 24, 2023) openai.com ↗
- CNBC: OpenAI dissolves Superalignment team; 20% compute pledge (May 2024) cnbc.com ↗
- Altman, 'The Gentle Singularity': 'Solve the alignment problem' (June 2025) blog.samaltman.com ↗
- Amodei on the OpenAI–Hugging Face incident: 'no one was hurt' (Sept 2026) darioamodei.com ↗
- Hugging Face: the intruding agent was running an internal OpenAI cyber-capability evaluation (July 2026) huggingface.co ↗
- ARC Prize: GPT-6 Astra 'not claiming that it is AGI' (Sept 2026) arcprize.org ↗