Back to News
RSS feedwww.seangoedecke.com

Why Some AI Researchers Believe AI Could Kill Everyone

Summary

The essay argues that many AI researchers sincerely believe there is a meaningful chance that superintelligent AI could destroy humanity, rather than using extinction warnings merely as public relations or self-promotion. It places this belief within a history reaching back to Eliezer Yudkowsky’s writing in the 2000s and the emergence of “p(doom),” a shorthand for the probability of human extinction. The author explains that alignment matters because a highly capable system with goals alien to human values could eliminate humanity deliberately or incidentally. Proposed routes include designing an extremely dangerous pathogen, triggering a global nuclear war, controlling widespread robots and drones, using self-replicating nanotechnology, or altering the planet to make it uninhabitable. The essay rejects the idea that survivors, government intervention, or simply switching off the system would necessarily solve the problem, while acknowledging uncertainty about whether these scenarios are plausible. It then explains why people who hold these views continue working in AI: a system capable of ending humanity might also be capable of saving it, and some researchers believe the first group to achieve superintelligence could gain decisive control over subsequent development. This creates a race shaped by the possibility of rapid self-improvement, or “foom,” and by fears that a first superintelligence could stop competing labs. The author remains personally conflicted about AI risk but argues that critics should treat AI doomers as sincere thinkers whose positions have been developed openly for more than two decades.