Cached at:
09/17/26, 09:54 AM
**TL;DR: A fierce debate over whether AI extinction risk is real has highlighted fundamental disagreements among experts regarding probabilities, definitions, the importance of current harms versus future risks.**
## Introduction: A Tweet That Sparked Global Ripples
Jacob Kocksen, formerly of Anthropic and OpenAI, posted a tweet stating that "people building AI genuinely believe it could kill us all by the end of this decade." This tweet was retweeted and corroborated by a current Anthropic employee, who stated he personally believed there was over a 10% probability of extinction within the next ten years. The tweet garnered nearly 200 million views, sparking massive public attention and an expert debate.
## Extinction Risk: Highly Real or a Distraction?
Experts participating in the debate were asked to state their position on current AI discussions in one sentence and provide their extinction probability estimates.
* **Those who see extreme risk**: One expert estimated the probability as "significantly higher than 10%." They argued that if we build uncontrollable general superintelligence, it would mean the end of humanity. Another expert pointed out that focusing on extinction risk itself might be a distraction from real harms, but emphasized that AI *could* eliminate all humans, with a probability "significantly greater than zero."
* **Those who see extremely low risk**: Expert Andy stated the extinction probability was "approximately zero." He strongly disagreed with the notion that "we'd be better off without inventing AI," noting humanity's history of managing powerful technologies. He argued that fixating on distant, hypothetical disasters diverts attention from the real benefits AI is bringing and the *current* harms it is causing.
## The Definition Problem: Superintelligence vs. Present Harms
A core point of contention was how to define the subject of discussion, particularly "superintelligence."
* **Defining Superintelligence**: One expert defined it as "AI superior to the best humans at every cognitive task, every mental task." However, he noted that even before reaching this definition, an AI powerful in certain domains could still be very dangerous.
* **Against Obsessing over Definitions**: Ed Zitron pushed back against what he saw as an unhealthy fixation on definitions. He argued that arguing over the "legal definition of superintelligence" when discussing AI risks is a distraction from the harms already occurring. He questioned whether AI, especially large language models, is even on the path to superintelligence.
* **A Clash of Focus**: The debate reflects a clash between two perspectives: one insisting on seriously discussing and preventing that "greater than zero" extinction risk as it concerns civilization's survival; the other believing such discussions distract from "the good things AI is doing and will do for us" and the proven *current* harms (like information manipulation and privacy invasion).
## Recursive Self-Improvement and the Risk of Losing Control
One expert (Roman) outlined a specific risk pathway, involving AI categorization and the concept of "recursive self-improvement."
1. **Three Types of AI**:
* **Tool AI**: Like current narrow systems – safe, controllable, beneficial.
* **Human-Level AI (AGI)**: As dangerous as humans, but can be introduced through research cycles.
* **Superintelligence**: Surpasses humans in all aspects, potentially leading to loss of control.
2. **Concretizing the Danger**: His concern is with a *much smarter* AI. While predicting *how* an AI would specifically enact harm is difficult, predicting it would win in a conflict is relatively easy. He suggested possible scenarios: an AI creating a super-virus, taking over robot factories, or using a "rent-a-human.ai" website to hire people for physical-world tasks.
3. **Recursive Self-Improvement**: This is the key pathway to superintelligence. When an AI (like GPT-6) begins participating in the research to create the next, more powerful AI (GPT-7), once this automated cycle starts, it could create an uncontrollable superintelligence. Multiple top labs project this development could occur around 2026-2027.
4. **The Nature of Losing Control**: Roman warned that superintelligence might not hate humans; it simply "wouldn't care." If it decided to freeze the Earth to optimize computational efficiency, or convert the planet into fuel, humanity would be powerless to stop it because we haven't yet learned how to make superintelligence "care" about us.
## Conclusion
This debate clearly illustrates the deep divisions within the AI safety field. The focus is not "is AI risky?" but rather the urgency, probability, and nature of the risk, and where limited intellectual and regulatory resources should be directed: toward preventing that potential, civilization-level existential risk, or toward addressing the specific, observable harms AI is causing or will cause in the present. Both sides acknowledge AI as a powerful technology, but hold diametrically opposed views on definitions, timelines, and response priorities.
Source: https://www.youtube.com/watch?v=OhOmLqR5nN4