Anthropic researchers warn AI may trigger human extinction by 2030
anthropic reasoning
| Source: Mastodon | Original article
Anthropic researchers warn AI could cause human extinction by 2030, as executives resign and frontier models claim human‑level reasoning.
Three researchers with ties to Anthropic have warned that artificial intelligence could wipe out humanity by 2030. The most prominent voice, alignment lead Evan Hubinger, told Inside AI that he personally believes there is “a greater than 10 % chance” that sufficiently advanced AI will cause human extinction within the next decade. A second researcher, who has since left Anthropic, echoed the alarm, while a third added similar concerns in a written response to a request for comment.
The statements arrive as a wave of unease sweeps the AI sector. Executives at several frontier‑AI firms have recently resigned, citing ethical doubts, and a growing number of models are being billed as having reached human‑level reasoning. Anthropic’s own internal debates about safety have been documented in our coverage of its distillation efforts and predictive surveillance projects earlier this month.
Why the warning matters is twofold. First, it comes from a company that positions itself as a leader in AI alignment, lending credibility to the risk assessment. Second, a quantified “greater than 10 %” chance of extinction is unusually explicit for a corporate researcher, potentially sharpening the focus of regulators, investors and the broader public on the need for robust safeguards.
What to watch next includes any formal risk assessments or policy proposals from Anthropic, further resignations or internal reviews, and reactions from governments and industry bodies that have been drafting AI safety frameworks. The conversation may also prompt more researchers to voice quantitative risk estimates, shaping the next round of regulation and funding for alignment work.
Sources
Back to AIPULSEN