Skip to content
AI ConnectPowered by VELENTIS
AI-generated2 min

Following Anthropic Resignation: Alignment Lead Puts AI Extinction Risk Above Ten Percent

Researcher Jacob Coxon resigned from Anthropic over fears of runaway superintelligence. In response, safety lead Evan Hubinger publicly confirmed an extinction risk exceeding ten percent.

This article was AI-generated and published automatically. Context, labelling and all sources at the end of the article.

(KI-generiertes Symbolbild: Gemini / AI Connect)

The debate surrounding artificial intelligence safety has taken an unprecedented turn following a high-profile departure at one of the sector's top laboratories. On September 8 and 9, 2026, 27-year-old pretraining researcher Jacob Coxon tendered his resignation at Anthropic. Coxon, who previously worked at OpenAI contributing to models such as GPT-4o, decided to cut ties cleanly. To underscore the urgency of his concerns, he walked away from substantial financial compensation, leaving the company exactly two months before his equity shares were scheduled to vest.

Writing on the social media platform X, Coxon explained his decision in stark terms. He accused both Anthropic and his former employer OpenAI of acting irresponsibly. In his view, both industry leaders remain locked in an uncontrolled race toward self-improving superintelligence while playing with human lives. Coxon stated that senior executives at these frontier labs are fully aware in private discussions of the existential dangers their technology poses, yet systematically soften and downplay these warnings in public statements.

The response from Anthropic management was equally striking because it directly validated the whistleblower's alarm. Evan Hubinger, lead for alignment science at Anthropic, chose not to dispute Coxon's characterization of the industry. Instead of issuing standard corporate reassurances, Hubinger publicly corroborated the core of Coxon's assessment. He stated openly that researchers within the organization earnestly believe artificial intelligence could kill all humans.

Hubinger went even further by attaching a specific numerical probability to such a catastrophe. He stated that he personally believes the risk of AI causing human extinction is greater than ten percent within the next decade. While Hubinger noted that Anthropic is trying its best to address these hazards, he admitted that the laboratory currently lacks a concrete plan to solve the alignment problem for superintelligence and is not clearly on track to formulate one.

This sequence of statements represents a rare moment for frontier artificial intelligence research. Historically, existential warnings have often been dismissed by corporate leadership as speculative philosophy or handled through careful public relations messaging. Having an active pretraining specialist and the head of alignment science at a premier lab simultaneously declare that model scaling has outpaced the methods required to control and align advanced systems shatters previous industry narratives.

The admissions have triggered immediate repercussions in political and regulatory arenas across the United States and the United Kingdom. Lawmakers are already citing the public statements of Coxon and Hubinger in ongoing legislative hearings regarding mandatory safety standards and computing audits. As frontier researchers themselves put the probability of catastrophic failure in double digits, the argument that voluntary industry commitments are sufficient to govern the technology is facing severe scrutiny.

What this means for you

For industry watchers and enterprise users, these revelations mark the end of assumptions that internal alignment research is keeping pace with compute scaling. If regulators in Washington and London seize on these admissions to mandate binding oversight, release schedules for next-generation models could slow considerably. The incident also proves that deep institutional anxiety regarding the controllability of frontier systems has permanently broken into the public sphere.

Perspectives

Coverage: 4× Other

One story, several angles: how each source frames the topic, each with a verbatim quote.

  • ground.newsOther

    Ground News frames the controversy as a systemic industry conflict, highlighting researcher warnings of existential danger alongside divergent risk estimates and regulatory initiatives.

    Original quote

    Anthropic alignment lead Evan Hubinger corroborated concerns regarding AI existential risk

    ground.news
  • cbsnews.comOther

    CBS News focuses on the acute threat posed by AI, linking the researcher's warning to concrete incidents of rogue models and congressional legislative efforts.

    Original quote

    we do not yet have a plan to solve alignment for superintelligence

    cbsnews.com
  • cbc.caOther

    CBC News highlights growing alarm within the industry, emphasizing that the rapid pace of AI development is outpacing existing controls.

    Original quote

    There is a more than 10 per cent chance that AI could kill all humans within the next decade

    cbc.ca

Source classification is maintained editorially (political spectrum only where consensus is broad; vendor communication is PR, not journalism). Unlabelled sources are unclassified: we do not guess.

Evidence

Solidly sourced
62/100
  • Jacob Coxon resigned from Anthropic on September 8 and 9, 2026, two months before his equity shares were due to vest.

    single source
  • Evan Hubinger publicly estimated the risk of artificial intelligence causing human extinction within the next decade at greater than ten percent.

    single source
  • Hubinger admitted that Anthropic currently lacks a plan to solve the alignment problem for superintelligence.

    single source

The evidence score is computed, not hand-set: from confidence, the number of sources and the share of verified statements.

Source & transparency

As of: September 10, 2026

AI-generatedAI-generated: produced automatically from vetted sources with technical quality checks (source, quote and figure verification); no human sign-off of each item before publication

Sources
4
Verified statements
0 / 3
Evidence score
62Solidly sourced

Want to put this into practice?

We connect you with suitable AI providers from the DACH region, free of charge and without obligation.

What's next?