Anthropic's >10% AI extinction odds trigger UK Cabinet shakeup
Anthropic's own safety researcher says there is a greater than 10% chance AI kills all humans within a decade, and the UK government is responding by elevating its AI safety minister to Cabinet. The warning comes amid reports that Anthropic withheld its latest model from the UK's AI Safety Institute.
Beat this week
Last 7 days · Policy & Regulation
Impact 6.4/10 (-0.6 vs prior). Counts are stories in our record, not a market forecast.
Open the change reportCoverage balance Negative coverage leads. Negative coverage exceeds positive coverage by 56 percentage points.
This story sits in Policy & Regulation — the counts compare this beat's last 7 days with the previous 7 in our verified record, not a market forecast.
Figures are computed live from our source-verified story record (as of ) The volume change compares this window with the prior 7 days in the same record. — see our methodology for how impact and sentiment are derived.
AI briefing
Key takeaways
- Anthropic's own safety researcher says there is a greater than 10% chance AI kills all humans within a decade, and the UK government is responding by elevating its AI safety minister to Cabinet.
- The warning comes amid reports that Anthropic withheld its latest model from the UK's AI Safety Institute.
- worcesternews.co.uk
- eadt.co.uk
In this briefing
Mentioned
Key Intelligence
Key Facts
- 1Anthropic safety researcher Evan Hubinger wrote on X that he personally believes there is a greater than 10% chance AI could pose a species-ending risk within 10 years.
- 2Hubinger stated Anthropic 'does not yet have a plan to solve alignment for superintelligence' and is 'not clearly on track' to do so.
- 3UK Prime Minister Andy Burnham acknowledged 'risks to national security' from AI during Commons questions on 9 September 2026.
- 4Burnham elevated AI safety minister Kanishka Narayan to a Cabinet-level role, saying the move had lifted the level of government conversation about AI.
- 5The Financial Times reported that Anthropic withheld its latest model from the UK's AI Safety Institute, a leading body for testing AI risks.
- 6Anthropic researcher Jacob Coxon resigned, warning that the race toward self-improving superintelligence was 'gambling with our lives.'
We really do earnestly believe AI could kill all humans! I personally think it is >10% within the next decade. I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to.
Posted on X on 9 September 2026 in response to Jacob Coxon's resignation
Analysis
For AI builders and researchers, this is not merely a policy story — it is an alignment story. An Anthropic safety researcher publicly stated the lab has no plan to solve superintelligence alignment, while another resigned over an industry race to self-improving systems. The PM's response signals that frontier labs may soon face harder external testing demands.
On 9 September 2026, the UK's AI safety debate crossed a significant threshold. Prime Minister Andy Burnham told the House of Commons that artificial intelligence poses "risks to national security" after Evan Hubinger, a top safety researcher at Anthropic, wrote on X that he personally assigns a greater than 10% probability to AI killing all humans within the next decade. Hubinger's statement was not an isolated comment: it came in direct response to his colleague Jacob Coxon's resignation, in which Coxon warned that the industry's race toward self-improving superintelligence is "gambling with our lives." The same day, Burnham confirmed that AI safety minister Kanishka Narayan had been elevated to a Cabinet-level role, a structural change the PM said had "lifted the level of the conversation about artificial intelligence within Government."
The Financial Times reported that Anthropic withheld its latest model from the UK's AI Safety Institute, one of the world's leading bodies for pre-deployment risk testing.
The convergence of internal dissent at Anthropic and a policy response in Westminster marks a new phase in the global conversation about frontier AI risk. Anthropic has long marketed itself as one of the most safety-conscious AI labs, making the public alarm from its own researchers particularly notable. Hubinger was explicit: "I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to." That admission undercuts the lab's public confidence and suggests that even the organizations most institutionally committed to AI safety are struggling to keep pace with the technical challenge of aligning systems more capable than their creators. For the UK government, the timing is difficult to dismiss. The Financial Times reported that Anthropic withheld its latest model from the UK's AI Safety Institute, one of the world's leading bodies for pre-deployment risk testing. The combination of internal warnings and reduced external transparency creates pressure on policymakers to treat engagement from frontier labs as neither sufficient nor reliable.
The elevation of Kanishka Narayan to Cabinet is more than symbolic. It places AI safety at the same table as traditional national security portfolios, signaling that existential risk from machine intelligence is now considered a first-order state concern rather than a niche policy issue. Burnham's framing to the Commons was carefully calibrated: he acknowledged the risk without endorsing the most alarming probability estimates, and he extended an offer to colleagues on all sides to participate in the safety discussion. That bipartisanship may be essential if the UK is to maintain credibility as a hub for AI governance while also pressing firms like Anthropic to submit models for independent evaluation. The fact that a US-based lab chose to withhold a model from the UK's safety institute raises immediate questions about whether voluntary testing regimes are viable at the frontier, or whether governments will need to establish mandatory access conditions for developers operating at the highest capability thresholds.
What to Watch
The remarks from Hubinger and Coxon also carry implications for the AI research ecosystem and talent market. Public resignations over safety concerns create reputational risk for labs and may accelerate the flow of safety-minded researchers toward government bodies, independent nonprofits, or rival firms promising stronger alignment commitments. If credible internal researchers are willing to state a greater than 10% chance of human extinction within a decade, the burden of proof on labs to demonstrate concrete alignment progress rises sharply. At the same time, the probabilistic framing itself is unstable: a 10% chance of species-ending catastrophe is not a normal risk metric, and it may distort public debate if treated as a precise, settled figure rather than a subjective expression of uncertainty.
Forward-looking, this episode is likely to intensify three pressure points. First, Parliament may escalate scrutiny of whether UK safety testing has real enforcement power, especially when labs can refuse to share models. Second, Anthropic's leadership will face renewed questions about how its safety commitments translate into verifiable alignment milestones, not just rhetorical caution. Third, the global regulatory conversation may shift from broad principles toward concrete mechanisms such as mandatory pre-deployment evaluation, whistleblower protections for researchers who raise existential concerns, and international agreements on model access. The UK's decision to give its AI safety minister a Cabinet seat is an early signal that governments are preparing to treat refusal to cooperate on safety as a strategic problem. Whether that preparation produces enforceable policy or simply more elevated conversation remains the central uncertainty of the next decade.
Timeline
Timeline
Hubinger posts extinction risk warning
Evan Hubinger writes on X that he believes AI has a greater than 10% chance of killing all humans within the next decade, responding to Jacob Coxon's resignation.
PM Burnham addresses Commons on AI risk
Andy Burnham acknowledges national security risks from AI and confirms AI safety minister Kanishka Narayan has been given a Cabinet seat.
Source cluster
Primary reporting
- worcesternews.co.ukPM acknowledges AI risk as Anthropic expert says tech could kill all humans
Cite This Page
"Anthropic's >10% AI extinction odds trigger UK Cabinet shakeup." AI Intelligence Brief, September 9, 2026. https://getaibrief.com/story/anthropic-10-percent-extinction-odds-uk-cabinet
How we covered this story
Every story in our AI coverage is assembled from multiple primary sources, cross-referenced for factual consistency, and scored along three independent dimensions: sentiment, operational impact, and source-cluster confidence. Single-source rumors and unverifiable claims do not pass our editorial gate. When a story shows "Verified by N sources" with N≥2, the development is independently corroborated; when N=1, we mark it explicitly so readers can weigh the signal accordingly.
Impact scoring uses a 1-10 scale weighted toward regulatory, financial, and operational consequence rather than coverage volume. A topic that runs in every outlet but moves no real decisions ranks lower than a niche regulatory filing that reshapes how operators in the AI space have to behave. Read our full methodology for the scoring rubric, our glossary for term definitions, and our trends index for the longitudinal view across the beat.
Sources are only linked to a story once they clear our classification pipeline at a minimum 35 percent relevance threshold. According to that methodology, reviewed July 2026, this follows multi-source corroboration standards recommended by journalism research bodies such as the Reuters Institute for the Study of Journalism.
See something wrong in this story — a wrong fact, a broken source link, a misattributed entity? Report a data issue.
| Signal on this page | What it tells you |
|---|---|
| Verified by N sources | Independent corroboration count. N≥2 is our confidence floor; N=1 is marked explicitly. |
| Impact score (1-10) | Regulatory + financial + operational weight. 8+ signals an experienced-operator action item. |
| Sentiment | Five-tier classification trained on labeled AI-specific corpora. |
| Timeline | Where applicable, the related-events sequence that contextualizes today's development. |