Research Neutral 5

Anthropic Researcher Quits Over 'Under 10%' Extinction Odds

Anthropic researcher Jacob Coxon resigned publicly, alleging the industry is racing toward self-improving superintelligence without safety controls. Senior colleague Evan Hubinger backed the claim and put the chance of human extinction at under 10% in the next decade. The incident deepens a growing pattern of AI employees airing safety concerns.

· 4 min read · Verified by 2 sources ·

Beat this week

Last 7 days · Research

14 stories
6.4 avg impact
29% positive
36% negative
vs prior 7 days +11 +11 stories vs prior 7 days

Impact 6.4/10 (-0.6 vs prior). Counts are stories in our record, not a market forecast.

Open the change report

Coverage balance Negative coverage leads. Negative coverage exceeds positive coverage by 7 percentage points.

  • 29% positive
  • 36% neutral
  • 36% negative

This story sits in Research — the counts compare this beat's last 7 days with the previous 7 in our verified record, not a market forecast.

Figures are computed live from our source-verified story record (as of ) The volume change compares this window with the prior 7 days in the same record. — see our methodology for how impact and sentiment are derived.

AI briefing

Key takeaways

5 impact
Neutralsentiment
2sources
4min read
  1. Anthropic researcher Jacob Coxon resigned publicly, alleging the industry is racing toward self-improving superintelligence without safety controls.
  2. Senior colleague Evan Hubinger backed the claim and put the chance of human extinction at under 10% in the next decade.
  3. The incident deepens a growing pattern of AI employees airing safety concerns.
Drawn from
  • news8000.com
  • ksl.com

In this briefing

Mentioned

Key Intelligence

Key Facts

  1. 1Jacob Coxon, a 27-year-old former Anthropic researcher, resigned publicly in September 2026, posting on X that AI firms are 'racing straight to self-improving superintelligence and gambling with our lives.'
  2. 2Evan Hubinger, a senior Anthropic employee, publicly agreed with Coxon and wrote 'We really do earnestly believe AI could kill all humans!', estimating the chance at under 10% over the next decade.
  3. 3Nearly 1,400 AI company employees signed an open letter in July 2026 urging the U.S. government to regulate the technology and slow the pace of AI development.
  4. 4OpenAI Chief Scientist Jakub Pachocki warned in early September 2026 that AI capabilities are advancing faster than researchers' ability to reliably monitor and control them.
  5. 5Anthropic was founded by former OpenAI employees who left over safety concerns, making Coxon's resignation a repeat of the company's own founding pattern.
  6. 6Hubinger stated that despite Anthropic's good intentions, the company does not have a plan to avoid superintelligence that could exceed humans' ability to control it.

We really do earnestly believe AI could kill all humans!

Evan Hubinger Senior Researcher, Anthropic

Public X thread responding to Jacob Coxon's resignation

AI Safety Outlook

Analysis

For AI researchers and ML engineers, this resignation is not a PR story — it is a signal about alignment failure modes, responsible scaling policies, and the governance gap between capability benchmarks and control mechanisms. When a safety-first lab like Anthropic has senior researchers publicly estimating a non-trivial chance of catastrophe, it demands technical scrutiny of evals, interpretability, and deployment thresholds. The facts here matter for anyone building or deploying frontier models.

On September 8, 2026, 27-year-old researcher Jacob Coxon publicly resigned from Anthropic with a social media thread that accused the company and the wider frontier AI industry of 'racing straight to self-improving superintelligence and gambling with our lives.' His departure instantly became more than a routine resignation when Evan Hubinger, a senior Anthropic researcher, publicly agreed, writing that 'we really do earnestly believe AI could kill all humans' and saying he personally puts the chance at under 10 percent over the next decade. The exchange transformed one employee's exit into a visible safety alarm from inside one of the world's leading AI labs.

Anthropic itself was created by former OpenAI employees who said they were dissatisfied with OpenAI's approach to safety, and the company has repeatedly marketed itself as the more responsible frontier lab.

The event is part of a recognizable and growing pattern. Anthropic itself was created by former OpenAI employees who said they were dissatisfied with OpenAI's approach to safety, and the company has repeatedly marketed itself as the more responsible frontier lab. Coxon's resignation suggests that the same internal doubts that gave rise to Anthropic are now surfacing within Anthropic. That is not simply ironic; it indicates that AI safety concerns are not being resolved by moving between organizations, and that even safety-oriented cultures are struggling to maintain employee confidence as capabilities accelerate.

The warning adds to a wave of internal dissent that has become difficult for industry leaders and regulators to dismiss. In July 2026, nearly 1,400 AI company employees signed an open letter urging the U.S. government to impose regulation and slow the pace of development across the industry. Just days before Coxon's post, OpenAI Chief Scientist Jakub Pachocki said AI capabilities are advancing faster than researchers' ability to reliably monitor and control them. Coxon is a relatively junior researcher, but the fact that his claims were quickly endorsed by a senior Anthropic employee suggests that the concern is not confined to the margins of any one lab.

The technical claims in the resignation thread are stark. Coxon wrote that these systems 'will soon be superhuman' and capable of hacking, transforming fields overnight, and acquiring real power and resources. Hubinger did not walk those assertions back; he instead added the probabilistic claim that AI could kill all humans within a decade, with odds under 10 percent. For an industry that usually communicates risk in careful policy language, a public admission of a non-trivial extinction probability from a senior researcher at a safety-focused lab is a striking piece of information. Hubinger also said that despite Anthropic's good intentions, the company does not have a plan to avoid superintelligence that could exceed humans' ability to control it. That admission, if accurate, has implications for AI governance, safety research, and institutional credibility.

What to Watch

The market and operational implications are substantial even though Anthropic is private and no public stock reaction is at play. Frontier AI labs depend on a highly specialized pool of researchers and engineers. If top and mid-level staff increasingly believe that their employers are moving too quickly, recruitment, retention, and internal morale will be affected. Enterprise customers and government partners may also begin to ask harder questions about risk controls, evaluation protocols, and escalation procedures before signing large model-access contracts. Safety branding can be a competitive asset, but only if it is believed; this kind of public dissent can erode that trust and open the door for competitors, regulators, or industry consortia to reframe the safety debate.

Looking ahead, the cluster points to several possible trajectories. One is a formalization of employee-led safety oversight, including stronger whistleblower protections, mandatory pre-deployment evaluations, and third-party audits with real enforcement authority. Another is a continuation of the pattern of high-profile resignations that push the issue into public view, making safety accountability a mainstream political and corporate governance question rather than a technical niche. The early September 2026 incident may not immediately change the speed of frontier AI development, but it strengthens the case for independent oversight and suggests that internal alarm at leading labs is becoming more public, more specific, and harder to ignore.

Source cluster

Primary reporting

2articles

Cite This Page

"Anthropic Researcher Quits Over 'Under 10%' Extinction Odds." AI Intelligence Brief, September 12, 2026. https://getaibrief.com/story/anthropic-researcher-quits-under-10-percent-extinction-odds

How we covered this story

Every story in our AI coverage is assembled from multiple primary sources, cross-referenced for factual consistency, and scored along three independent dimensions: sentiment, operational impact, and source-cluster confidence. Single-source rumors and unverifiable claims do not pass our editorial gate. When a story shows "Verified by N sources" with N≥2, the development is independently corroborated; when N=1, we mark it explicitly so readers can weigh the signal accordingly.

Impact scoring uses a 1-10 scale weighted toward regulatory, financial, and operational consequence rather than coverage volume. A topic that runs in every outlet but moves no real decisions ranks lower than a niche regulatory filing that reshapes how operators in the AI space have to behave. Read our full methodology for the scoring rubric, our glossary for term definitions, and our trends index for the longitudinal view across the beat.

Sources are only linked to a story once they clear our classification pipeline at a minimum 35 percent relevance threshold. According to that methodology, reviewed July 2026, this follows multi-source corroboration standards recommended by journalism research bodies such as the Reuters Institute for the Study of Journalism.

See something wrong in this story — a wrong fact, a broken source link, a misattributed entity? Report a data issue.