FULL STORY
Anthropic Resignations Ignite AI Extinction Risk Firestorm
A wave of resignations from Anthropic researchers warning of catastrophic AI risks sparked political and media fallout, a criticized official response, and a string of insider risk estimates ranging from 10% to 70%.
2026-09-09 ~ 2026-09-11 · 11 episodes · 258 posts
Episode 1 · Anthropic Researcher Resigns Warning Labs Are "Gambling With Our Lives" (2026-09-09, 217 posts)
A string of researchers from OpenAI and Anthropic have resigned and publicly voiced concerns about the dangers of AI, sparking heated debate in the AI safety community over whether "quitting and speaking out" is the best way to deliver safety warnings. The current discussion focuses not on the content of the warnings, but on why the act of resigning carries unique signal value and credibility.
Confirmed
- Kelsey Tuoc pointed out that the huge attention drawn by Anthropic employees leaving and publicly saying "what we're doing is dangerous and not worth it" should give pause to other worried employees who believe resigning is pointless; allTheYud went further, calling for another Anthropic capabilities researcher to do the same now.
- A view relayed by birchlse holds that safety warnings from those still inside a company are easily dismissed as marketing or PR theater, whereas alarms from those who have quit cannot be casually attributed to business motives—this is the key source of their credibility. He also argued that resigning is "an honest signal of serious risk that people outside AI can also understand," more intuitive than technical argumentation.
- jachiam0 explained why these resignation warnings keep going viral: the warning is aimed not at any one or two labs but at the entire industry—many who resign promptly join another lab, which comes across as cheap and hollow; moreover, timing and credibility align, and public receptiveness has been latent all along.
Unconfirmed
- zck is skeptical, arguing that every tech cycle produces people who make a name by "whistleblowing"—first from Meta, now from OpenAI and Anthropic. He sees this not as an organized manipulation campaign but as the liberal "New York Times class" having an endless appetite for anti-tech testimony, with lab employees' sincere idealism being exploited. This remains one individual's skeptical take.
Why it matters
- The debate touches a core question of AI safety governance: whether internal whistleblowing or public resignation is more effective at transmitting risk signals. Resignation's "cannot be attributed to business motives" quality makes it the most intuitive narrative for the public and the outside world to understand AI risks, and puts worried employees inside labs under real pressure over whether to "vote with their feet."
- WSJ: Anthropic researcher quits AI industry over fears of uncontrollable AGI race — peterwildeford · 2026-09-09
- Anthropic researcher quits AI industry over fears of out-of-control race, WSJ reports — Miles_Brundage · 2026-09-09
- Researcher quits Anthropic after 3 years in pretraining at OpenAI and Anthropic, blasting both labs — peterwildeford · 2026-09-09
- Exec quits over industrywide rush to build self-improving AI, citing humanity-ending risk — Miles_Brundage · 2026-09-09
- Ex-pretraining researcher quits Anthropic, accusing OpenAI and Anthropic of racing recklessly to superintelligence — peterwildeford · 2026-09-09
- Anthropic researcher Jacob Coxon quits AI over fears self-improving models could go uncontrollable by 2027 — rohanpaul_ai · 2026-09-09
- Anthropic researcher Jacob Coxon quits AI over fears self-improving models could go uncontrollable by 2027 — rohanpaul_ai · 2026-09-09
- Anthropic Researcher Quits Over AI Fears, WSJ Reports — Bubbly-Air7302 · 2026-09-09
- Anthropic researcher Jacob Coxon resigns, warning superintelligence could kill us all by 2030 — Polymarket · 2026-09-09
- Anthropic researcher: >10% chance AI kills all humans within a decade, alignment unsolved — EvanHub · 2026-09-09
- Anthropic researcher Jacob Coxon quits AI industry over out-of-control AGI fears — AGI Hunt · 2026-09-09
- Anthropic researcher: >10% chance AI kills all humans within a decade, no alignment plan — AICopyLab · 2026-09-09
- Researcher quits Anthropic after 3 years: both OpenAI and Anthropic are 'racing to self-improving superintelligence' — Miles_Brundage · 2026-09-09
- Anthropic researcher: >10% chance AI kills all humans within a decade, no alignment plan yet — EvanHub · 2026-09-09
- Anthropic-affiliated researcher: >10% chance AI kills all humans within a decade, no alignment plan yet — ramagetime · 2026-09-09
- Anthropic Employees Can't Stop Predicting AI Will Eliminate All Jobs, Joke Goes — toptickcrypto · 2026-09-09
- Anthropic Researcher: >10% Chance AI Kills All Humans Within a Decade — ccerrato147 · 2026-09-09
- Anthropic researcher: >10% chance AI kills all humans, no alignment plan yet — JosephJacks_ · 2026-09-09
- Anthropic's Evan Hubinger puts >10% odds on AI killing all humans within a decade — Bonecondor · 2026-09-09
- Anthropic Researcher Jacob Resigns, Sparking Sarcastic Debate Over AI Safety Exit Strategy — ctjlewis · 2026-09-09
Episode 2 · Anthropic's doomsday talk ahead of IPO draws wave of mockery (2026-09-09, 10 posts)
As Anthropic approaches its IPO, executives' public statements that AI "could destroy humanity" have drawn a wave of mockery and pushback on X, with the fight centered on the contradiction between AI safety narratives and profit-seeking.
Confirmed
- Marc Andreessen (relayed by beffjezos) mocked Anthropic employees for claiming their product "might kill us all" while investors excitedly line up to buy the stock; developer Chetan Puttagunta quipped that pre-IPO statements have escalated from "destroy most jobs" to "destroy most companies" to "potentially kill all of humanity," joking that the next step is "eating the universe"
- a16z partner John Allard scoffed that "sincere conviction outweighing financial returns" is impossible, aiming squarely at existential-risk statements issued right before going public
- Responding to the defense that lab employees genuinely believe the work is dangerous but that responsible people should do it, with enterprise contracts and an IPO merely serving as fundraising, intellectrona asked: if so, why do employees still get equity incentives
- Azeem proposed a test: if Anthropic is intellectually honest, it should put extinction risk in its S-1 filing; another user (relayed by m1) noted that omitting it could constitute misleading disclosure, giving investors grounds to sue
- Taylor Lorenz criticized such self-aggrandizing doom talk for merely scaring the public and producing bad regulation, urging executives to be more precise and proactively hedge regulatory consequences; MoonL88537 strongly agreed, arguing that if they truly believe there's even a 1% chance the company destroys humanity, the right move is to report it to authorities immediately (e.g., the FBI) with concrete evidence, not issue vague warnings
- Investor firstadopter cited similar "destroy humanity" remarks by OpenAI executive Evan, calling such statements weeks before a trillion-dollar IPO "braindead" — if you truly believe it, resign immediately — and noting it hands ammunition to opponents and politicians attacking AI infrastructure buildout
Why it matters
- The fight has pushed "is AI safety discourse just doom marketing?" to the forefront: the same statements are seen as responsible warnings within safety circles, but as pre-IPO self-contradiction or even a regulatory risk source by investors and some journalists
- The S-1 disclosure proposal offers an actionable test: whether a company's risk rhetoric matches its formal statements to investors could directly shape future litigation and regulatory conditions
- a16z partner mocks pre-IPO AI firms warning their product 'will kill us all' — john__allard · 2026-09-09
- Taylor Lorenz slams alarmist AI exec post for scaring public into bad policy — sebkrier · 2026-09-09
- Joke With Teeth: Anthropic's AI Doom Warnings Should Appear in Its IPO Prospectus — examachine · 2026-09-09
- AI x-risk debate: if you truly believe your lab endangers humanity, go to the FBI with receipts — MoonL88537 · 2026-09-09
- Debate: If Anthropic Believes Its Own Doom Risk, It Should Call the FBI — MoonL88537 · 2026-09-09
- OpenAI exec's doomsday remark weeks before trillion-dollar IPO sparks investor backlash — GarrisonLovely · 2026-09-09
- Anthropic Employees Warn AI Could Kill All Humans as Firm Eyes $2.5T IPO — scaling01 · 2026-09-09
- The equity problem: why AI-lab doomer employees profiting from stock undermines the safety narrative — intellectronica · 2026-09-09
- Marc Andreessen mocks Anthropic's doom marketing: investors just want to buy the stock — beffjezos · 2026-09-10
- As IPO Nears, Anthropic's Doom Claims Draw Mockery: "Next They'll Eat the Universe" — chetanp · 2026-09-10
Episode 3 · AI Insiders Fear Loss of Control as Anthropic Races Toward IPO (2026-09-09, 2 posts)
A viral thread documents the AI industry's collective fear of racing toward self-improving systems, arguing that competing on an unrecoverable technology while the finish line is an IPO roadshow makes the race itself fundamentally wrong.
- Thread: insiders' shared fear as Anthropic races toward self-improving AI ahead of ~$2T IPO — eyishazyer · 2026-09-09
- Racing toward a system you can't recall, with an IPO as the finish line: the race itself is the mistake — eyishazyer · 2026-09-09
Episode 4 · AI Extinction Warnings Spark UK Parliament Debate (2026-09-09, 2 posts)
Anthropic safety experts told UK lawmakers that AI has a greater-than-10% chance of causing human extinction, with warnings it could happen by 2030, sparking fierce parliamentary criticism of AI companies and renewed regulatory debate.
- Anthropic safety expert puts odds of AI wiping out humanity above 10% as UK Parliament debates ASI ban — nordicinst · 2026-09-09
- UK MP Al Carns: AI has a 10% chance of wiping out humanity, unchecked rollout 'irresponsible' — connoraxiotes · 2026-09-10
Episode 5 · Anthropic Alignment Lead Cites 10% Extinction Risk; US Lawmaker Proposes Five AI Rules (2026-09-10, 2 posts)
US Rep. Ro Khanna cited an Anthropic alignment lead's estimate of roughly 10% extinction risk from AI, proposing five regulatory measures including a federal AI agency modeled on nuclear and aviation oversight, pre-certification of containment capabilities, and kill switches.
- Anthropic alignment lead puts AI extinction risk at 10%; lawmaker proposes 5-point federal oversight plan — ShakeelHashim · 2026-09-10
- US lawmaker proposes nuclear-style federal AI agency with kill switches and criminal penalties — ShakeelHashim · 2026-09-10
Episode 6 · Anthropic Researcher Says AI Developers Believe Extinction Risk Is Real (2026-09-10, 2 posts)
An Anthropic researcher's personal essay argues that AI developers genuinely believe their technology could cause human extinction within years, and that the industry must articulate concrete asks to governments — sparking heated debate over AI risk narratives.
- Anthropic staffer on AI-risk open letter: industry must say what it wants from government — clarejtbirch · 2026-09-10
- Anthropic researcher: AI devs believe their tech could cause human extinction within years — xuanalogue · 2026-09-10
Episode 7 · Ex-OpenAI/Anthropic researcher's doom-laden resignation post goes viral, then faces backlash (2026-09-10, 9 posts)
On September 10, the account @hilbertspaess posted that they were leaving Anthropic that day, claiming to have spent the past three years doing pretraining research at OpenAI and Anthropic, and accusing both companies of recklessly racing toward self-improving superintelligence — "gambling with our lives." The post went viral, reportedly drawing over 100 million views and dominating that day's AI discourse.
Confirmed
- @hilbertspaess published the departure post, claiming three years of pretraining research across OpenAI and Anthropic, and asserting that "everyone inside" Anthropic knows its AI could end humanity
- The post spread massively, with users reporting over 100 million views
- The author previously worked at OpenAI, and reportedly complained that OpenAI colleagues largely dismissed their doomer views
Not Yet Confirmed
- Some users pointed out that their tenure at Anthropic lasted only about six weeks (other accounts say two months), but the exact duration cannot be independently verified
- Multiple posters suspect the account is a sockpuppet and the whole episode a staged drama, though these remain speculation with no hard evidence
- Claims such as "three years of pretraining research" and an "internal Anthropic consensus" remain unverified
Why It Matters
- The episode highlights the reach and credibility problems of AI doom narratives: a single explosive accusation can rack up over 100 million views without any credentials to back it up
- The swift community backlash — fixating on the whistleblower's tenure length rather than the substance of the claims — reflects the intense battle over positioning and source credibility in AI safety debates, with both sides more inclined to question motives than engage the arguments
- If the account turns out to be a stunt, it will further erode public trust in "former employee whistleblower" content
- Doomer warning 'AI will kill everyone' revealed to have worked at Anthropic for just two months — TheMoonMidas · 2026-09-10
- Skeptic mocks ex-Anthropic doomer: he only worked there a couple months — TheMoonMidas · 2026-09-10
- The viral 100M-view AI doom post came from someone who lasted six weeks at Anthropic — Dr_Singularity · 2026-09-10
- Ex-OpenAI/Anthropic Researcher's Viral Resignation Post Suspected to Be a Psyop — TinfoilTricorn · 2026-09-10
- Ex-OpenAI/Anthropic pretraining researcher quits, says labs are 'gambling our lives' on superintelligence — inductionheads · 2026-09-10
- Ex-OpenAI/Anthropic researcher quits, says both labs are gambling lives on superintelligence race — SatelliteNetSec · 2026-09-10
- Anthropic researcher Jacob Coxon quits, says it's 'crunch time for humanity' on AI safety — DavidLinthicum · 2026-09-11
- Ex-OpenAI/Anthropic pretraining researcher resigns, says labs are gambling lives on superintelligence — ShakeelHashim · 2026-09-11
- Ex-Anthropic researcher quits over 'reckless' superintelligence race, fueling AI doomerism debate — benfielding · 2026-09-11
Episode 8 · Anthropic's Risk Statement Slammed as Corporate PR (2026-09-10, 5 posts)
On September 10, Anthropic issued an official statement addressing the AI-safety controversy that erupted over the past 24 hours, saying AI will bring both enormous benefits and unprecedented risks. The company said it has always been transparent, that its models come with the industry's strongest safety guardrails, and that it pioneered the Responsible Scaling Policy (RSP) and mechanistic interpretability research—a field now being used to analyze and prevent alignment failures.
Confirmed
- Anthropic's official statement was shared and amplified by journalist Hadas Gold; the wording emphasizes "enormous benefits alongside unprecedented risks," but stays vague on the specifics of those risks.
- In the statement, Anthropic highlights its pioneering role in mechanistic interpretability and the Responsible Scaling Policy.
Unconfirmed
- The details of the specific safety incident referenced (the claims made by Hubinger and Coxon) were not elaborated in the post, and observers remain divided over whether the company's risk framing is internally consistent.
Why it matters
- Shakeel Hashim wrote a scathing critique, calling the statement "really bad": it devotes most of its space to self-congratulation, fails to discuss risks candidly, and doesn't restate the explicit claims of researchers Hubinger and Coxon—arguing that a company known for candor delivered cowardly "corporate speak."
- Others in the safety community similarly questioned Anthropic's image as a safety pioneer, saying the statement dodged substance.
- There are dissenting views too: @menhguin argues Anthropic has been consistent on risk, with Dario and other co-founders' public statements over the years aligning closely with the official line, showing no sign of internal policy conflict; Hesamation raised questions about whether "talking about risk" and "their own policies" are consistent—and the debate continues.
- Anthropic's risk statement slammed for self-praise over frank talk on AI risks — ShakeelHashim · 2026-09-10
- Anthropic slammed for "corporate comms" statement that praises itself instead of stating AI risks — ShakeelHashim · 2026-09-10
- Anthropic issues safety statement amid controversy; investors mock the bland language — Miles_Brundage · 2026-09-10
- Anthropic's vague AI-risk statement slammed as 'mealy-mouthed corporate speak' — sjgadler · 2026-09-10
- Hesamation Knocks Anthropic Over Gap Between Risk Talk and Actual Behavior — Hesamation · 2026-09-10
Episode 9 · Wildeford Urges Anthropic to Replace PR Team with Researchers (2026-09-10, 3 posts)
Peter Wildeford criticized Anthropic's official statements as overstating safety confidence, saying internal researchers disagree, and urged the company to let researchers speak directly instead of PR staff. He even used a Claude model to evaluate Anthropic's statement, which called it 'a resume rather than a response.'
- Peter Wildeford: Anthropic should ditch comms team, let researchers speak authentically — peterwildeford · 2026-09-10
- Wildeford: Anthropic should let researchers, not comms, speak authentically — peterwildeford · 2026-09-10
- Wildeford: Anthropic's Corporate Safety Messaging Doesn't Match Reality — peterwildeford · 2026-09-10
Episode 10 · Seven AI Insiders Warn of Extinction Risk Within a Decade (2026-09-10, 4 posts)
Within four days, seven AI insiders issued stark warnings: an Anthropic researcher put extinction risk within a decade above 10%, while an OpenAI employee cited 70% odds within three years without regulation, a figure other researchers disputed but still saw exceeding 10%.
- Seven AI insiders warn in four days that AI could kill everyone, citing extinction fears — sebpaquet · 2026-09-10
- Anthropic's Evan Hubinger says AI could kill all humans with >10% odds this decade — wfithian · 2026-09-10
- OpenAI Employee Puts 70% Odds on Human Extinction Within 3 Years Without AI Regulation — peterwildeford · 2026-09-11
- Researcher Pushes Back on OpenAI Employee's 70% Extinction Claim, Still Sees >10% Risk — peterwildeford · 2026-09-11
Episode 11 · Ex-Anthropic researcher warns AI could 'kill us all' on CNN (2026-09-10, 2 posts)
Former OpenAI and Anthropic researcher Jacob Coxon appeared on CNN's Anderson Cooper to warn AI could "kill us all" after publicly resigning from Anthropic, though some commentators dismissed the remarks as recycled doomsaying.
- Former Anthropic Researcher Tells CNN AI Could 'Kill Us All' — Bubbly-Air7302 · 2026-09-10
- Ex-OpenAI/Anthropic Researcher's AI Safety TV Warning Dismissed as Nothing New — deliprao · 2026-09-10