"I resigned from Anthropic today. I spent the last three years doing pretraining research at both OpenAI and Anthropic. Neither company is acting responsibly. They are racing straight to self-improving superintelligence and gambling with our lives. More thoughts below."
On 8 September 2026 (US time) a researcher almost nobody had heard of posted seven paragraphs and quit his job, and by the end of the week the story had reached the network news. The single most repeated claim about Jacob Coxon, that he put the odds of AI killing everyone above ten percent, is not something he said. His thread contains no numbers at all.
That matters for reasons beyond pedantry. The number belongs to somebody else, somebody who did not resign, and understanding who said what turns a scary headline into a considerably more interesting story. This guide from Best Answer Hub sets out the full text of what Coxon wrote, separates his claims from those attributed to him, and covers what has been verified about him and what has not.
Why did Jacob Coxon resign?
Because he believes both laboratories he worked for are racing toward self-improving artificial intelligence without an adequate plan, and he no longer wanted to take part. In his own words, in the first of seven posts: "I resigned from Anthropic today. I spent the last three years doing pretraining research at both OpenAI and Anthropic. Neither company is acting responsibly. They are racing straight to self-improving superintelligence and gambling with our lives."
His specific objection is not that the technology is evil. It is that a decision of this consequence is being made in the wrong place, by people with no mandate to make it. That argument lands in his fifth post, and it is the line most likely to outlive the news cycle.
Accepting this race and entering the endgame is a hubristic gamble that should not be launched from a private company's Slack.
What did Jacob Coxon say about AI?
The thread ran to seven posts. Because so much coverage has paraphrased it into something it is not, here it is as written, in order.
"Do not underestimate the power of this technology. These will soon be superhuman systems that can hack anything, revolutionize any field overnight, and acquire real power and resources. We have all witnessed the progress in each of these domains, and progress is not slowing."
"The people building AI earnestly believe that it could kill us all by the end of the decade. This is not a marketing stunt. If anything, many executives and senior researchers will couch their phrasing in the press to sound sensible, but I hear the same people express fear privately. No other human activity poses this level of danger."
"A common response is 'if they truly believe this, why are they still building it?' At OpenAI, many have not deeply internalized the civilizational stakes. At Anthropic, the stakes are well-understood, but they are locked in a race to get there first. They believe no one else will act responsibly, so they must do it themselves, despite the risk."
"Accepting this race and entering the 'endgame' is a hubristic gamble that should not be launched from a private company's Slack. Attempting to speedrun alignment should require extraordinary confidence that there are no better trajectories available."
"I am optimistic about the potential for coordination. Warning shots like the Hugging Face attack have made pacing agreements between U.S. labs more viable. I don't feel like we're on track to prevent a global race, which may require costly actions such as a temporary ban on improving model capabilities."
"If you are a lab researcher, I urge you to consider what the next few years will actually feel like. Do you want to kick off a superintelligent RL run without a rigorous understanding of its mind? Should you put your head down because 'it's happening anyway' or take this moment to call for different conditions?"
Read it as a whole and the shape is different from the headlines. Post three is the load-bearing one, and it is not a prediction. It is testimony about other people: what colleagues say privately versus what they say to journalists. He is reporting a credibility gap, not forecasting a date.
Did Jacob Coxon say there is a 10% chance AI kills everyone?
No. He gave no percentage, no probability and no statistic of any kind. The figure attached to his name belongs to Evan Hubinger, who leads alignment science at Anthropic, who did not resign, and who still works there. Hubinger posted it as a reply later the same day, and hedged it explicitly as personal.
| Claim | Who said it | Exactly |
|---|---|---|
| "Racing straight to self-improving superintelligence" | Coxon | Post one of his resignation thread |
| Insiders privately fear AI could kill everyone "by the end of the decade" | Coxon | Post three, reporting what colleagues say, not his own forecast |
| ">10% within the next decade" | Hubinger | "I personally think it is >10% within the next decade." A reply, explicitly personal, from a serving employee |
| "We do not yet have a plan to solve alignment for superintelligence" | Hubinger | Same reply. An Anthropic alignment lead saying so in public |
| Extinction "by 2030" | Nobody | A headline artifact. Neither man wrote the year 2030 |
Here is how the error spread, because the mechanism is worth recognizing. Both CBS News and CNBC ran the figure under headlines saying a "researcher" said it. Both article bodies attribute it correctly to Hubinger. Aggregators copied the headline rather than the body, and within a day the number was Coxon's in dozens of places. The original error was ambiguity, not invention, and it still ended up as a false quote.
Coxon's central claim was that senior people say one thing in public and a more frightening thing in private. The same day, a serving Anthropic alignment lead replied in public with a number above ten percent and the words "we really do earnestly believe AI could kill all humans." That is not a rebuttal of the resignation. It is confirmation of its main allegation, from inside the company, on the record. The misattribution has obscured the strongest fact in the whole episode.
What did Jacob Coxon do at Anthropic?
Pretraining research, for about four months. That last part is where most coverage goes wrong, and the confusion is understandable because his own first sentence invites it.
He wrote that he spent "the last three years doing pretraining research at both OpenAI and Anthropic." Several outlets rendered that as three years at Anthropic. Axios, which interviewed him, reports he was at Anthropic for four months. The three years is the two jobs added together, which puts roughly two years and eight months at OpenAI and four months at the company he denounced.
No outlet reports a job title for him at either laboratory. "Pretraining researcher" is his own description of the work, not a title, and "core researcher" appears only on a stock-analysis site that paired it with speculation about an Anthropic listing. Separately, Fast Company reported that neither Anthropic nor OpenAI confirmed his employment when asked. Treat any article that gives him a precise title as having made it up.
The split matters for how you weigh him. As an Anthropic insider he is thin, four months in a pretraining role. As an OpenAI alumnus he is substantial, and his harshest judgment is aimed there: at OpenAI, he wrote, "many have not deeply internalized the civilizational stakes," while at Anthropic "the stakes are well-understood" but the company is trapped in a race. Coverage that flattens this into "both companies bad" loses his actual argument.
Why did Jacob Coxon leave Anthropic when he did?
He left before any of his equity vested, which he raised himself as a reason to believe him. Anthropic employees reach their first vesting cliff at six months. He resigned at four, and told Axios: "I no longer have anything to gain by juicing up Anthropic's valuation. I left before any of my equity vested."
Two precisions are worth keeping, because the headline version overstates it in one direction and understates it in another. He did not give up vested equity, because he had none yet. He forfeited the opportunity to reach the cliff, which is a real cost but a different one. And he still holds equity in OpenAI, so he is not financially neutral about the sector, only about the company he was criticizing. His own phrasing was narrower and more accurate than most of the coverage of it.
This story is often paired with the OpenAI equity controversy of May 2024, where departing staff faced losing already vested equity unless they signed a lifelong non-disparagement clause. That is a different thing: vested versus unvested, and an agreement versus a vesting schedule. No non-disparagement agreement or NDA has been verified in Coxon's departure at all. Drawing the parallel loosely gets it wrong.
What did Anthropic say in response?
It issued a general statement through an unnamed spokesperson. No named executive responded. The statement says the company has "always been transparent that AI will bring both enormous benefits and unprecedented risks," and that "to address these risks, we continue to build models with some of the strongest safeguards in the industry," citing its work on mechanistic interpretability. In a statement reported by the Associated Press, it also said the world would benefit from a "lawful, verifiable way to work together to pace how we release powerful models."
Notice what the statement does not do. It does not name him. It does not dispute a single factual claim in his thread. It does not confirm or deny that he worked there. Several outlets, including TIME, reported that neither Anthropic nor OpenAI responded before their deadlines on the day; Anthropic's statement arrived later, and OpenAI's position remains that it has not commented on the resignation.
On 12 September, four days after the thread, Anthropic's chief executive Dario Amodei published an essay calling for exactly the kind of slowdown Coxon asked for, stating "We must slow the pace at which we improve the capabilities of AI models." Sam Altman said he agreed. No source states that these events are connected, and this guide does not claim they are. The dates are 8, 10 and 12 September, and you can draw your own conclusions.
Is Jacob Coxon a doomer?
He says not, and his thread supports him. Post six opens "I am optimistic about the potential for coordination," and he told Axios that "the word doom is kind of silly." He is not predicting the end. He is making a policy argument, and it has a specific ask in it.
- 1Pacing agreements between US laboratories. His central proposal, and he argues recent security incidents have made them "more viable" rather than less.
- 2A temporary ban on improving model capabilities, if needed. He calls this a "costly action" and raises it as something a global race might require, not as a first resort.
- 3Researchers inside the labs to speak up. His seventh post is addressed to colleagues, asking whether they will "put your head down because it's happening anyway" or "call for different conditions."
- 4The decision taken somewhere other than a company. The Slack line is the whole argument compressed: this is a question about authority and mandate, not about whether the technology works.
Whether you agree or not, that is a governance position rather than a prophecy, and it is the part of the story that almost every headline dropped in favor of the number he did not give.
Who else has left an AI lab over safety?
Two more researchers went public two days later, and they are frequently merged with Coxon into a single event. They are related but distinct, and being precise about it matters if you want to judge whether this is a pattern.
| Who | Lab and role | What happened |
|---|---|---|
| Jacob Coxon | Anthropic, pretraining research, previously OpenAI | Resigned and posted the same day, 8 September 2026 (US time). Leaving the industry. |
| Joe Benton | Anthropic, led a safety research team | Recently left, went public 10 September in an NBC News interview. Joined METR. |
| Josh Engels | Google DeepMind, AGI safety team | Also recently left, went public 10 September in a separate interview for the same NBC News report: "There are no adults in the room." Joined METR. |
Benton and Engels spoke to the same outlet on the same day, so they are one coordinated disclosure rather than two independent signals. Calling all three a spontaneous wave overstates it. Three frontier-lab researchers going public in one week is still notable, and two of the three went to the same evaluations organization rather than out of the field.
Anthropic's AI safety lead Mrinank Sharma resigned in February 2026, seven months earlier, writing that "the world is in peril." Headlines reading "Anthropic's AI safety head just resigned" are from that event, not this one. Coxon was never a safety lead, and Evan Hubinger, who does lead alignment science, has not resigned at all.
How good is your AI knowledge, really?
Stories like this one turn on whether you can tell a measured claim from a stated one, and a person's opinion from a finding. The free Best Answer Hub AI Knowledge Test checks what you actually know across five domains in about 20 minutes. No signup, no email, nothing uploaded.
Take the free AI Knowledge TestCommon questions about the Coxon resignation
Sources
- Axios, Madison Mills, Anthropic whistleblower gave up his equity to leave the company, 9 September 2026 (the interview, and the origin of the four-month and equity facts).
- TIME, Harry Booth, He helped build powerful AI at OpenAI and Anthropic. Now he is afraid it could kill us, 9 September 2026.
- TechCrunch, Rebecca Bellan, Gambling with our lives: Anthropic researcher quits, warns against self-improving AI, 9 September 2026.
- CBS News, Megan Cerullo and Faris Tanyos, Ex-Anthropic researcher Jacob Coxon warns AI could grow 'smart enough to kill us', updated after publication to carry the Anthropic spokesperson statement.
- CBS News, AI could kill all humans, Anthropic researcher says, 9 September 2026 (the Hubinger quote; note the headline says researcher while the body attributes it correctly).
- Associated Press via PBS NewsHour, Kaitlyn Huamani, Anthropic researcher's resignation sends warning about the dangers of AI development, 9 September 2026 (AP's updated version of this story adds Anthropic's later statement, including the "lawful, verifiable" line).
- NBC News, Jared Perlo, Two AI researchers leave Anthropic, Google over safety concerns, 10 September 2026 (Joe Benton and Josh Engels).
- Forbes, Conor Murray, Anthropic AI safety researcher warns of world in peril in resignation, 9 February 2026 (the earlier Mrinank Sharma departure).
- Dario Amodei, We Must Pace the Frontier, September 2026.
- Axios, Amodei calls for pacing the frontier, 12 September 2026.
- Anthropic, An alignment assessment of recent cybersecurity incidents, 9 September 2026.
People also read these
More from Best Answer Hub: How good is your AI knowledge, How to tell when AI is wrong, The biggest myth about AI, and the free AI Knowledge Test.