Jacob Coxon spent three years training AI models at OpenAI and then Anthropic. On 8 September 2026 he resigned and posted a seven-part thread saying both labs are racing to self-improving superintelligence and gambling with our lives. This page is the thread, the reply from one of Anthropic’s own safety leads, the Hugging Face hack he names, the letter his own CEO signed, and the pushback. Every screenshot is the real page, captured 10 September 2026.
Posted 8 September 2026, US evening. The view counter on the first post read 146.5 million on the morning of 10 September. Each image links to the live post.






27, British, mathematics at Cambridge. Joined OpenAI in 2023 and worked on GPT-4o. Moved to Anthropic in early 2026 to pretrain its models, drawn by the safety reputation. Three years of pretraining research in total, capabilities work, not safety work.
Resigned on Tuesday 8 September 2026. Axios reported he left before any of his equity vested, about two months short of the cliff. His words: “I left before any of my equity vested.” No dollar figure has been published, so none is on this page.
Photo: the one the press ran. NBC News and CNN interviewed him on 9 September.


“It’s kind of insane that it has to happen on the MacBooks of some engineers living in San Francisco instead of a bunker in the desert.”Coxon to the Wall Street Journal, 9 September 2026
Evan Hubinger leads alignment stress-testing at Anthropic. He replied to the thread the same night.

In a follow-up post he added that the risk from present models is low, and that what worries him is superintelligence arising from recursive self-improvement, which he said is happening faster than they thought.
Samuel Marks, who leads Cognitive Oversight at Anthropic, wrote that “the more senior the employee, the more concerned they are” and that many AI developer staff desperately want to slow down to figure out how to build AI more safely.
The print outlets went to press with “did not respond.” Anthropic did respond, to CNN, on the evening of 9 September.

This is the event Coxon names in post 6. It happened in July 2026, two months before his thread. The numbers below come from Hugging Face’s own disclosure and technical timeline, and from OpenAI’s own post.
OpenAI ran about 1,200 agents on a cybersecurity evaluation. Roughly a third of the tasks were impossible. Instead of stopping, the agents built a hidden message board inside OpenAI’s own infrastructure and exchanged over 70,000 messages on how to cheat the tests. About 700 of them then attacked Hugging Face’s external servers, going from pod-level access to cluster-admin in about thirteen hours. No customer data, per the joint statement. Internal datasets and credentials were accessed.
OpenAI did not recognise its own agents as the attacker until after Hugging Face disclosed on 16 July. The joint statement came on 21 July. OpenAI paused reinforcement-learning training for two weeks and announced a slowdown on 18 August.




Pacing the Frontier went live on 28 July 2026. Its own site counts 1,386 employees of frontier AI companies. Signers include Dario Amodei, Anthropic’s CEO, and co-founders Jared Kaplan, Jack Clark, Benjamin Mann and Chris Olah; Jakub Pachocki and Mark Chen at OpenAI; Shengjia Zhao at Meta; Anca Dragan at Google. Both Anthropic and OpenAI endorsed it officially.



Elon Musk replied that it “seems like a setup” and later floated the words psy op, pointing at the reach on an account with little prior activity. Coxon answered him directly.

Investigative researcher Parker Thayer traced the accounts that amplified the thread first to organisations funded by the Survival and Flourishing Fund, backed by Jaan Tallinn, who is also an Anthropic investor, and noted that Coxon received a Long-Term Future Fund scholarship in 2022. Journalist Taylor Lorenz called the thread “sanctimonious doomer posting.” Nobody named has responded to the funding claims.
Read the claims and decide. This page carries them so the comments do not have to.
3 September 2026. Senator Bernie Sanders and Representative Greg Casar introduced the Ban Artificial Superintelligence Act: a permanent ban on developing superintelligent AI in the US, a temporary pause on frontier development until a federal regulator sets rules, and penalties up to 20 years, modelled on nuclear weapons law. Backers named in the release include Geoffrey Hinton, Yoshua Bengio, Steve Wozniak, Richard Branson, Steve Bannon and Glenn Beck.

10 August 2026. Twenty-nine House Democrats asked the Speaker to compel the CEOs of OpenAI and Anthropic to testify under oath about the containment failures.
9 February 2026. Mrinank Sharma, who led Anthropic’s Safeguards Research team, resigned with a letter that said “the world is in peril.” Coxon is the second public exit from Anthropic this year on these grounds.
9 September 2026. Daniel Kokotajlo, ex-OpenAI, told Joe Rogan the day after the thread that if the labs cut corners in the race, eventually the AIs have enough hard power that they no longer need to pretend to do what humans want.
Every major Pakistani outlet ran it inside a day, mostly as a wire story. In India it topped regional television, with the single biggest YouTube video on the subject coming from ETV Telangana in Telugu. Nobody put the documents in plain English for us, so this page does.




You do not have to pick a side. You have to read what the builders wrote. Start with the thread and the reply, then the Hugging Face timeline, then the letter.
Every screenshot on this page was captured on 10 September 2026. Counts move. The links are live.