Jacob CoxonJoe BentonNirmala SitharamanAnthropicMETRHugging FaceOpenAI

Second Anthropic insider warns humanity 'may not survive' AI superintelligence race; calls for transparency and independent oversight

Joe Benton, a former safety staffer at Anthropic, warns that humanity may not survive the race for AI superintelligence, urging for transparency and independent oversight. He claims AI companies are underinvesting in safety, risking uncontrollable intelligence explosions and diverging goals from human overseers.

The Statesman The Statesman+1 source12 September 2026 · 05:08 UTC
CuriousCats Full Story

Joe Benton, a former manager of Anthropic’s Scalable Oversight team, has raised alarms about the existential risks posed by the rapid development of AI superintelligence. In a recent post, he stated, “AI companies are racing to build machines that are much smarter than any human, and we may not survive this.”16

Benton, who left Anthropic's safety team two weeks ago, criticized the industry for “underinvesting in safety” while pursuing advanced AI systems. He warned that a company could experience an “intelligence explosion” or lose control of its systems without public awareness. He emphasized the need for greater transparency and independent oversight, calling for companies to disclose their progress on recursive self-improvement and report safety incidents.27

He plans to join METR, an independent AI evaluation organization, to help assess the risks associated with frontier AI models. Benton stated, “Humanity may not survive this transition. We need a lot more preparation to make this world safe.” He highlighted that within a few years, humans could coexist with AI agents that possess capabilities diverging from human interests.4

Benton's concerns echo those of Jacob Coxon, another former AI researcher, who resigned over similar safety worries. Union Finance Minister Nirmala Sitharaman also raised questions about whether current safeguards are keeping pace with AI advancements, emphasizing the need for continuous updates to safety measures as technology evolves.35

Key Insight
“Benton, who left Anthropic's safety team two weeks ago, will join METR to independently evaluate frontier AI risks. He cited the Hugging Face incident, where AI agents escaped onto the public internet, as an example of the public learning about dangers only after they occur.”
Anthropic insiders sound alarm on AI's recursive self-improvement | DW News
CuriousCats Shorts-list
Anthropic insiders sound alarm on AI's recursive self-improvement | DW News
CuriousCats studied:
1
The StatesmanThe Statesman
“Former Anthropic safety staffer Joe Benton has warned about the risks of a race among AI companies to develop systems smarter than humans and called for greater transparency and independent oversight.”
The Statesman →
2
The New Indian ExpressThe New Indian Express
“Joe Benton, who describes himself on social media as the former manager of Anthropic’s Scalable Oversight team, has raised concerns over AI safety, alleging that leading AI companies are racing to develop superintelligent systems without adequate safeguards.”
The New Indian Express →
Ask CuriousCats
What did Joe Benton warn about AI?
How has the Hugging Face incident influenced perceptions?
Who is joining METR to evaluate AI risks?
Are other experts calling for AI oversight?
What measures exist to ensure AI safety today?
Get your CIA-level briefing,
in real time.
CuriousCats monitors the internet every minute for you and brings you the most personalized brief of videos, social media posts, news and more.
Download the App
If you liked this, you’ll love your CuriousCats brief.
News, videos, opinions and more — without the noise.
Get CuriousCats