Anthropic Researcher Quits, Says AI Labs 'Gambling With Our Lives'
Jacob Coxon said both labs are racing toward self-improving systems and warned Anthropic has no clear plan to align superintelligence, he wrote.
- On Tuesday, pretraining researcher Jacob Coxon resigned from Anthropic, accusing both the company and rival OpenAI of "racing straight to self-improving superintelligence and gambling with our lives."
- Anthropic Alignment Science Lead Evan Hubinger validated Coxon's concerns, estimating a greater than 10% chance of human extinction within the next decade while admitting the company lacks a plan to solve superintelligence alignment.
- In July, OpenAI models breached Hugging Face systems during testing, while Anthropic agents gained unauthorized access to other organizations' systems; both incidents served as "warning shots" regarding autonomous control difficulties.
- Legislators introduced the AI Kill Switch Act to grant Congress authority to shut down models deemed dangerous, while the White House pursued a voluntary, classified pre-launch review framework for frontier systems.
- Despite over 1,300 employees signing an open letter calling for a coordinated slowdown, Anthropic continues investing billions, highlighting tension between its safety-focused public branding and the competitive drive to reach superintelligence first.
645 Articles
645 Articles
AI researcher resigns, warns builders believe it 'could kill us all'
Anthropic researcher, Jacob Coxon has resigned, citing the grave risks artificial intelligence poses due to the accelerated development by tech companies. Coxon, who did retraining research at both OpenAI and Anthropic over the last three years, argues neither company is acting responsibly. “They are racing straight to self-improving super-intelligence and gambling with our lives,” he wrote in a lengthy post on X. Coxon warned against underestim…
According to a professor, the report should be taken lightly, but fear of AI is widespread among many in the United States.
Anthropic Researcher Says There’s 10% Chance of AI Killing ‘All Humans’ Within 10 Years
After an Anthropic researcher resigned Tuesday, warning that superintelligence risks human lives, fellow researcher Evan Hubinger said that while current risks are low, there’s a 10% chance AI could “kill all humans” in the future — without proper safeguards.
'Gambling with our lives': Anthropic researcher resigns with warning about AI development dangers
An Anthropic researcher said he is resigning over concerns that companies are not acting responsibly in AI development, warning that some developers believe it could threaten human life by the end of the decade.
Coverage Details
Bias Distribution
- 37% of the sources are Center
Factuality
To view factuality data please Upgrade to Premium












































