“OpenAI safety leader quits, warning AI company’s culture is ‘broken’”
What actually happened
David Robinson resigned from OpenAI and published an essay in The Atlantic arguing that AI firms, including his former employer, are moving too fast and lack the culture needed to handle dangerous technology safely. The Guardian's report accurately summarises his central claims and situates them alongside similar warnings from a former Anthropic researcher and an ex-OpenAI/DeepMind scientist, while also including OpenAI's on-record response.
Key facts
- Robinson's own essay is titled "I quit OpenAI because its culture is broken": Robinson wrote that a cultural overhaul was needed at cutting-edge AI firms and incidents such as a "swarm" of OpenAI agents attacking Hugging Face were "typical of the industry, given the speed and flexibility with which people operate."
- His actual role at OpenAI was narrower than "safety leader" implies: he led the writing of safety reports that accompanied the ChatGPT developer's product releases.
- The article includes OpenAI's recent caution, which the headline omits but the body does not bury: OpenAI announced it was scrapping the release of a next-generation AI model after researchers raised safety concerns during internal testing, and had also paused training of its most advanced models.
- OpenAI's spokesperson response is included in full context, not cherry-picked against the company: an OpenAI spokesperson said the company was continuing to "strengthen our safety and security practices to address the risks we see today," while working on dealing with risks that might be created by future AI breakthroughs.
- The story is framed within a wider pattern rather than as an isolated claim: Robinson's essay follows the resignation of Jacob Coxon, a researcher at Anthropic who quit last month and warned AI "could kill us all by the end of the decade," followed by Anthropic warning there was a more than 10% chance AI would wipe out humanity within the next decade.
What to watch for
Watch whether OpenAI's scrapped model release and paused training get framed by other outlets as proactive caution or as reactive damage control after Robinson's essay. Also watch how "50% chance we all die" style estimates from Geoffrey Irving are reported, since the article itself flags that such probabilistic doom claims are contested as unfalsifiable.
