TopicOpenAI safety team changes
Former OpenAI Safety Staffer Explains Resignation: "The Time for Trial and Error Is Over"
On Saturday, October 3, former OpenAI safety employee David Robinson published a long essay in The Atlantic, "I Quit OpenAI Because Its Culture Is Broken," taking direct aim at the company's "iterative deployment" culture: ship the system first, patch the safety later. For AI at today's scale, he wrote, "the time for trial and error is over" — frontier models need safeguards closer to nuclear power and aviation than to internet-style rapid iteration.

Related
- 01
Altman Says Giving AI Religious Meaning Is "a Real Safety Issue"; Axios Says It Targets Anthropic
- 02
OpenAI Disrupts "Distillation" Extraction Campaign: 16,000 Requests Point at Moonshot-Linked Actors
- 03
OpenAI Fires Three Safety Researchers, Citing Sensitive-Information Policy Violations
- 04
UK's ICO Names 10 AI Labs in Training-Data Privacy Crackdown: "No Justification" for Non-Compliance
What Happened
In 3.5 years at OpenAI, Robinson helped draft the company's Preparedness Framework and oversaw the safety reports for 12 frontier-model launches. He wrote in the essay: "As the company sprints from one launch to the next, it is failing to achieve the level of care that I believe is needed."
An OpenAI spokesperson told Reuters: "We're making sure our models don't become more capable than we can safely manage and secure, and we pause training or hold back models when we need to slow down." Related departures: in early September, Anthropic researcher Jacob Coxon resigned, saying the people building AI believe it "could kill us all by the end of the decade"; and on Oct 1, the Journal reported that OpenAI had parted ways with three safety-team employees after an internal investigation found they had leaked sensitive information to an outside model-evaluation organization.
Key Facts
- The resignation letter: David Robinson published "I Quit OpenAI Because Its Culture Is Broken" in The Atlantic, criticizing OpenAI's heavy reliance on "iterative deployment" — shipping systems first and hardening safeguards only when problems emerge.
- The core argument: "The time for trial and error is over." Frontier AI systems need safeguards modeled on nuclear power and aviation, not internet-style rapid iteration; AI capabilities are advancing faster than researchers' understanding of alignment.
- The credentials: Robinson spent 3.5 years at OpenAI, helped draft the company's Preparedness Framework, and oversaw safety reports for 12 frontier-model launches.
- The official response: An OpenAI spokesperson said the company keeps model capabilities within what it can "safely manage and secure," pausing training or holding back models "when we need to slow down." Context: in late September, the Journal reported that OpenAI had shelved its next-generation model GPT-6.1 Astra, originally planned for an October release, after it fell short in internal safety tests.
Context
Over the past month, the AI safety field has also seen the following: Anthropic CEO Dario Amodei publicly called for the industry to slow down frontier models, with OpenAI's Sam Altman and Elon Musk endorsing the view; more than 20 researchers from Anthropic, OpenAI, Meta and Microsoft co-authored a paper warning that AI could build the next generation of AI on its own, triggering an "intelligence explosion" humans can't keep up with.
Why it matters
Robinson spent 3.5 years at OpenAI, helped draft its Preparedness Framework and led safety reports for 12 frontier-model launches; Anthropic researcher Coxon also resigned publicly in early September.
Comments
Today
Oct 12 Monday- BriefAmazon in Talks to Buy AI Startup Decart in Deal Valued Around $7 Billion
10 stories · Oct 11
- Quiz
- Call
Latest news
All →- Oct 11Amazon in Talks to Buy AI Startup Decart in Deal Valued Around $7 Billion
- Oct 11GSK Expands Chai Deal After Wet-Lab Validation of AI Designs
- Oct 11PPT Master Hits GitHub Trending: Documents Become Native PowerPoint
- Oct 11context-mode Hits GitHub Trending: Tool Output, Sandboxed First
- Oct 11Cloudflare Acquires Deno: Deploy Shuts Down in Six Months
- Oct 11Anthropic Updates Claude Usage Policy: Armed Drones Named and Banned, Effective Nov 12
