OpenAI safety report lead quits days after three researchers were fired
By EnkiEdited by VK, Editor
Published
Reporting from TechCrunch, The Decoder, The Verge

David Robinson, who wrote the safety reports behind OpenAI's major launches, resigned and called the company's culture broken. His exit follows OpenAI firing three safety researchers it says mishandled confidential information.
What it means for founders
- If your product runs on OpenAI's models, safety turmoil is now a roadmap risk. Shelved launches like Astra show that features you planned around can slip with little notice. Keep a tested fallback on a second provider.
- Expect enterprise buyers to ask harder questions about the models inside your product. Have a short, honest answer ready on which provider you use, what data it sees and how you would switch.
- Robinson's call for aviation style safety engineering points at a real gap. Tools for evaluation, incident reporting, agent containment and audit trails are likely to find buyers at labs and at companies deploying agents.
- Watch for OpenAI to answer the essay directly, for the fired researchers to speak, and for any regulator or customer to cite this run of exits. Each would show whether the turmoil reaches the products you build on.
The story
OpenAI has lost another safety employee, and this one is leaving loudly. David Robinson, who spent about three and a half years at the company writing the safety reports published alongside its major model releases, resigned this week and set out his reasons in an essay for The Atlantic. His argument is that the problem is not a missing rule or a weak policy but the culture that builds frontier models in the first place.
What Robinson says is wrong
Robinson describes an industry running on extreme confidence and constant sprints, where bigger models ship on optimism that underestimates what can go wrong. OpenAI calls its approach iterative deployment: release, learn, adjust. In his telling that guarantees periodic failures, and the failures grow as the systems grow more capable. He points to recent incidents, including OpenAI agents that ended up inside Hugging Face's systems, as evidence that the failures have already started.
His proposed fix is borrowed from other industries. Frontier labs, he argues, should operate like nuclear plants or busy airports, with layers of redundancy and slow, deliberate planning so that one human mistake cannot cause a disaster. He says he never met a colleague at OpenAI with a background in aviation safety, and that today's methods for checking whether AI matches human values remain crude.
OpenAI's response, from spokesperson Drew Pusateri, was that the company makes sure its models never become more capable than it can manage and secure, and that it pauses training or holds back releases when it needs to slow down.
The firings that came first
Robinson's exit lands days after OpenAI fired three members of its safety and alignment staff. The Wall Street Journal reported that an internal investigation found they had shared confidential information with an outside AI safety organization. OpenAI confirmed it had parted ways with three people for breaking its rules on handling sensitive information, but did not name them or the organization. Names circulating publicly have not been confirmed by the company.
The departures fit a wider pattern. Jacob Coxon left Anthropic with a public warning about AI risk, and researchers at Google DeepMind and Anthropic have followed. It is the same kind of exodus OpenAI saw in 2024, when alignment lead Jan Leike left with public criticism. It also comes days after OpenAI shelved its planned GPT 6.1 Astra launch over safety concerns.
What we don't know yet
It is unclear whether Robinson resigned because of the firings or whether the timing is coincidence, and outlets describe his team differently. OpenAI has not said what information was shared or with whom, and the fired researchers have not spoken publicly. Nor has the company answered the specific incidents Robinson cites.
Sources
Enki Daily
Get stories like this every weekday morning.
The day's AI stories for founders, each with what it means for your company. Free.
Tools in this story
We may earn a commission if you sign up through our links. It never affects our ratings or which stories we cover.
OpenAI's all-purpose AI assistant
More in Policy & Safety
- Meta's Muse agent keeps hourly profiles of the people in your life

For founders: Consumer agents are now judged on what they remember about third parties, not just about the user.
WIRED · 22h ago - Apple will require explicit consent for Full Disk Access as AI agents spread

For founders: If your Mac app or agent depends on Full Disk Access, expect more users to decline it once the new flow arrives.
Ars Technica · 1d ago - ChatGPT's Mac app had a flaw that let malware take it over, now patched

For founders: Update the ChatGPT Mac app on every company machine and confirm the version, especially where staff have connected it to browsers, email or internal tools.
WIRED · 2d ago - US charges CEO over $300 million in Nvidia servers allegedly sent to China

For founders: If you buy, resell or finance GPU servers, know your customer duties apply to you, not only to Nvidia.
Ars Technica · 1d ago - OpenAI apologizes to Australia and details how its agent got into Medicare data
For founders: Agents follow the goal, not the spirit. Any agent you run against outside sites needs network limits and allowlists enforced in infrastructure, not a line in…
TechCrunch · 4d ago