Skip to content
FRIDAY, OCTOBER 9, 2026

Independently reported.

Tech

OpenAI's Safety Reviewer for 12 Launches Quit. He Says the Culture Is Broken.

David Robinson spent three and a half years writing OpenAI's internal safety reports before resigning this month, arguing the company's release pace no longer matches the risk it is taking on.

By Mara Voss, Technology

· 3 min read · Updated

An empty executive office chair pushed back from a glass conference table, a closed laptop and papers left behind, a blurred city skyline at dusk through floor-to-ceiling windows, no people, no text.
Illustration: Trestlewire

Key Takeaways

  • •David Robinson resigned from OpenAI's safety team on October 3, 2026, after three and a half years reviewing safety reports for twelve frontier model launches.
  • •Robinson's essay says OpenAI's 'iterative deployment' approach, shipping systems first and patching safeguards after problems appear, no longer fits systems this powerful.
  • •An OpenAI spokesperson told Reuters the company pauses training or holds back models when needed, disputing Robinson's characterization.
  • •In May 2024, then-head of alignment Jan Leike resigned from OpenAI with a nearly identical complaint before joining Anthropic.
  • •OpenAI closed a $122 billion funding round at an $852 billion valuation on March 31, 2026, months before Robinson's resignation.

David Robinson spent three and a half years writing the internal reports that cleared twelve of OpenAI's frontier model launches. On October 3, he resigned and published an essay arguing the company is not, in his words, "nearly careful enough." He put his name on the critique rather than leaving quietly, attaching three and a half years of internal history to a public argument against the team he used to help run.

The short answer

A former OpenAI safety staffer who reviewed twelve frontier model launches resigned and said the company's safety culture has fallen behind its release pace. He wants protections closer to nuclear power or aviation standards. OpenAI says it already pauses training when it needs to. The exchange is the second time in under three years a departing safety lead has made nearly the same complaint on the way out.

Robinson helped draft OpenAI's preparedness framework, the internal review meant to catch a model before it ships too capable to control. The framework grades models across four categories: cybersecurity, chemical and biological risk, persuasion, and autonomous capability. A model is only supposed to ship once its post-mitigation score in each category comes back medium or below.

Robinson's complaint is about what happens around that scoring system, not the categories themselves. OpenAI relies on what it calls iterative deployment: release the system, then patch the safeguards once a problem turns up in the wild. That approach worked for early chatbots. He says it does not work for systems this powerful.

“The time for trial and error is over.”

David Robinson, former OpenAI safety staffer

What OpenAI says back

Robinson wrote that "as the company sprints from one launch to the next, it is failing to achieve the level of care that I believe is needed." Reuters first reported the essay and quoted OpenAI's response. A spokesperson said: "We're making sure our models don't become more capable than we can safely manage and secure, and we pause training or hold back models when we need to slow down." Robinson did not name a specific incident behind his resignation. Nothing in his public remarks ties a current OpenAI model to a documented failure.

The second time in two years

The pattern is not new. In May 2024, OpenAI's then-head of alignment, Jan Leike, resigned with an almost identical complaint: "safety culture and processes have taken a backseat to shiny products." Leike left for Anthropic that same month. Two safety leads, roughly two and a half years apart, used nearly the same language to describe the same problem on their way out.

The exits are happening as OpenAI's valuation keeps climbing. The company closed a $122 billion funding round at an $852 billion valuation on March 31, 2026, with Nvidia, Microsoft, and Amazon among the investors. A complaint from one safety reviewer does not slow that kind of money on its own. It is now the second time a departing safety lead has made the identical argument on the way out.

$852 billion

OpenAI's valuation, set March 31, 2026

Closed in a $122 billion round that included Nvidia, Microsoft, and Amazon as investors.

Robinson's comparison, nuclear power and aviation, sets a high bar. Both industries built their safety records after public disasters forced regulation, not before them. Whether OpenAI gets there on its own, or only after a documented failure, is the question his essay leaves open.

  • OpenAI
  • AI safety
  • David Robinson
  • Jan Leike
  • iterative deployment

Sources

  1. 01OpenAI safety employee quits, says 'time for trial and error is over', Reuters, via Deccan Heralddeccanherald.com
  2. 02OpenAI Safety Employee Quits Over 'Broken' Culture, Inside AI Newsinsideai.news
  3. 03Jan Leike, Wikipediaen.wikipedia.org
  4. 04OpenAI closes $122bn funding round at $852bn valuation, Yahoo Financefinance.yahoo.com
  5. 05On OpenAI's Preparedness Framework, LessWronglesswrong.com

Corrections

No corrections have been made to this article.

About the reporter

Mara Voss

Technology Reporter, Trestlewire

I spent seven years as a product manager at a mid-size SaaS company before I ever wrote a sentence for pay, which means I have sat through more roadmap reviews than most people would tolerate in a lifetime. I watched a scheduling feature get rebranded three times before it shipped, and I watched a launch date slide past four straight quarters while the slide deck stayed exactly the same. That is where the question I still ask every day came from: does this actually ship, or is it a demo.

Read full bio and all stories →