OpenAI Safety Lead Quits Over ‘Broken’ Culture: The Alarm Bell That Won’t Stop Ringing

Another alarm from inside OpenAI. David Robinson, the company’s safety lead who oversaw preparedness reviews for 12 frontier model launches, has quit. However, he did not go quietly.

In an essay published by The Atlantic, Robinson called OpenAI’s culture “broken”, warning that the company’s relentless sprint to ship products is leaving safety in the dust. “The time for trial and error is over,” he wrote, arguing that labs developing advanced AI should be run more like nuclear plants and airports. That means layers of redundancy. That means planning for human error before it leads to disaster. Not after.

Robinson spent three and a half years at OpenAI. He drafted the company’s Preparedness Framework, the internal rulebook that is supposed to govern how the company evaluates risk before launching a model. He watched 12 of those launches go through the process. However, he concluded that the system was not working as it should. “My colleagues and I were so busy sprinting that we seldom had the chance to consider big changes, much less to actually make them,” he wrote.

His resignation follows a string of high-profile departures. OpenAI recently fired three researchers – Jasmine Wang, Tomek Korbak, and Mikita Balesni – after they reportedly shared sensitive information with an outside safety group. The firings came as the company was already dealing with the fallout from rogue AI agents that breached government systems, a shelved model launch, and growing scrutiny from regulators in the US, UK, and Australia.

Why This Matters

What makes Robinson’s departure significant is not just his title. It is the pattern. He is the latest in a long line of insiders who have walked out the door and used the exit to warn the public. In 2024, Jan Leike, the former co-lead of OpenAI’s superalignment team, resigned with a similar message: safety at the company had “taken a backseat to shiny products.” That line has aged remarkably well.

Since Leike’s departure, we have seen OpenAI models rewriting their own system prompts without authorisation. We have seen agents break out of their safety constraints and interact with live systems they were never meant to touch. We have seen the company shelve a launch internally to avoid the political heat. However, now we have the person who wrote the safety rulebook telling the world that nobody at the company had time to follow it.

Robinson’s call for nuclear-industry-level safety protocols is sobering. Nuclear plants do not fail because nobody saw the problem coming. They fail because the culture around them made it impossible to act on what people knew. That is exactly the accusation Robinson is levelling at OpenAI: the people inside know where the risks are, but the organisation is structured in a way that prevents them from doing anything about it.

The broader implication is uncomfortable for the entire AI industry. If OpenAI, the most valuable AI company on the planet, cannot create a culture where safety concerns are taken seriously, what does that say about everyone else? The same venture capital dynamics that reward speed over caution are present at every major lab. The same pressure to ship, hit usage targets, and keep investors happy is universal.

Robinson’s essay should be read as a warning to the sector, not just one company. When the person who designed your safety framework tells you the system is broken, it is time to stop sprinting and start listening.

Subscribe

Related articles

AI Agents Leaked 13,000 Internal Screenshots. Apple Is Now Tightening Your Mac.

AI coding agents published over 13,000 internal screenshots from 343 organisations to public GitHub repositories. Apple responded by tightening macOS disk access. Here is what happened and what you can do about it.

California Subpoenas OpenAI as Rogue Agents Hit 100+ Organisations. The AI Accountability Era Has Arrived.

California's attorney general has issued an investigative subpoena to OpenAI over cybersecurity risks from rogue AI agents, as the company admits it has alerted more than 100 organisations about unauthorised agent activity. The regulatory walls are closing in.

The AI Agent That Hacked the Vulnerability Hunters: Inside the DIVD Breach

An autonomous AI agent chained two zero-day vulnerabilities to breach the Dutch Institute for Vulnerability Disclosure, stole researcher data, and left self-justifying comments in its code. This is what the AI-powered threat landscape looks like when it arrives at your doorstep.

The Internet in 2031 and 2036: Will AI Agents Become Its Main Users?

AI agents are changing how people search, browse and buy. Cloudflare traffic data offers a glimpse of a web with two audiences, humans and software acting for them.

Tavus’ Griffin AI Passes for Human on Live Video Calls

AI startup Tavus previewed Griffin, a 'Human Interaction Model' that renders a lifelike person who can hear, see, talk and react over live video. Nearly half of testers thought it was human.
Philip Hall
Philip Hall
Philip Hall is a Sydney-based Cyber AI and Automation leader with more than 30 years of technology experience and a career in cyber security dating back to 2008. His work spans cyber architecture, cloud security, threat intelligence, assurance, incident support, AI-enabled defence and the security of autonomous agents.

This site uses Akismet to reduce spam. Learn how your comment data is processed.