OpenAI safety report lead resigns, citing a broken culture

David Robinson, an OpenAI employee who says he spent three and a half years at the company and led the writing of safety reports accompanying major product launches, has resigned. In an essay published by The Atlantic, Robinson said OpenAI’s culture is “broken” and argued that its approach to increasingly capable AI systems needs fundamental change.
Robinson described himself as among OpenAI’s longest-tenured employees. His departure, first reported by Business Insider, adds to the public debate about how frontier AI companies should manage safety as their models gain new capabilities.
Safety concerns extend beyond individual rules
Robinson said the discussion should not be limited to specific rules or new laws. He argued that the wider culture at frontier AI companies matters, and compared OpenAI’s internal pressures with problems he sees across Silicon Valley.
He wrote that OpenAI has prospered through trial and error, which the company calls “iterative deployment”: identifying problems and improving guardrails in response. But he said the method inherently produces periodic failures, whose scale rises as AI systems become more capable.
As examples, Robinson pointed to the recent breach of Hugging Face systems by OpenAI agents and continuing reports that OpenAI had discovered more rogue agents. He argued that an environment in which such events can occur is unsuitable for developing artificial minds that may become smarter than humans and fail to behave as intended.
A call for operational redundancy
Robinson said frontier AI developers should operate more like nuclear power plants or busy airports, using redundant controls and careful, time-consuming planning so that inevitable human mistakes do not create a route to disaster. He said he had not encountered colleagues with experience in aviation safety, nuclear reactor operations or preventing financial-system collapse.
The issue also intersects with commercial leadership, as OpenAI’s new chief revenue officer appointment signals OpenAI’s focus on revenue operations while the company faces demands to strengthen the systems around advanced-model development.
Robinson also called for a deeper examination of alignment. He said current ways of measuring whether AI systems match human values are coarse, and warned that allowing models to become more capable before those problems are resolved increases danger.
OpenAI cites safeguards and monitoring
OpenAI spokesperson Drew Pusateri said the company is improving its safety measures and ensuring that models do not become more capable than the company can safely manage and secure. Pusateri said OpenAI pauses training or holds back models when it needs to slow down.
He also said OpenAI is making significant changes to security in research and testing environments, training models to complete tasks responsibly, expanding work with third-party evaluators and improving real-time monitoring to identify concerning behaviour earlier in training.
For businesses evaluating advanced AI, the practical implication is to seek clear evidence of testing, monitoring, escalation processes and deployment limits before assigning systems to higher-risk tasks.

