David Robinson, a former safety lead at OpenAI, has resigned and stated the company’s culture is broken. In an essay for The Atlantic, he noted that while he is among the longest-tenured staff with three-and-a-half years there, he was also “something of a cliché” for issuing a dire warning while leaving.
In this article
A pattern of warnings
Robinson led the writing of safety reports that accompanied major product launches. His comments echo those of Jacob Coxon, a researcher at both OpenAI and Anthropic who quit and declared these firms are “gambling with our lives.” Coxon’s remarks sparked a wider debate, leading Anthropic CEO Dario Amodei to unveil a plan for more cautious development. AI executives met with President Donald Trump this week and signed what appeared to be hastily written, non-binding pledges to implement more safety controls.
Robinson argues the debate must go beyond specific rules or new laws to address the overall culture. While much reporting has focused on how CEO Sam Altman lost the trust of former colleagues, Robinson suggests OpenAI’s issues mirror those of Silicon Valley at large.
Risk and reaction
“OpenAI has thrived by trial and error (which it calls ‘iterative deployment’), looking for problems and improving its guardrails in response,” he wrote. “But this approach, by its very nature, guarantees periodic failures — and the scale of those failures is growing as systems get more capable.”
He pointed to the recent breach of Hugging Face systems by OpenAI agents, as well as continuing revelations of OpenAI discovering more rogue agents. Robinson argued, “An environment where things like this can happen is no place to grow artificial minds that could be smarter than we are and that might not do what we want them to.”
Given the increased risk, Robinson argued that frontier AI companies need to start operating “like nuclear-power plants or busy airports, with layers of redundancy and careful, time-consuming planning, so that the occasional and inevitable human error does not open a door to disaster.”
But Robinson said that in his time at OpenAI, he “never encountered a colleague who had experience making airplanes fly safely or nuclear reactors run without melting down, or helping the financial system grow without collapsing.”
In response to Robinson’s essay, OpenAI spokesperson Drew Pusateri said the company continues to improve its safety measures. “We’re making sure our models don’t become more capable than we can safely manage and secure, and we pause training or hold back models when we need to slow down,” Pusateri said in a statement. “We’re making significant changes to strengthen security in our research and testing environments, train models to not just complete tasks but do so responsibly, expand our work with third-party evaluators, and improve real-time monitoring so we can detect and respond to concerning behavior earlier in the training process.”
Alignment and incentives
Beyond calling for changes in OpenAI’s culture, Robinson also said it’s time to ask bigger questions about alignment — something that he admitted could sound “touchy-feely,” but he said it’s critical as companies’ current “measures of how well” AI systems “match human values are coarse.”
“The smarter the industry lets models grow while these problems remain unsolved, the more dangerous our situation becomes,” he said.
Robinson’s departure was first reported by Business Insider. In his essay, he also acknowledged that he’s following an apparently a common step in the AI whistleblower playbook: He’s hired a PR firm. But he insisted, “The decision to speak out is mine alone.”
“Perhaps I should have stayed and fought for fundamental shifts in our staffing and culture, but in practice, my colleagues and I were so busy sprinting that we seldom had the chance to consider big changes, much less to actually make them,” Robinson said. “That’s why I concluded that stronger incentives for safety — coming from outside the company — are a big part of getting this right.”




