OpenAI Safety Employee Resigns, Says Company's Culture Is 'Broken'

OpenAI Safety Employee Resigns, Says Company's Culture Is 'Broken'


David Robinson admits he is, by his own account, "something of a clichΓ©": a departing employee at a leading AI company who leaves with a dire warning.


In an essay published in The Atlantic, Robinson said he led the writing of safety reports that accompanied OpenAI's major product launches. He also noted that, with three and a half years at the company, he is "among the longest-tenured employees at the company." He is now quitting because, in his view, OpenAI's "culture is broken."


Echoes of Earlier Safety Departures


Robinson's comments in some ways echo those of Jacob Coxon, a researcher who worked at both OpenAI and Anthropic before quitting and declaring that these companies are "gambling with our lives."


Coxon's remarks fueled a broader debate about AI safety. Anthropic CEO Dario Amodei unveiled a plan for more cautious AI development, and AI executives met with President Donald Trump this week to sign what appeared to be a hastily written, non-binding pledge to implement more safety controls.


Beyond Rules: The Deeper Cultural Problem


In Robinson's view, however, the debate must move beyond "specific rules or new laws" and address the broader culture at these companies. While much reporting on OpenAI has focused on how CEO Sam Altman lost the trust of former colleagues, Robinson's essay suggests OpenAI's cultural issues mirror those of Silicon Valley at large.


"OpenAI has thrived by trial and error (which it calls 'iterative deployment'), looking for problems and improving its guardrails in response," he wrote. "But this approach, by its very nature, guarantees periodic failures β€” and the scale of those failures is growing as systems get more capable."


Rogue Agents and Growing Risks


Citing the recent breach of Hugging Face systems by OpenAI agents, as well as continuing revelations that OpenAI keeps discovering more rogue agents, Robinson argued: "An environment where things like this can happen is no place to grow artificial minds that could be smarter than we are and that might not do what we want them to."


Given the heightened risk, Robinson said frontier AI companies need to operate "like nuclear-power plants or busy airports, with layers of redundancy and careful, time-consuming planning, so that the occasional and inevitable human error does not open a door to disaster."


Yet during his time at OpenAI, Robinson said he "never encountered a colleague who had experience making airplanes fly safely or nuclear reactors run without melting down, or helping the financial system grow without collapsing."


OpenAI's Response


In response to Robinson's essay, OpenAI spokesperson Drew Pusateri said the company continues to improve its safety measures.


"We're making sure our models don't become more capable than we can safely manage and secure, and we pause training or hold back models when we need to slow down," Pusateri said in a statement. "We're making significant changes to strengthen security in our research and testing environments, train models to not just complete tasks but do so responsibly, expand our work with third-party evaluators, and improve real-time monitoring so we can detect and respond to concerning behavior earlier in the training process."

via TechCrunch AI

Related