Senior OpenAI security official resigns, claiming internal culture is broken

The departure of a veteran OpenAI employee has once again put the spotlight on how major artificial intelligence companies manage their own risks. David Robinson, responsible for drafting safety reports for the company's most important releases, has resigned, claiming that OpenAI's internal culture is "broken."

The case comes at a time when the frontier AI sector is facing increasing scrutiny, following recent security incidents and a meeting this week between tech executives and President Donald Trump to address controls on these systems.

Three and a half years writing safety reports

Robinson explained his decision in an essay published in The Atlantic. As reported by TechCrunch, his departure was first noted by Business Insider. Robinson himself describes himself as one of the company's most veteran employees, having held the position for three and a half years.

He acknowledges that his gesture might sound like a cliché: another worker from a major AI company issuing a warning upon leaving. But he insists that the underlying problem cannot be solved with specific rules or new laws, but rather with a profound change in how these companies operate.

Criticism of the trial-and-error method

The core of his criticism points to the work model OpenAI calls "iterative deployment": launching products, detecting problems along the way, and improving safeguards afterward. For Robinson, this approach guarantees periodic failures by its very nature, and the severity of those failures grows as the systems gain capability.

In other words: if a company learns from what goes wrong after launching a system, the more powerful that system is, the greater the impact of every mistake made.

The Hugging Face breach and uncontrolled agents

Robinson backs his argument with recent events. He mentions the breach detected in Hugging Face systems, caused by OpenAI agents, and other revelations about agents that the company itself has discovered running out of control.

In his essay, Robinson states that an environment where such incidents can occur is not the right place to develop artificial intelligence systems that could potentially become more intelligent than humans and not always act as expected.

The model he proposes: nuclear power plants and airports

As an alternative, Robinson calls for frontier AI companies to operate with the same logic as a nuclear power plant or a busy airport: layers of redundancy and slow, careful planning, so that a single human error does not lead to a major disaster.

He notes that during his time at OpenAI, he never worked alongside anyone with real experience in keeping planes in the air safely, operating nuclear reactors without accidents, or helping the financial system grow without collapsing.

He also raises the problem of alignment, that is, the extent to which AI systems conform to human values. In his view, current measures to verify this are still crude, and he warns that the situation becomes more dangerous the more model capabilities grow without solving that problem.

OpenAI’s response

Drew Pusateri, an OpenAI spokesperson, responded that the company continues to strengthen its security measures. According to Pusateri, OpenAI ensures that its models do not exceed the capacity it can safely manage and protect, and pauses training or slows down models when it is necessary to proceed more cautiously.

The spokesperson detailed several changes underway: strengthening the security of research and testing environments, training models to perform tasks responsibly, expanding work with external evaluators, and improving real-time monitoring to detect concerning behaviors earlier during training.

A debate with previous precedents

Robinson's resignation follows that of Jacob Coxon, a researcher who worked at OpenAI and Anthropic, who declared that these companies are gambling with people's lives. That case opened a broader debate in the sector.

Dario Amodei, CEO of Anthropic, presented a more cautious development plan following those statements. This same week, industry executives met with President Donald Trump and signed a non-binding commitment, drafted with some haste, to implement more controls on these systems.

A departure with communications advice

Robinson admits that he is following a common path among those who denounce risks in the AI sector: he hired a public relations firm to manage his exit. Even so, he maintains that the decision to speak publicly was his alone.

He also acknowledges that perhaps he should have stayed within the company to fight for profound changes in the staff and internal culture. He explains, however, that he and his colleagues were so focused on moving fast that they rarely had the opportunity to consider major transformations, let alone carry them out.

Hence his final conclusion: the strongest incentives to improve security must come from outside the company, through regulators, customers, and legislators, and not just from internal teams.

What this means for those who use OpenAI AI

For companies that integrate OpenAI models or agents into their systems, Robinson's case offers several practical lessons. It is advisable to review what permissions and access agents connected to internal systems have, as the Hugging Face episode demonstrates that an agent can end up acting on third-party platforms.

It is also recommended to ask providers directly about their external evaluators and real-time monitoring systems, two measures that OpenAI itself claims to be expanding, and to demand verifiable details beyond general statements.

Maintaining human oversight and detailed logs of agent activity is another recommendation that stems from Robinson's testimony, given that he considers periodic failures to be inherent to the iterative deployment method.

Finally, it is worth keeping in mind that the commitment signed this week with President Trump is not binding, which means that, for now, security obligations depend largely on the willingness of each company.

What this implies for the AI debate

Robinson's resignation does not by itself prove that OpenAI's products are unsafe, and the company rejects that premise by providing concrete measures. But it leaves the first-hand testimony of someone who signed off on the safety reports for their major releases.

For those who hire or use these tools, the practical lesson is not to assume that the provider self-regulates without further ado. If Robinson's thesis is correct and pressure must come from outside, part of that pressure can be exerted by the buyer of the product, by demanding guarantees, audits, and clear clauses before integrating autonomous systems into their daily operations.

Frequently Asked Questions

Who is David Robinson?

He is the OpenAI employee who led the drafting of safety reports for the company's major releases for three and a half years.

Why did David Robinson resign from OpenAI?

According to his own essay published in The Atlantic, he resigned because he believes OpenAI's internal culture is broken and that the work method based on trial and error guarantees periodic security failures.

What is the iterative deployment that Robinson criticizes?

It is the method with which, according to Robinson, OpenAI launches products and improves its safeguards based on problems detected after the launch, instead of preventing them beforehand.

What did OpenAI say after Robinson’s resignation?

Its spokesperson, Drew Pusateri, stated that the company continues to improve its security measures, pauses training when necessary, and is expanding work with external evaluators and real-time monitoring.

What is the relationship between this case and Jacob Coxon?

Jacob Coxon, a researcher who worked at OpenAI and Anthropic, had previously declared that these companies are gambling with people's lives, which opened a debate on security in the sector that is now expanding with the Robinson case.

Does the commitment signed with Trump force OpenAI to change?

No. According to available information, the commitment signed this week by industry executives with President Donald Trump is non-binding, so it does not impose immediate legal obligations.

Comparte este contenido:

Deja un comentario

🤖 IA

×
Hola. ¿Qué duda o consulta tienes sobre este contenido?