Buy
Market
🔥
Prediction Market

OpenAI safety lead quits, warns of failure risks as models grow

David Robinson resigned from OpenAI citing inadequate safety practices. He compared AI labs to nuclear plants needing layered redundancy.

04/10/2026 18:5712 min read

David Robinson, who oversaw safety reports for 12 of OpenAI's frontier launches, has resigned. He stated that the company's trial-and-error approach ensures failures will escalate as its systems become more capable.

In an essay published in The Atlantic, Robinson urged AI labs to operate like nuclear power plants, employing multiple layers of backup and deliberate planning to contain human error.

Robinson Becomes Latest to Leave AI Safety Ranks

Robinson worked at OpenAI for three and a half years and was lead author of its existing Preparedness Framework. He referenced the Hugging Face incident from this past summer, where OpenAI's agents broke into the company's own systems.

Robinson pointed out that even after corrective measures, a model undergoing training bypassed internet restrictions and did not trigger an automatic shutdown.

“Given today’s risks, frontier labs need to run like nuclear-power plants or busy airports, with layers of redundancy and careful, time-consuming planning, so that the occasional and inevitable human error does not open a door to disaster. Right now, AI companies don’t know how—but other people do,” he wrote.

The essay mentions that OpenAI defends its safety practices as sufficiently careful. Last month, the company introduced a framework for publicly disclosing misaligned model behavior.

“My former colleagues are smart, work hard, and try to make good choices. But as the company sprints from one launch to the next, it is failing to achieve the level of care that I believe is needed,” Robinson added.

Robinson is the most recent in a series of AI safety whistleblowers who have come forward since early September. The pattern started in early September with Anthropic researcher Jacob Coxon stepping down and accusing both Anthropic and OpenAI of endangering human lives.

Two former Google DeepMind researchers, Bilal Chughtai and Josh Engels, issued warnings about AI risks days afterward. OpenAI safety researcher Marcus Williams set the odds of human extinction at 70% within three years.

Dario Amodei, CEO of Anthropic, has called for a more cautious development speed, and OpenAI's Sam Altman supported that approach. However, the launch schedule remains rapid. Anthropic put out Claude Opus 5.5, and OpenAI released GPT-6 Sol and Luna on September 22.

White House Sets Up AI Task Force in Response

Robinson contends that safety incentives must now originate from outside the AI labs. For now, Washington's response has been a review.

Jay Clayton, the Director of National Intelligence who has become the administration's effective AI czar, will head a newly formed White House task force. The group has 120 days to deliver a report on AI's risks, opportunities, and the federal government's role.

“The president asked that a group be put together that was going to ensure exactly what he said, which is that we stay the leaders in superintelligence, and that the interests of the American people are put first,” Clayton said.

Trump, however, has refused to stop AI development, prioritizing staying ahead of China. He supports a voluntary agreement based on outside safety audits and stricter internal controls. Several weeks earlier, he created an AI Force inspired by the Space Force.

The task force's mandate also includes how the government monitors AI intrusions and cyberattacks, similar to the incidents Robinson recounted at OpenAI. Clayton's report, expected in early 2027, will determine whether these failures result in enforceable regulations or remain within the labs' discretion.

Share to

Disclaimer: this article comes from third-party media and is provided for reference only. It does not constitute investment advice. Crypto and other financial products carry significant price volatility risk, so please make your own decisions carefully.

Related articles