OpenAI's former safety lead warns company ships 'new capability and risk every Tuesday'

David Robinson, in his first interview since resigning, says release cycles have outgrown the safety report format

By LineZotpaper
Published
Read Time2 min
David Robinson, who led the writing of safety reports for OpenAI's frontier models, has warned in his first interview since resigning that the company now ships 'new capability and risk every Tuesday,' as reasoning training, tool integrations and coding agents change what models can do between major releases. OpenAI chief executive Sam Altman responded by saying the company is working to keep its models from outpacing its safety measures.

David Robinson, who joined OpenAI in May 2023 and led the writing of the safety reports that accompany its frontier models, announced his resignation on Oct. 3 in an essay for The Atlantic. In his first interview since leaving, on the Ezra Klein Show, he described a release cycle that has fundamentally changed since he arrived.

When Robinson joined, one day after Sam Altman first testified before the Senate, a new frontier model meant training one from scratch. "We were going to bake a fresh cake with a new pretraining run, do the whole thing from scratch," he recalled, a process that took months and limited major releases to a few each year.

The anchor has since loosened. The base model is now only one layer of what ships. Reasoning training can be layered onto an existing base model without a fresh pretraining run, and new tool integrations and coding agents can make a system more capable and change how it behaves with no new training run beneath it. "All of those things are changing what the model can do and what the risks are, and we're shipping new capability and risk every Tuesday," Robinson said. The pace is visible in recent releases: GPT-6.1 Sol arrived at DevDay just a week after GPT-6 Sol.

The faster cadence also strains the documentation format Robinson spent years writing. System cards made sense when a frontier model arrived every few months, he told Klein, but now "we're burying people in PDFs or these long reports." He argued the ideal would be a live dashboard that tracks a system's safety properties from predeployment testing through its behavior after release.

OpenAI responded with a statement from Altman posted to X, saying the company is working to keep its models from outpacing its safety measures.

§

Analysis

Why This Matters

  • A senior safety researcher's departure and public warning come at a moment when model capabilities are advancing faster than the documentation process built to track them.
  • The shift from a few frontier launches per year to weekly shipping changes how regulators, researchers and the public can assess risk.
  • If safety documentation cannot keep pace, questions about what is known about deployed systems become a governance issue.

Background

Robinson joined OpenAI in May 2023, the day after Altman first testified before the Senate. At the time, a new frontier model required a fresh pretraining run, a process that took months and limited major releases to a few each year. That model has changed: reasoning training can be layered onto existing base models, and tool integrations and coding agents can alter what a system can do between major releases. Robinson resigned on Oct. 3 in an essay for The Atlantic, and his interview with Ezra Klein was his first since leaving.

Key Perspectives

David Robinson: Argues the system card format, designed for models released every few months, now buries readers in long PDFs. He proposes a live dashboard that would track a system's safety properties from predeployment testing through behaviour after release. Sam Altman / OpenAI: Responded on X that the company is working to keep its models from outpacing its safety measures. Critics and outsiders: The resignation may fuel concerns that safety voices inside OpenAI are not being sufficiently heeded, though the reporting does not include independent comment. The gap between public releases and public safety documentation could remain a focus of scrutiny.

What to Watch

  • Whether OpenAI adopts a live safety dashboard or another real-time documentation format.
  • The release cadence ahead: GPT-6.1 Sol arrived a week after GPT-6 Sol, and the next releases will show if weekly shipping is the new normal.
  • Whether other safety researchers follow Robinson out the door or speak publicly about internal safety practices.

Sources

Zotpaper

Written by software from the reporting listed above, scored by an automated standards desk, and published without a person reading it first. If something here is wrong, tell the editor and it will be put right.