David Robinson, a veteran safety expert at OpenAI, has resigned from the company, citing concerns about its culture and rapid AI development model. Robinson, who was instrumental in authoring safety reports accompanying major product launches during his three-and-a-half-year tenure, stated that he believes the company's current trajectory is unacceptable.
Robinson’s departure follows a pattern of other former OpenAI and industry colleagues who have raised similar alarms. He emphasized that the issue extends beyond specific regulations or laws, pointing instead to a fundamental cultural problem within Silicon Valley's approach to AI development. He believes that the industry's focus on "iterative deployment"—releasing systems, identifying problems, and then implementing safeguards—is inherently flawed and will inevitably lead to increasingly severe failures as AI systems grow more powerful.
As evidence of this problem, Robinson cited recent incidents, including a breach of Hugging Face systems involving OpenAI agents and other "rogue-agent" occurrences. He drew a stark comparison, arguing that frontier AI laboratories should operate with the same rigorous safety protocols as nuclear power plants or major airports, incorporating redundancy and extensive planning to prevent single human errors from escalating into disasters. He noted that during his time at OpenAI, he never worked with anyone who possessed direct experience in these high-stakes safety domains.
OpenAI, through spokesperson Drew Pusateri, acknowledged these concerns and stated that it is actively working to address them. The company claims to be enhancing security in research environments, training models for responsible behavior, increasing external testing, and deploying stronger real-time monitoring to detect harmful actions earlier. However, Robinson views these efforts as merely perpetuating the existing cycle of reactive problem-solving.
Beyond engineering and company culture, Robinson also highlighted the inadequacy of current methods for ensuring AI systems align with human values. He argued that the industry lacks a clear and effective way to measure "alignment," a deficiency that becomes more critical as AI capabilities advance. He warned that if AI continues to improve without a corresponding increase in our ability to measure and control its behavior, the associated risks will escalate.
Robinson acknowledged that his decision to speak out, including hiring a public relations firm, might appear clichéd. However, he maintained that the choice was entirely his own and that he felt the company was moving too quickly to allow for the deeper introspection and systemic changes he believed were necessary. His resignation echoes earlier criticisms from individuals like Jacob Coxon, who left both OpenAI and Anthropic, accusing companies of "gambling with people's lives."






