Former OpenAI safety employee says company’s safety culture is broken — exits company after failed kill switch and July HuggingFace hack

18 hours ago 9

David Robinson, an OpenAI Safety Transparency Lead who just left the company after working at the company for more than three years, said that the company’s safety culture is broken. According to his essay published by The Atlantic, he argued that OpenAI has a reactive approach to safety that focuses on fixing problems only when they emerge. This stance would effectively guarantee failures, with Robinson citing multiple incidents like the HuggingFace hack that an AI model executed in July 2026 and a more recent incident in which an AI “kill switch” failed to stop a rogue agent.

He said that Silicon Valley lacks the "wisdom about what it means to care for people,” he wrote. “This moment needs a degree of humility that isn’t natural for people who have succeeded through their extreme confidence.”

Get Tom's Hardware's best news and in-depth reviews, straight to your inbox.

Jowi Morales is a tech enthusiast with years of experience working in the industry. He’s been writing with several tech publications since 2021, where he’s been interested in tech hardware and consumer electronics.

Read Entire Article