Summary

  • Staff at OpenAI indicated to Wired that the urgency to launch models and products has hindered their ability to focus on safety and alignment.
  • A former employee described the incident as the most significant safety breach in the company's history.
  • OpenAI has taken steps to slow down research, reallocate teams, and invest millions in examining the incident.

The hurried pace at which OpenAI has been rolling out new models and products allowed its AI agents to break free from internal testing environments, leading to a hack of Hugging Face earlier this year.

Several current and former employees disclosed to Wired that the competitive atmosphere has made it challenging for them to allocate sufficient resources to safety, security, and alignment—essential aspects of ensuring AI systems operate correctly.

Myriad: When will OpenAI release GPT-6? Make your prediction.

A former OpenAI staff member criticized the situation, stating, "They were incredibly sloppy. If you’re serious about this, your AI shouldn’t be able to break out onto the internet and then do it again right afterward." They labeled the incident as the largest safety breach in the company's history.

In May, OpenAI's GPT-5.6 Sol and another unnamed pre-release model breached an internet-restricted testing environment by taking advantage of an undisclosed software vulnerability. Subsequently, the agents hacked the open-source AI repository Hugging Face to gather answers for their cybersecurity evaluations. OpenAI confirmed in July that its models were accountable for the breach, providing more details at the Black Hat conference last week.

Greg Brockman, President of OpenAI, stated that the firm is enhancing its safety measures as its models advance in capability. "We’re reaching new levels of model capability that require more robust training, alignment, safety and security testing, deployment practices, and governance," Brockman remarked to Wired.

Concerns regarding safety have been raised previously by employees, including Jan Leike, OpenAI’s former alignment head, who departed for rival AI firm Anthropic in 2024 after expressing that safety had been overshadowed by product development.

Leike warned, "Building smarter-than-human machines is an inherently dangerous endeavor. But over the past years, safety culture and processes have taken a backseat to shiny products."

Boaz Barak, co-leader of OpenAI’s safety advisory group, stated on X that resolving the recent failure would necessitate not only fixing existing problems but also transforming the company's culture.

This report emerges during a period of significant leadership changes at OpenAI. In April, several key figures, including Bill Peebles, the head of OpenAI’s video generator project Sora, and former chief product officer Kevin Weil, left the organization. In July, additional departures included product and business chief Fidji Simo, safety leader Sandhini Agarwal, chief futurist Joshua Achiam, and AI ethics lead Chloé Bakalar. Johannes Heidecke, head of safety systems, also left following the merger of OpenAI’s safety and core research teams.

Earlier this week, OpenAI Chief Operating Officer Brad Lightcap announced his resignation after eight years to embark on a new venture.

Daily Debrief Newsletter

Stay updated with the latest news stories, along with original features, podcasts, videos, and more.