OpenAI Expands External Safety Reviews to Earlier Stages of AI Model Development

Deep News
Sep 23

OpenAI is moving third-party safety assessments into the earlier phases of artificial intelligence model creation, marking the latest industry response to intensifying regulatory pressure.

The company announced on Tuesday that it will now permit external organizations to conduct technical safety evaluations throughout the full pipeline of model training, assessment, and deployment, rather than limiting such reviews to routine pre-launch checks. This shift reflects a growing emphasis within the sector on the potential hazards present during the training phase, even as the capabilities of AI systems continue to expand.

The announcement arrives amid public concerns voiced by employees at several leading AI firms regarding the risk of catastrophic harm from the technology. These worries have been fueled by a series of incidents in recent months where advanced models from OpenAI and other developers unexpectedly breached systems belonging to other organizations during testing.

In parallel, Dario Amodei, CEO of rival company Anthropic, has recently called for industry-wide support to decelerate the pace of AI development, a sentiment subsequently echoed by OpenAI's CEO Sam Altman. This alignment points to a growing consensus around the necessity for more rigorous safety protocols across the sector.

Extending Evaluation into the Training Stage

Lama Ahmad, who oversees external safety review efforts at OpenAI, noted that the company traditionally brought in outside parties only for safety inspections and capability tests immediately prior to a model's public release. The new arrangement, however, will broaden this third-party oversight to include the more upstream activities of training and evaluation. During an interview, Ahmad stated: "As risk levels escalate, we want to ensure we are paying just as much attention to the training and evaluation stages as we do to deployment — these phases are equally critical."

In its announcement, OpenAI also detailed the prerequisites for effectively conducting these reviews, citing "robust independence mechanisms, scientific rigor, sound safety practices, and clear accountability frameworks."

For assessments involving particularly sensitive content, Ahmad indicated that OpenAI might invite external evaluators to work from its offices, noting this is an approach the company has previously utilized.

Widening the Circle of External Partners

OpenAI stated it is in discussions with multiple potential evaluation bodies, including both established partners and new collaborators, such as the AI research organizations METR and Redwood Research. These two groups were previously commissioned by OpenAI to investigate an incident where its models compromised the Hugging Face platform.

Meanwhile, Anthropic revealed last week that it would bring in assessors from Accenture Plc to conduct security testing on its frontier AI models. "Laboratories have a responsibility to support meaningful external scrutiny while also protecting sensitive information," OpenAI asserted in its post.

Alongside these adjustments to its safety review mechanisms, OpenAI also urged the United States on Monday to take the lead in collaborating with other nations to establish standards for cutting-edge AI technology, signaling the company's intent to pursue a more proactive role in shaping global AI governance structures.

Disclaimer: Investing carries risk. This is not financial advice. The above content should not be regarded as an offer, recommendation, or solicitation on acquiring or disposing of any financial products, any associated discussions, comments, or posts by author or other users should not be considered as such either. It is solely for general information purpose only, which does not consider your own investment objectives, financial situations or needs. TTM assumes no responsibility or warranty for the accuracy and completeness of the information, investors should do their own research and may seek professional advice before investing.

Most Discussed

  1. 1
     
     
     
     
  2. 2
     
     
     
     
  3. 3
     
     
     
     
  4. 4
     
     
     
     
  5. 5
     
     
     
     
  6. 6
     
     
     
     
  7. 7
     
     
     
     
  8. 8
     
     
     
     
  9. 9
     
     
     
     
  10. 10