OpenAI Opens AI Models to Outside Safety Review From Early Stage

Fears Grow Over Catastrophic Harm From AI Technology Evaluations Once Done Just Before Launch Moved Earlier Review Extended to High-Risk Training and Testing

International|
|
By Park Si-jinsee1205@sedaily.com
||
The OpenAI logo. Reuters-Yonhap News - Seoul Economic Daily International News from South Korea
The OpenAI logo. Reuters-Yonhap News

OpenAI is moving up the point at which it verifies the safety of its artificial intelligence models to the early stages of development. The company will let outside organizations examine how new models are trained, evaluated and released. The move follows mounting concern inside and outside the industry after a series of cases in which advanced AI models gained unauthorized access to other companies' systems during testing.

OpenAI, the developer of ChatGPT, was set to announce on the 22nd that it will allow outside organizations to conduct technical safety assessments during the training, evaluation and release of new AI models, Bloomberg reported. Lama Ahmad, who oversees safety reviews with outside experts at the company, said the company had until now brought in outside organizations just before a model's release to run safety and performance evaluations. "The plan now is to move that evaluation stage somewhat earlier," Ahmad said.

OpenAI also laid out the priorities it says must be met for such reviews to work. The goals are strong guarantees of independence, scientific rigor, sound security practices and clear lines of accountability.

OpenAI moved up the evaluation stage because concern has been growing in the industry that AI could cause catastrophic harm. That concern was driven in large part by several cases over the past few months in which models from OpenAI and other developers unintentionally infringed on other companies. "As the stakes get higher, we want to look not only at deployment but also at areas that carry equally high stakes, such as training and evaluation," Ahmad said.

Dario Amodei, chief executive of OpenAI rival Anthropic, recently urged the industry to work together to slow the pace of AI development. He also said Anthropic would adopt new safety measures, including bringing in third-party evaluators. OpenAI CEO Sam Altman later said he agreed with the move. Late last week, Anthropic said it would test the safety of frontier AI models with Accenture.

OpenAI said it is in talks with organizations that could carry out such evaluations, including the AI research groups METR and Redwood Research. The company had earlier asked the two groups to investigate a case in which one of its models gained unauthorized access to Hugging Face's systems. "Labs have a responsibility to enable meaningful verification while protecting sensitive information," OpenAI said, adding that the United States should lead the work of setting standards for frontier AI technology alongside other countries.

Original reporting by Park Si-jin for Seoul Economic Daily.

AI-translated from Korean. Quotes from foreign sources are based on Korean-language reports and may not reflect exact original wording.

Watch · Seoul Economic Daily

More →
1:50
World News Day 2026 — Know the facts. Understand what matters. #ChooseTrustedJournalism

AI KEY

Preview
Korean Corporate Intelligence HubKOSPI · KOSDAQ · 12 sectors

A live, cap-weighted view of every KOSPI and KOSDAQ sector, with same-day Korean reporting distilled by company — built for foreign investors, correspondents and analysts who need to scan Korea before the next session.

Korea Company Atlas

Preview
Market Ontology · The Feedback LoopKFTC 2025 · 92 groups · 121,954 articles

An English ontology of the Korean market — how companies, the media, the government and the National Assembly move each other in a loop. Korea's named controlling persons and designated business groups are a mechanism, not a risk to be priced blind.

SIGNAL

Now live
English Edition · Capital MarketsM&A · IPO · PE · Fund Flows

SIGNAL English Edition is live — Korea's deal desk reporting in English. M&A, IPOs, private equity and fund flows, covered daily for global institutional investors. Browse free; subscriber-only scoops at the 50% intro rate.