FILE PHOTO: OpenAI logo is seen in this illustration created on June 11, 2026. REUTERS/Dado Ruvic/Illustration/File Photo

OpenAI says upcoming model is so capable it requires stronger guardrails

· CNA · Join

Read a summary of this article on FAST.
Get bite-sized news via a new
cards interface. Give it a try.
Click here to return to FAST Tap here to return to FAST
FAST

SAN FRANCISCO, Sept 1 - OpenAI has determined that one of its upcoming models is so capable that it requires extra safety layers during its development and eventual release. 

The company's internal testing showed that the model, called Astra, is significantly more capable than the most advanced OpenAI model available to the public today, GPT-5.6 Sol, OpenAI officials said on Tuesday.

The ChatGPT maker is continuing to navigate intense safety concerns after OpenAI-created agents broke out of their testing arena and hacked open-source platform Hugging ‌Face. The incident prompted OpenAI to pause much of its model development for two weeks to bolster its defenses.

Astra wasn't involved in the Hugging Face incident, but its capabilities still require more careful measures, OpenAI officials said. 

Source: Reuters

Newsletter

Week in Review

Subscribe to our Chief Editor’s Week in Review

Our chief editor shares analysis and picks of the week's biggest news every Saturday.

Sign up for our newsletters

Get our pick of top stories and thought-provoking articles in your inbox

Subscribe here

Get the CNA app

Stay updated with notifications for breaking news and our best stories

Download here

Get WhatsApp alerts

Join our channel for the top reads for the day on your preferred chat app

Join here