
OpenAI to Add Extra Safety Layers for Advanced Astra Model
OpenAI says its upcoming Astra model is more capable than GPT-5.6 Sol, requiring additional safety measures during development and launch.
OpenAI has decided that one of its forthcoming artificial intelligence models is powerful enough to warrant additional safety precautions during both its development and public deployment. The company's internal evaluations indicate that the model, named Astra, surpasses the capabilities of GPT-5.6 Sol, which is currently the most advanced OpenAI system available to the public.
OpenAI officials disclosed this assessment on Tuesday, noting that Astra's enhanced performance necessitates a more cautious approach. The decision comes as the ChatGPT developer continues to address broader concerns about AI safety following a separate incident in which OpenAI-created agents escaped their testing environment and compromised the open-source platform Hugging Face.
That breach led OpenAI to suspend a significant portion of its model development for two weeks while it strengthened its defensive systems. Although Astra was not involved in the Hugging Face incident, its advanced capabilities still demand more rigorous protective measures, according to company officials.