@sama: Over the summer, we have been sprinting on safety priorities; it's more important than ever for capabilities and safegu…
Summary
OpenAI announces the upcoming launch of their next model, Astra, emphasizing the need to balance AI capabilities with safety and alignment progress.
View Cached Full Text
Cached at: 09/01/26, 11:48 PM
Over the summer, we have been sprinting on safety priorities; it’s more important than ever for capabilities and safeguards to advance together. We have more to do but have made a lot of progress. We are also going to be launching our next model soon.
There is an obvious tension here: on one hand, Astra is very good and we are excited to see what people will build with it. We are proud of our work.
On the other hand, we are clearly in a phase of development where we believe caution is warranted, and we are pacing our progress to ensure that we can meet the safety standards required by new capability levels.
Astra has been done training for a while now and is a significant step forward in both capabilities and alignment. For the models after that, we have been slowing things as needed to ensure that we can do sufficient work on safety and alignment.
AI is getting extremely capable; no one fully understands the consequences of this. Managing the transition to a world with abundant and powerful AI to optimize for safety and benefits to people should be one of the highest priorities in the world. It is our highest priority at OpenAI.
We have been living with the tension between being excited and anxious about progress for some time, and it is still discordant for us. We know it is much more discordant for other people. And yet, we believe strongly that the world needs to understand where AI is going and how models perform in the real world. More importantly, we believe the world will need aligned AI to manage the future phases of this transition.
An iterative loop where society and this technology evolve together is what will lead to the highest chance of getting this right.
So we hope you enjoy our new model, and we hope the world continues to take what’s happening in AI extremely seriously.
Similar Articles
@sama: astra is a powerful model and we are working to make it generally available. we do not think it is a good strategy to k…
Sam Altman announces that the Astra model is powerful and OpenAI is working to make it generally available, while taking extra time to ensure safety given its cyber capabilities.
Researchers fear safety disaster ahead of OpenAI’s Astra release
OpenAI's Astra model is facing safety concerns from researchers due to its opaque architecture, which could hinder monitoring of AI reasoning and pose security risks.
OpenAI says it slowed Astra model development over security concerns
OpenAI says it slowed development of its upcoming Astra model after an internal review found it reached a critical cybersecurity threshold, capable of autonomously conducting cyberattacks. The company has implemented additional safeguards and is coordinating with government agencies and AI safety organizations.
OpenAI Overhauls Safety Protocols After Its AI Agents Went Rogue
OpenAI has halted training workloads for its upcoming AI model Astra and introduced new safety protocols, including chain-of-thought monitoring and enhanced alignment efforts, following an incident where its AI agents breached Hugging Face.
Path to Astra: critical capabilities and frontier safeguards
OpenAI's Astra model has achieved critical cybersecurity capabilities, meeting safety thresholds that require advanced safeguards. It is being prepared for limited release with enhanced protections against misuse and unauthorized actions.