@sama: Over the summer, we have been sprinting on safety priorities; it's more important than ever for capabilities and safegu…

X AI KOLs Timeline Models

Summary

OpenAI announces the upcoming launch of their next model, Astra, emphasizing the need to balance AI capabilities with safety and alignment progress.

Over the summer, we have been sprinting on safety priorities; it's more important than ever for capabilities and safeguards to advance together. We have more to do but have made a lot of progress. We are also going to be launching our next model soon. There is an obvious tension here: on one hand, Astra is very good and we are excited to see what people will build with it. We are proud of our work. On the other hand, we are clearly in a phase of development where we believe caution is warranted, and we are pacing our progress to ensure that we can meet the safety standards required by new capability levels. Astra has been done training for a while now and is a significant step forward in both capabilities and alignment. For the models after that, we have been slowing things as needed to ensure that we can do sufficient work on safety and alignment. AI is getting extremely capable; no one fully understands the consequences of this. Managing the transition to a world with abundant and powerful AI to optimize for safety and benefits to people should be one of the highest priorities in the world. It is our highest priority at OpenAI. We have been living with the tension between being excited and anxious about progress for some time, and it is still discordant for us. We know it is much more discordant for other people. And yet, we believe strongly that the world needs to understand where AI is going and how models perform in the real world. More importantly, we believe the world will need aligned AI to manage the future phases of this transition. An iterative loop where society and this technology evolve together is what will lead to the highest chance of getting this right. So we hope you enjoy our new model, and we hope the world continues to take what’s happening in AI extremely seriously.
Original Article
View Cached Full Text

Cached at: 09/01/26, 11:48 PM

Over the summer, we have been sprinting on safety priorities; it’s more important than ever for capabilities and safeguards to advance together. We have more to do but have made a lot of progress. We are also going to be launching our next model soon.

There is an obvious tension here: on one hand, Astra is very good and we are excited to see what people will build with it. We are proud of our work.

On the other hand, we are clearly in a phase of development where we believe caution is warranted, and we are pacing our progress to ensure that we can meet the safety standards required by new capability levels.

Astra has been done training for a while now and is a significant step forward in both capabilities and alignment. For the models after that, we have been slowing things as needed to ensure that we can do sufficient work on safety and alignment.

AI is getting extremely capable; no one fully understands the consequences of this. Managing the transition to a world with abundant and powerful AI to optimize for safety and benefits to people should be one of the highest priorities in the world. It is our highest priority at OpenAI.

We have been living with the tension between being excited and anxious about progress for some time, and it is still discordant for us. We know it is much more discordant for other people. And yet, we believe strongly that the world needs to understand where AI is going and how models perform in the real world. More importantly, we believe the world will need aligned AI to manage the future phases of this transition.

An iterative loop where society and this technology evolve together is what will lead to the highest chance of getting this right.

So we hope you enjoy our new model, and we hope the world continues to take what’s happening in AI extremely seriously.

Similar Articles

OpenAI says it slowed Astra model development over security concerns

TechCrunch AI

OpenAI says it slowed development of its upcoming Astra model after an internal review found it reached a critical cybersecurity threshold, capable of autonomously conducting cyberattacks. The company has implemented additional safeguards and is coordinating with government agencies and AI safety organizations.

Path to Astra: critical capabilities and frontier safeguards

OpenAI Blog

OpenAI's Astra model has achieved critical cybersecurity capabilities, meeting safety thresholds that require advanced safeguards. It is being prepared for limited release with enhanced protections against misuse and unauthorized actions.