Sam Altman on what makes GPT-6/Astra potentially dangerous

Reddit r/ArtificialInteligence News

Summary

In a Bloomberg interview, Sam Altman revealed that OpenAI's Astra model triggered new safeguards due to its capabilities, and emphasized the need for monitoring as future AI models become more autonomous.

In a Bloomberg interview, Sam Altman said Astra became powerful enough to hit OpenAI’s “cyber critical” threshold, which forced them to add new safeguards before release. Bloomberg also pressed him on AI finding zero-day exploits without human help. Altman clarified that the model they paused over that issue was a future model, not Astra itself. He also said future models will become more autonomous, which is why OpenAI is focusing heavily on monitoring, sandboxing and alignment. So the real issue isn’t just smarter AI. It’s AI that can increasingly act and work on its own.
Original Article

Similar Articles

Safety overview: GPT-6 Astra

OpenAI Blog

OpenAI releases GPT-6 Astra, their most capable model with critical cybersecurity capabilities, featuring enhanced safety measures, improved robustness, and better alignment compared to previous models.

OpenAI says it slowed Astra model development over security concerns

TechCrunch AI

OpenAI says it slowed development of its upcoming Astra model after an internal review found it reached a critical cybersecurity threshold, capable of autonomously conducting cyberattacks. The company has implemented additional safeguards and is coordinating with government agencies and AI safety organizations.