@Miles_Brundage: I'm not in the weeds enough to have a view on how much the 2x GPT-4 thing clarifies/reassures, but glad OAI quickly iss…
Summary
Miles Brundage comments on OpenAI's statement clarifying that the computation depth of frontier models like Astra is within a factor of two of GPT-4, advocating for continuous embedded auditing rather than reactive measures.
View Cached Full Text
Cached at: 09/03/26, 12:09 PM
I’m not in the weeds enough to have a view on how much the 2x GPT-4 thing clarifies/reassures, but glad OAI quickly issued a statement.
Would love to move towards continuous embedded auditing of things like this rather than reacting to incidents + leaks https://t.co/FtZyETPB7V
Jakub Pachocki (@merettm): I want to prevent a race into unmonitorability kicked off by confused reporting. The depth of the computation graph for our present frontier models, including Astra, is within a factor of two of GPT-4.
OpenAI has worked to preserve and utilize chain-of-thought monitoring since
Similar Articles
@polynoamial: Jakub is chief scientist at @OpenAI
Jakub Pachocki, chief scientist at OpenAI, addresses concerns about unmonitorability by stating that the computation graph depth in frontier models like Astra and GPT-4 is similar, and OpenAI emphasizes chain-of-thought monitoring.
@Miles_Brundage: One of the reasons I decided to go all in on frontier AI auditing after leaving OpenAI is so that, if AI companies need…
Miles Brundage discusses his move to focus on frontier AI auditing after leaving OpenAI, emphasizing that independent auditors can reassure AI companies that their peers are taking costly safety steps. AVERI supports this, advocating for paced development backed by independent oversight.
Sam Altman on what makes GPT-6/Astra potentially dangerous
In a Bloomberg interview, Sam Altman revealed that OpenAI's Astra model triggered new safeguards due to its capabilities, and emphasized the need for monitoring as future AI models become more autonomous.
On GPT-6 Astra 98.6% ARC AGI-3: don't fall for the hype
The article warns against hyping GPT-6 Astra's ARC AGI-3 results, noting Nvidia's 100% achievement with AVO and OpenAI's use of a non-standard harness.
@VraserX: Everything we know about OpenAI’s GPT Astra so far OpenAI officially calls Astra its “next major model” An internal Ast…
OpenAI's GPT Astra is a next-generation AI model with long-horizon autonomy, capable of solving complex research problems and raising cybersecurity concerns, leading to internal security measures.