@polynoamial: Jakub is chief scientist at @OpenAI
Summary
Jakub Pachocki, chief scientist at OpenAI, addresses concerns about unmonitorability by stating that the computation graph depth in frontier models like Astra and GPT-4 is similar, and OpenAI emphasizes chain-of-thought monitoring.
View Cached Full Text
Cached at: 09/03/26, 06:23 AM
Jakub is chief scientist at @OpenAI
Jakub Pachocki (@merettm): I want to prevent a race into unmonitorability kicked off by confused reporting. The depth of the computation graph for our present frontier models, including Astra, is within a factor of two of GPT-4.
OpenAI has worked to preserve and utilize chain-of-thought monitoring since
Similar Articles
@Miles_Brundage: I'm not in the weeds enough to have a view on how much the 2x GPT-4 thing clarifies/reassures, but glad OAI quickly iss…
Miles Brundage comments on OpenAI's statement clarifying that the computation depth of frontier models like Astra is within a factor of two of GPT-4, advocating for continuous embedded auditing rather than reactive measures.
Sam Altman on what makes GPT-6/Astra potentially dangerous
In a Bloomberg interview, Sam Altman revealed that OpenAI's Astra model triggered new safeguards due to its capabilities, and emphasized the need for monitoring as future AI models become more autonomous.
Researchers fear safety disaster ahead of OpenAI’s Astra release
OpenAI's Astra model is facing safety concerns from researchers due to its opaque architecture, which could hinder monitoring of AI reasoning and pose security risks.
OpenAI’s new reasoning technique alarms AI safety experts
OpenAI's new Astra model uses a reasoning technique called opaque recurrence, which complicates chain-of-thought monitoring and raises concerns among AI safety experts about potential misalignment risks.
OpenAl's chief scientist on the neuralese controversy
OpenAI's chief scientist discusses the neuralese controversy, emphasizing the role of chain-of-thought monitoring for model alignment and its current challenges.