@NatPurser: over the past week, I’ve gotten a lot of questions about what independent AI evaluations should actually look like in p…
Summary
The tweet discusses the need for minimum conditions for independent AI evaluations to ensure credibility, highlighting principles endorsed by over 100 experts to standardize safeguards across the industry.
View Cached Full Text
Cached at: 09/19/26, 06:50 AM
over the past week, I’ve gotten a lot of questions about what independent AI evaluations should actually look like in practice.
what about conflicts of interest between labs and evaluators? will evaluators actually get meaningful access? what if companies just block the publication of unfavorable findings?
these are good and important questions. and we shouldn’t assume voluntary arrangements alone will produce the conditions that credible independent evaluation requires.
so i’m very excited to see 100+ researchers, evaluators, and experts — including us at @AVERIorg — laying out minimum conditions for embedded evaluation, including:
meaningful independence, strong access, transparency / editorial control, protections against retaliation, and a diversity of perspectives and competencies.
and we shouldn’t try to incorporate those principles in ad hoc, one-off agreements, policy is critical to standardizing these safeguards across the industry.
glad to see @aievalforum leading on putting these principles on paper.
AI Evaluator Forum (@aievalforum): Today, more than 100 leading AI experts endorsed a set of minimum requirements to take seriously AI companies’ recent call to embed external evaluators.
These evaluators need to be genuinely independent, transparent, and represent a range of expertise areas. They also need to
Similar Articles
@NatPurser: a couple thoughts after chats w/ some non-ai-safety DC friends: i think there’s a misperception that independent evalua…
The author discusses a misperception that independent AI safety evaluators aim to replace government oversight, emphasizing their preference for clear regulatory rules and their role as a complement to government efforts.
@RepLoriTrahan: .@Fathom_org is right. Bringing in independent evaluators is a real step. But a voluntary commitment can be dropped the…
The article advocates for independent evaluators in AI safety via the FRONTIER Act to ensure transparency and accountability, arguing that voluntary commitments are insufficient compared to legal requirements.
Anthropic and OpenAI want to embed safety evaluators. Will they really be independent?
Anthropic and OpenAI propose embedding independent safety evaluators within their companies to assess AI models during training, but details need to be clarified to ensure true independence and effective oversight.
Closing the AI Trust Gap: The Case for Independent Certification for Trustworthy AI
This paper argues that current responsible AI practices fail to create a market that rewards trustworthiness, proposing independent, outcome-oriented certification to close the 'trust gap' by making AI trustworthiness measurable, comparable, and commercially rewarded.
@andykonwinski: 3 take-aways from chatting w/ top AI researchers last month: - Evals are the “source code” of AI agents (48:35) - BigAI…
Summary of three key takeaways from conversations with leading AI researchers at CAISconf, covering the importance of evaluations for AI agents, the trade-offs between industry and academia, and a novel pedagogical RL approach.