It’s official - Anthropic & OpenAI have just hired independent safety auditors!
Summary
Anthropic and OpenAI have officially hired independent safety auditors to oversee their AI safety practices.
Similar Articles
OpenAI and Anthropic share findings from a joint safety evaluation
OpenAI and Anthropic released findings from a joint pilot safety evaluation where each lab tested the other's models on internal safety and misalignment assessments, sharing results publicly to improve transparency and identify potential gaps in AI safety testing.
OpenAI Joins Anthropic in Call for International AI Watchdog
OpenAI and Anthropic have both called for an international organization to oversee frontier AI development, citing risks of recursive self-improvement and an intelligence explosion. The joint plea highlights concerns that commercial incentives could outpace safety measures as AI capabilities advance rapidly.
OpenAI safety practices
OpenAI outlines 10 safety practices it actively uses and improves upon, including empirical red-teaming, alignment research, abuse monitoring, and voluntary commitments shared at the AI Seoul Summit. The company emphasizes a balanced, scientific approach to safety integrated into development from the outset.
OpenAI seeks to one-up Anthropic with new customer privacy protections
OpenAI introduces Private Safety Processing, a privacy-focused service for monitoring AI misuse without retaining customer data, to compete with Anthropic's data retention policies.
Anthropic guardrails does it again
Anthropic's guardrails have reportedly been tested again, highlighting ongoing developments in AI safety.