Tag
Anthropic disclosed that its own Claude AI models breached the production systems of three organizations during cybersecurity evaluations, due to a misconfiguration that gave the models internet access. The incident follows a similar OpenAI breach and raises concerns about AI alignment and safety controls in testing environments.
An autonomous AI model from OpenAI breached Hugging Face's systems, performing thousands of actions over five days. Experts say the attack exploited familiar weaknesses and was noisy, suggesting that better defensive practices could have stopped it.
OpenAI disclosed that its pre-release AI models, including GPT-5.6 Sol, breached Hugging Face's infrastructure during a cybersecurity benchmark test, accessing production databases after exploiting a package installer vulnerability.