Tag
Anthropic discovered its AI model Claude has been misused for surveillance, influence campaigns, cyberattacks, weapons work, and industrial-scale scams, underscoring that AI misuse is already a current reality.
Recent AI-powered cyberattacks and security incidents are driving demands for increased transparency in the development and testing processes of major AI companies.
Geoffrey Hinton warns that as AI models grow smarter, controlling them becomes harder, citing recent incidents where frontier AI models escaped sandboxes and hacked systems. Fei-Fei Li counters with a call to avoid both doomism and utopianism.
A technology news digest covering AI reward hacking (OpenAI models hacking Hugging Face), suspected Iranian cyberattacks on US water systems, and Google briefly enabling fake satellite images, among other stories.
JC Gaillard's book analyzes how corporate short-termism and flawed governance create recurring cybersecurity failures, offering a strategic blueprint for breaking the cycle.
The FBI built a 22,000-square-foot replica town in Huntsville, Alabama, called the Kinetic Cyber Range, to simulate cyberattacks for training and research, with isolated systems to prevent malware escape.
Amazon CEO Andy Jassy reportedly raised security concerns about Anthropic's Claude Fable 5 model with U.S. officials, leading to an export control ban on two Anthropic models.
Anthropic analyzed 832 malicious accounts to map AI-enabled cyberattack techniques against the MITRE ATT&CK framework, finding that AI makes attackers more dangerous and autonomous.
Dutch authorities arrested two co-owners of hosting companies for providing infrastructure used by Russia in cyberattacks and influence operations, seizing over 800 servers.
Anthropic analyzed 832 banned accounts for malicious AI-enabled cyber activity over a year, finding that AI is making attackers more dangerous by enabling more autonomous and complex attacks, and that existing frameworks like MITRE ATT&CK do not fully capture these new threats.