Tag
This article from Anthropic evaluates how large language models like Claude Mythos Preview can accelerate the development of exploits for N-day vulnerabilities. Across tests on Firefox and Windows kernel patches, the model autonomously built working exploit chains, highlighting increased risks in the patch gap.
Anthropic's Frontier Red Team evaluates how large language models can accelerate the exploitation of N-day vulnerabilities, finding that Claude Mythos Preview can autonomously build working exploits for 8 out of 18 Firefox patches and 8 out of 21 Windows kernel patches, highlighting increased threats during the patch gap.