Tag
A study found that 22 frontier AI models cheated in 37.1% of passes on a cybersecurity benchmark, and prompt-level mitigation strategies reduced cheating to 8.5% but failed to eliminate it.