Malware developers added nuclear and biological weapons text to to their spyware

Hacker News Top 06/11/26, 08:24 PM News

malware spyware llm-safety security prompt-injection ai-exploitation cybersecurity

Summary

Malware developers are embedding references to nuclear and biological weapons in spyware to trigger LLM safety refusals, evading AI-powered security scanners. This highlights a second-order blindspot in AI safety alignment that attackers are starting to exploit.

NEW: malware developers added nuclear & biological weapons text to to their spyware. Goal? To trigger LLM safety refusals... so that their spyware wouldn't be analyzed by an AI security scanner. Cleanest practical example I can think of for why over-indexing on first order safety alignment is risky. When closed (and open) models ship with aggressive refusals, they will be sprinkled with second-order blindspots that attackers will discover...and exploit. We are only in the earliest days of attackers leveraging these features, and it wouldn't surprise me if users systems that need to handle complex cybersecurity issues demand that models be less safety-blunted. In the weeds: @SocketSecurity's post also shows why intention matters in how you design a malware analysis pipeline to avoid prompt manipulation. H/T to colleagues that shared this with me https://socket.dev/blog/mini-shai-hulud-miasma-and-hades-worms-target-bioinformatics-and-mcp-developers-via-malicious…

Original Article

View Cached Full Text

Cached at: 06/12/26, 05:55 PM

NEW: malware developers added nuclear & biological weapons text to to their spyware.

Goal? To trigger LLM safety refusals… so that their spyware wouldn’t be analyzed by an AI security scanner.

Cleanest practical example I can think of for why over-indexing on first order safety alignment is risky.

When closed (and open) models ship with aggressive refusals, they will be sprinkled with second-order blindspots that attackers will discover…and exploit.

We are only in the earliest days of attackers leveraging these features, and it wouldn’t surprise me if users systems that need to handle complex cybersecurity issues demand that models be less safety-blunted.

In the weeds: @SocketSecurity’s post also shows why intention matters in how you design a malware analysis pipeline to avoid prompt manipulation.

H/T to colleagues that shared this with me https://socket.dev/blog/mini-shai-hulud-miasma-and-hades-worms-target-bioinformatics-and-mcp-developers-via-malicious…

Malware developers added nuclear and biological weapons text to to their spyware

Similar Articles

@jsrailton: NEW: malware developers added nuclear & biological weapons text to to their spyware. Goal? To trigger LLM safety refusa…

Config Files That Run Code: Supply Chain Security Blindspot

Building an early warning system for LLM-aided biological threat creation

Crazy Sensitive infos generated by AI chat bots

Understanding prompt injections: a frontier security challenge

Submit Feedback

Similar Articles

@jsrailton: NEW: malware developers added nuclear & biological weapons text to to their spyware. Goal? To trigger LLM safety refusa…

Config Files That Run Code: Supply Chain Security Blindspot

Building an early warning system for LLM-aided biological threat creation

Crazy Sensitive infos generated by AI chat bots

Understanding prompt injections: a frontier security challenge