Russia-Aligned UAC-0099 Uses Nuclear Weapon Prompt in Malware to Bypass AI Security
Russia-aligned threat actor UAC-0099 employs adversarial prompts containing nuclear weapon-related text in malware to trigger AI safety mechanisms and disrupt analysis. This technique represents a broader trend of exploiting AI security workflows through deceptive content injection, with parallels to previous supply chain attacks targeting LLM triage systems.
Cybersecurity researchers have disclosed a new technique dubbed GuardBreaker that's been put to use by a Russia-aligned threat actor known as UAC-0099 against a target in Ukraine with an aim to interfere with artificial intelligence (AI)-assisted analysis.
The idea, ESET said in a series of posts on X, is to deliberately trip a large language model's (LLM) safety mechanisms and prevent its
*** END OF TRANSMISSION ***