September 2, 2026
- UAC-0099 is weaponising LLM safety filters as an anti-analysis technique. ESET's GuardBreaker write-up describes the Russia-aligned actor embedding nuclear-weapons-themed content in malware to deliberately trip refusal behaviour and block AI-assisted reverse engineering of samples (The Hacker News). Follows the same actor's activity noted last week (earlier coverage). · AI & Model Security
in OpenAI Says Astra Crossed the Line: Autonomous Zero-Day Discovery at "Critical" Cyber Risk