Visual Style Jailbreak: How Stylistic Triggers Can Bypass AI Safety
Researchers find that multimodal LLMs robustly understand content regardless of visual style, yet their safety mechanisms can be easily bypassed by specific stylistic triggers. The new ASO approach automates adversarial image creation by optimizing s...