
Anthropic's Opus 5 Nearly Immune to Prompt Injection Attacks
Anthropic claims its Opus 5 model is nearly immune to prompt injection attacks, with a zero percent success rate in browser agent tests across 129 scenarios. The model also leads the Gray Swan IPI benchmark, reducing attacker success to 2.0 percent after 15 attempts. The protection relies on a combination of model improvements and Auto Mode defenses.

