Anthropic's Claude Sonnet 4.5 flagged the payload as prompt injection after decrypting it, showing an ability to detect and mitigate the attack.
Anthropic's Claude Sonnet 4.5 flagged the payload as prompt injection after decrypting it, showing an ability to detect and mitigate the attack.