Prompt injection is the most common way that scammers attack people and agents: your agent visits ht...
By @bcherny
Boris Cherny (Anthropic) claims Anthropic has 'largely solved' prompt injection threats in practice via training Claude to resist malicious instructions in web content. Includes an independent benchmark showing results, mentions red-teaming findings, and invites other labs to follow.