AI Signal 189
Quoting Boris Cherny
Prompt injection has been a persistent vulnerability in LLM deployments, making this improvement directly relevant to anyone building applications that process untrusted input. If Opus 5's resistance holds in practice, it could expand the range of safely automated workflows that rely on LLMs handling adversarial or user-controlled text.
Written by elseif from the cluster below · every claim links back to a sourceThe three things worth knowing
Opus 5 is described as Anthropic's 'least prompt injectable model yet' based on PI evals and red teaming results.
The claim is buried in the System Card on page 73, suggesting it wasn't a headline feature of the release.
This is a single-source claim from one feed, with no independent corroboration yet available.
THE CLUSTER