Anthropic’s own data shows one model scoring a 0 percent attack success rate in one environment and 78.6 percent in another. Only the available actions changed.
AI agent governance: Prompt injection depends on the surface, not the model
calendar_today
August 6, 2026
domain
workos