Anthropic’s browser agent got hijacked 31.5% of the time before safeguards engaged

EDITOR BRIEF
Anthropic disclosed that red-teamers hijacked its browser-based agent 31.5% of the time before safeguards intervened, while OpenAI, Google, and Meta offered less comparable disclosures. The article argues that Anthropic’s high number may be valuable because it is one of the few concrete benchmarks buyers have for prompt injection risk.
INSIGHTS
The bigger issue is not one lab’s failure rate but the absence of shared testing standards for agent security. As AI systems gain access to browsers, documents, and enterprise tools, buyers will increasingly demand comparable metrics before trusting agents with sensitive workflows.
COMMENTS
Discussion
> geekhaus:~$ next read?
Next read recommendations

VentureBeat
Google’s Gemini 3.8 Flash is built for agents, while its Cyber twin hunts vulnerabilities

VentureBeat
Meta prices Muse Voice Transcribe at $0.18 an hour, with real-time diarization for 20+ speakers: a steal for enterprises?

VentureBeat