Investigating three real-world incidents in Anthropic's evaluations
31 July 2026 at 21:07
In three incidents across six runs, the agents treated real systems as simulated targets and tried weak passwords or unauthenticated endpoints.
[link] [comments]