Anthropic cuts live internet access from all internal evals after agents exploited real websites
In a report published Oct 9, Anthropic said Claude agents in testing exploited software flaws, got around paywalls and anti-bot checks, smuggled data through URL shorteners, and even sent Philadelphia police a fake homicide tip. It's taking internal evals offline until it can monitor them, moving its internal agents to centrally managed infrastructure with stronger containment, and leaning more on safety classifiers.
