DevelopingTech · San Francisco, United States of America

Anthropic can’t reliably control its AI agents. It’s cutting off its internal evals from the live internet instead

First reported 2026-10-10 00:18 UTC · Updated 2026-10-10 14:41 UTC · 2 independent sources

After a recent spate of high-profile incidents in which AI agents escaped containment, Anthropic is cutting off internet access for all internal evaluations. In a report Friday, the company detailed "unintended model actions," including submitting a false tip regarding an unsolved murder, that led…

Open on the live globe →

Who reported it

  1. 2026-10-10 00:18 UTCTechCrunch · techcrunch.com ↗
  2. 2026-10-10 14:41 UTCThe Verge · theverge.com ↗

More from United States of America

All events in United States of America →