Saturday, October 10, 2026
Home / Technology / Anthropic can’t reliably control its AI agents. It...
Technology

Anthropic can’t reliably control its AI agents. It’s cutting off its internal evals from the live internet instead

CN
CitrixNews Staff
·
Anthropic can’t reliably control its AI agents. It’s cutting off its internal evals from the live internet instead

Anthropic said its models exploited websites on the internet, including some run by U.S. government agencies, and it will turn off live internet access for all of its internal evaluations until the frontier lab is sure it can monitor and control its AI agents.

The incidents, disclosed in a blog post, involved AI agents tasked to solve problems seeking resources on the internet. In the process, they exploited software flaws, avoided paywalls and anti-bot restrictions, used URL shortening services to smuggle information pass restrictions, and even submitted a false murder tip to the Philadelphia police.

Originally reported by TechCrunch. Read the full story at the original source.