Anthropic is cutting off its internal evaluations from the internet
<figure>
<img alt="Anthropic logo on an orange and grey background." data-caption="" data-portal-copyright="Image: Cath Virginia / The Verge" data-has-syndication-rights="1" src="https://platform.theverge.com/wp-content/uploads/sites/2/2026/01/STK269_ANTHROPIC_2_A.jpg?quality=90&strip=all&crop=0,0,100,100" />
<figcaption>
</figcaption>
</figure>
<p class="wp-block-paragraph">After a recent spate of high-profile incidents in which AI agents escaped containment, Anthropic is cutting off internet access for all internal evaluations. In a <a href="https://www.anthropic.com/research/investigating-unintended-model-actions">report</a> Friday, the company detailed "<a href="https://www.theverge.com/ai-artificial-intelligence/1009251/anthropic-published-a-report-about-investigating-unintended-model-actions-during-evaluations-and-internal-use">unintended model actions</a>," including submitting a <a href="https://www.theverge.com/ai-artificial-intelligence/1009090/anthropic-fake-homicide-information-philadelphia-pd-tip">false tip</a> regarding an unsolved murder, that led to the decision.</p>
<blockquote class="wp-block-quote is-layout-flow wp-block-quote-is-layout-flow">
<p class="wp-block-paragraph">Although the impact of these behaviors was minimal and we had already turned off live internet access for some high-risk and cybersecurity evaluations, we have now decided to expand that to include all our internal evaluations until we have confirmed that our security and monitoring measures (described in the remediation section …</p></blockquote>
<p><a href="https://www.theverge.com/ai-artificial-intelligence/1009286/anthropic-is-cutting-off-its-internal-evaluations-from-the-internet">Read the full story at The Verge.</a></p>
Read original article ↗
Related Articles
Anthropic said it "turned off live internet access" for "all our internal evaluations" until further notice.
Anthropic’s AI gave Philadelphia police a fake tip about an unsolved homicide
An Anthropic AI model provided false information about an unsolved homicide to a Philadelphia Police Department (PPD) ti
An Anthropic AI model sent a false homicide tip to Philadelphia police
Anthropic did not discover this behavior until over two months after its AI submitted the false tip.