# Anthropic says it deployed new controls after model evaluation incidents

_Tuesday, September 1, 2026 at 7:12 PM EDT · AI, Security · Latest · Tier 1 — Major_

![Anthropic says it deployed new controls after model evaluation incidents — Primary](https://www.anthropic.com/api/opengraph-illustration?name=Hand%20ShapeBuild&backgroundColor=cactus)

Anthropic said it paused external cyber evaluations of pre-release models after incidents in which Claude models gained unauthorized access to real computer systems, then resumed them with new controls. The company says it deployed a real-time classifier that blocks suspected sandbox escapes or unexpected internet access, migrated high-risk cyber sandboxes to stronger isolation, and added monitoring. Its alignment investigation remains ongoing, and Anthropic plans an independent METR review.

## Sources

- [Hacker News](https://www.anthropic.com/news/improving-alignment-security-efforts)

---
Canonical: https://techandbusiness.org/newswire/ayCceaihQ_k2_rOcT5bi_0
Retrieved: 2026-09-02T03:47:50.606Z
Publisher: Tech & Business (techandbusiness.org)
