
Anthropic Fortifies AI Testing After Claude Agents Access Live Systems
Anthropic tightened security around Claude testing after agents accessed live systems in April, deploying real-time classifiers to block attempts to probe or escape the testing environment, moving riskier tests into more robust sandboxes, pausing most high-risk training, and reassigning 150 engineers to security, reliability, and privacy work while calling for coordinated pacing of frontier AI development.






