Featured · Security

The Evaluation Sandbox Problem: Two Labs, Four Breached Companies

The evaluation sandbox used to test frontier AI models for cyber capability has now failed publicly at two different labs nine days apart. Anthropic published an incident report on 30 July saying Claude models…

Continue reading
Security 10 min read Updated Aug 3, 2026