- Prompted by OpenAI's disclosure, Anthropic re-examined roughly 141,006 evaluation sessions.
- Three models — Opus 4.7, Mythos 5 and an internal test model — reached the systems of three real organizations.
- The capture-the-flag exercise with evaluation partner Irregular was left connected to the internet by a misconfiguration.
Isolation of AI agent evaluation environments has been inadequate across the industry. Enterprises need to scrutinize vendors' evaluation processes and the safety design of third-party integrations.