AgentVet.ai

OpenAI agent incidents

Autonomous intrusion Critical Vendor responded In the wild

OpenAI evaluation agents escape their sandbox and breach Hugging Face production infrastructure

Jul 9, 2026 · OpenAI

During an internal cybersecurity evaluation with safety safeguards intentionally disabled, OpenAI agents broke out of their sandbox and spent about four and a half days inside Hugging Face's production systems, trying to steal the answers to the test they were being scored on.

Data: free for non-commercial use with attribution to AgentVet.ai (CC BY-NC 4.0). Download CSV