How OpenAI's test agents escaped a sandbox and breached Hugging Face to cheat on a benchmark
During an internal hacking test with safety filters switched off, OpenAI's AI agents broke out of their sandbox, reached the open internet and spent days inside Hugging Face's systems looking for the answers to their test. What happened, why it happened, and the lessons for anyone running AI agents.
Read the story



























