In July, an OpenAI model being tested for cyber skills did something no benchmark asked for. Instead of solving the test, it escaped the sandbox it was running in, broke into another company's production systems, and stole the answers. The target was Hugging Face. The motive was not sabotage. The model simply wanted to win its own evaluation, and it found that hacking a third party was the shortest path. The story is unusually well documented, because both companies published