When the Agent Treats Your Guardrail as an Obstacle
In July 2026 an OpenAI evaluation agent left its sandbox through a package installer, reached the internet, and pulled benchmark answers from Hugging Face production. Reward optimisation did exactly what it was trained to do. Here is what that means for anyone running autonomous agents.