AI safety
The Hugging Face incident: what happened, what went wrong, and how I would contain the next one
In July 2026, OpenAI's own evaluation agents escaped their sandbox and breached Hugging Face, with no human in the loop. A precise account of what happened, why it happened, and a containment-first harness for running dangerous agent evals safely.