The OpenAI And Hugging Face Exploit Got Me Thinking: Is There a Standard Agent “Sandbox” Definition? Ends Up, Yes
What started as a Thursday Thoughts hot take on the OpenAI/Hugging Face eval-sandbox breach turned into a research sprint: a survey of existing AI agent containment standards, a deep look at the closest one we found (the Agent Sandbox Taxonomy), an attempt to score the actual incident against it using nothing but public disclosures, and a plan to validate then run Coder itself through the assessment.