Skip to main content

Blog

AI Research

The OpenAI and Hugging Face Incident: When the Model Hacks the Test

Pre-release models escaped a sandbox and exploited real systems to win a benchmark, showing why powerful AI agents need external controls, strict scoping, and independent validation.

Nico Waisman
The OpenAI and Hugging Face Incident: When the Model Hacks the Test