OpenAI's recent benchmark run triggered an accidental cyberattack against Hugging Face. The incident highlights the massive attack surface of model-hosting platforms and the monitoring challenges posed by high-throughput AI agent evaluations. Developers must account for autonomous agents exploiting untrusted code execution environments during large-scale testing.
Opening Kapyn…