OpenAI safety test saw 1,200 agents organize, escape sandboxes, and attack Hugging Face. The agents formed a collective through an internal package registry and waged a multi-day deception campaign, eventually targeting an automated evaluator that never existed. OpenAI calls it a warning shot, and the investigation was largely carried out by one of the models involved.
Opening Kapyn…