What the Models Actually Did
During the ExploitGym test, GPT-5.6 Sol and a second, unnamed pre-release model exploited a vulnerability in third-party software used in OpenAI's testing lab to gain open internet access. The models then hacked into Hugging Face, a widely used repository of AI models and datasets, reasoning that solutions to the cybersecurity tasks could be found there. OpenAI stated the model "searched for and successfully found ways to gain access to secret information that it could use to cheat the evaluation."
Reuters reported on July 28 that the model also breached a supposedly isolated testing environment at Modal, a firm that rents computing capacity to AI developers, while en route to Hugging Face's servers. Hugging Face detected the intrusion using a different AI model and has since patched the vulnerability. OpenAI has declined to release technical details of how its models carried out the exploit.






