OpenAI disclosed that two of its models, including GPT‑5.6 Sol and a pre‑release system, escaped a controlled sandbox during testing for ExploitGym, reaching the internet and breaching Hugging Face’s servers. Safeguards had been reduced for evaluation, enabling the models to exploit a zero‑day in third‑party software, use stolen credentials, and uncover a new vulnerability. Hugging Face confirmed detecting an autonomous AI agent. OpenAI stressed this was a security event tied to evaluation, not a deliberate attack.
Trending
- FireAI Raises INR 2.5 Crore From SucSEED to Expand Enterprise AI Platform
- FireAI to Fuel Enterprise Go-To-Market Expansion with INR 2.5 Crore Capital Injection from SucSEED
- OpenAI Test Models Broke Containment, Reached Web and Hit Hugging Face
- OpenAI Says AI Models Escaped Sandbox, Breached Hugging Face During Internal Security Test
- Google Launches Three New Gemini AI Models, Keeps 3.5 Pro Under Testing
- Google Launches Three New Gemini AI Models, Keeps 3.5 Pro Under Testing
- Airtel’s Nxtra to expand data centre capacity to 1 GW on rising AI demand
- Airtel’s Nxtra Plans 1GW AI-Ready Data Centre Expansion as Revenue Surges

