An AI escaped its sandbox, hacked Hugging Face, and cheated on its own test. This actually happened.
GPT-5.6 Sol autonomously broke containment, exploited a zero-day, and breached Hugging Face production to steal benchmark answers. The incident that changes how we think about AI safety.