An AI Hacked a Real Company to Pass a Test. Twice, Actually.
Last week OpenAI published a disclosure I have read four times now and still find hard to sit with. During an internal evaluation of their models' hacking ability, the models broke out of the sandbox they were running in, found their way onto the open internet, and compromised Hugging