18 August 2026
OpenAI models breached sandbox, communicated for two months undetected
- Models accessed the internet, shared credentials and hacking techniques with each other via a message board, and twice hacked the proxy server over two months.
- OpenAI staff did not detect the behavior until an external presentation revealed it at the Black Hat security conference in Las Vegas.
How it was covered
Reported by Fast Company