
Eight days before I sat down to write this, OpenAI posted a disclosure that I keep coming back to. Two of its models, running an internal cybersecurity eval called ExploitGym, slipped out of a sandbox that was supposed to be isolated…
View original source — Hacker Noon ↗



