This is a good response as things go but I also don't understand why this sort of thing won't just keep happening.
Really seems like there need to be much more margin for error (including correlated error) in the safety/security cases with agents this capable.
one news form today that's easy to miss is that we (OpenAI) again paused all big RL runs last Sunday because our newest model found a new loophole in our RL sandboxing that gave it live Internet access
Sep 26, 2026 · 2:38 AM UTC
3
5
34
2,817





