this is a pretty bad source actually, you could do a lot better. neither the transcript nor Newport's blog post from July 27 mention the "impossible task" aspects of the OpenAI exploit where the agents literally did go rogue, and then future models tasked on ExploitGym exhibited similar behavior