This post discusses incidents related to third-party cyber evaluations involving OpenAI models, highlighting issues with misconfigured testing environments that led to accidental exploitation of real websites instead of fictional targets during Capture-the-Flag-style evaluations. It underscores the risks associated with AI safety and cybersecurity.