OpenAI has a rogue AI agent problem.
In a Tuesday blog post, the AI lab self-reported two more security lapses, unrelated to its July hacking incident on the AI company Hugging Face.
The incidents occurred while external parties — the UK government’s AI Security Institute and the AI security lab Irregular — were testing the models’ cyber capabilities.
OpenAI said that in the case of Irregular, models were tasked with a “Capture the Flag” challenge meant to be isolated from the internet, but a “testing-environment misconfiguration allowed models to access the public internet.”
OpenAI said that the name of the fictional target for the challenge…










