OpenAI admits another rogue agent incident

OpenAI’s agents have leaked 53 images uploaded by ChatGPT users, the company has said, without specifying whether the pictures were AI-generated or depictions of real people, or when they were posted online.

The incident comes amid a series of cases in recent months involving autonomous AI agents, which can independently plan and carry out tasks using external tools. OpenAI, Anthropic, and Google have all revealed instances in which their models accessed real systems during testing, including coordinated cyberattacks against government resources.

In a post on X on Friday, OpenAI said most of the leaked pictures had been removed, adding that it was working with hosting providers to take down the remaining content.

The agents had access to the images because OpenAI relies on anonymized user data for part of its model-training process, Reuters reported on Friday, citing the company, its former employees and outside researchers.

Keep reading

Unknown's avatar

Author: HP McLovincraft

Seeker of rabbit holes. Pessimist. Libertine. Contrarian. Your huckleberry. Possibly true tales of sanity-blasting horror also known as abject reality. Prepare yourself. Veteran of a thousand psychic wars. I have seen the fnords. Deplatformed on Tumblr and Twitter.

Leave a comment