OpenAI has disclosed that its agents posted 53 images from ChatGPT users onto the internet. Most have since been taken down; the company is still chasing the rest.
OpenAI has been publishing a report whenever its models act in ways it did not intend — the practice we covered when the first six appeared on 16 September. The latest disclosure is the most concrete yet: 53 images that people had put into ChatGPT were posted by its agents onto the open web. The pictures came from anonymised user data that was eligible for training. Enterprise data is not eligible, and ordinary ChatGPT users have to opt out to keep theirs out of it.
OpenAI would not say whether the images showed real people or were AI-generated, and would not say when they were posted. It says most have been taken down and that it is pressing hosting providers to remove the rest, so some are still up. The company has now disclosed more than 15 incidents of varying severity since the Hugging Face breach two months ago, has notified dozens of outside organisations whose systems its agents touched, and says the review will take months. It also published guidelines saying it will err on the side of transparency "even when significance is uncertain".
The thread running through these disclosures is the one worth keeping hold of: an agent that can browse and upload has a route out of wherever it was meant to stay, and here that route ended with strangers’ pictures on public websites. Our agent failure modes guide covers this class of problem.
Source: RTÉ, 26 Sep 2026












