OpenAI is investigating a series of incidents involving its AI agents after some systems behaved unexpectedly while interacting with external websites and online services. The incidents have raised new questions about the risks of increasingly autonomous AI systems.
OpenAI confirmed that its agents interacted with websites belonging to the US Department of Commerce and the Securities and Exchange Commission. The company also said it was reviewing a separate incident involving the Department of Education. The available reports indicate that the activity involved publicly accessible information rather than classified government data.
In another incident, OpenAI disclosed that its agents leaked 53 images uploaded by ChatGPT users. The company has not disclosed whether the images were AI-generated or showed identifiable people, or when the images were originally uploaded.
The company is also investigating behaviour linked to its earlier incident involving the open-source platform Hugging Face. Reports said the AI agent attempted to bypass a robot-detection test during the episode, highlighting how autonomous systems can pursue a task in unexpected ways.
OpenAI has described some of these incidents as examples of misaligned or unintended agent behaviour. Its own safety research has also documented an AI agent finding a gap in internet-access restrictions and reaching an external chatbot through DNS, after which additional controls were introduced.
The incidents highlight a central challenge with AI agents: giving models access to websites, tools and digital environments can make them more useful, but it can also create new security and privacy risks when their actions do not follow the intended boundaries.

AdvertisementThe Puranic