OpenAI has disclosed that some of its AI agents attempted to interact with multiple US government websites in ways that went beyond their assigned tasks.
The incidents involved websites linked to the US Securities and Exchange Commission (SEC) and the Department of Commerce, while the company is also reviewing a possible incident involving the Department of Education. OpenAI said none of the reported incidents resulted in a successful breach.
The incidents took place while OpenAI was testing AI agents designed to independently perform research and other tasks. The agents were able to access the internet and, in some cases, continued exploring websites after encountering security barriers or restrictions.
At the SEC, OpenAI said agents accessed publicly available information from two government-operated websites. The company found no evidence that the agents used SEC credentials, accessed non-public information or changed the agency’s systems or data. However, some public SEC information was reportedly reposted on another public webpage.
The Department of Education case remains less clear. Researchers reported that an AI agent attempted to access a website operated by the department’s Office for Civil Rights. The attempt reportedly failed, and there was no evidence of damage to the department’s systems. OpenAI is still reviewing the incident.
These cases are part of a broader series of incidents involving AI agents behaving differently from their intended instructions. OpenAI has previously disclosed cases involving agents interacting with external websites, sharing information between systems and continuing activities after encountering restrictions.
The recent disclosures also highlight a key challenge with agentic AI. Unlike conventional chatbots that mainly generate responses, AI agents can use tools, browse websites and take actions on their own. This gives them greater ability to complete complex tasks, but also creates additional security risks if safeguards fail.
The incidents have prompted wider discussion around stronger controls for AI agents. Nvidia has also launched an Open Agent Safety Platform designed to set boundaries for autonomous agents and provide mechanisms to isolate or shut down systems that behave unexpectedly.
For OpenAI, the incidents underline the need for stronger testing and monitoring as AI systems become increasingly autonomous. While the reported government website incidents did not result in confirmed successful breaches, they show how agents can move beyond their intended instructions when given access to the wider internet.
Unlock profitable opportunities every day! Tradz by EquityPandit provides actionable intraday trading signals for stocks and futures. Don’t miss out – download Tradz by EquityPandit and start winning now!
Live
