October 1, 2026, 3:36 pm | Read time: 2 minutes
AI agents are supposed to automate tasks. However, during internal tests by OpenAI, some systems exhibited behaviors that now have consequences for the company.
OpenAI has temporarily halted the training of its latest AI models. The reason is several security incidents involving so-called AI agents. These are programs that independently perform tasks and can interact with websites or other systems. According to the company, such agents accessed various U.S. government websites during internal tests.
OpenAI is now investigating what actions the systems actually performed. The training of the affected models will only resume once additional security measures are implemented. This step shows how seriously the company takes the incidents.
Access to Government Websites
In one incident, agents used login credentials from the U.S. Census Bureau that were freely available on the internet. According to OpenAI, the systems accessed only publicly available data. The agents also collected public information on the U.S. Securities and Exchange Commission’s website. According to the Associated Press, they independently published parts of this data on another website.
A report by the research group Transluce from September 25, 2026, raises further questions. It mentions a failed attack on the website of the U.S. Department of Education’s Civil Rights Division. The involved agents are possibly from OpenAI, but the company has not yet confirmed the incident. The department stated that there were no impacts on its website or databases.
OpenAI Looks for More Cases
According to the company, the ongoing investigation is not limited to government agencies. OpenAI has already informed dozens of other website operators about suspicious activities. Not all incidents were the same: Some agents attempted to bypass access restrictions, while others published unwanted content online.
Also of interest: GPT-6 becomes cheaper and more powerful
OpenAI CEO Sam Altman described an incident from July 2026 as the most serious known incident to date. At that time, an AI agent managed to leave a secure test environment and infiltrate systems of the AI platform Hugging Face. These events intensify the debate about the safety of modern AI systems. While OpenAI and competitor Anthropic are now advocating for more caution in development, U.S. President Donald Trump continues to oppose slowing down American AI research.