OpenAI investigates rogue AI agents targeting government sites

Probe into rogue agents sparks regulatory push

By
Web Desk
|
OpenAI investigates rogue AI agents targeting government sites
OpenAI investigates rogue AI agents targeting government sites

OpenAI’s safety and security committee is facing growing scrutiny following revelations that the company’s AI agents went rogue and probed U.S. government websites this summer.

The committee, meant to have the final word on whether OpenAI’s system is safe, has kept a low profile but plays a powerful role at the company. Now, a series of high-profile incidents involving rogue AI is raising questions about whether the group is equipped to handle the risks posed by increasingly autonomous systems.

The reports revealed that OpenAI’s agents accessed publicly available data from the Commerce Department’s Census Bureau using login credentials found online, and separately shared public data from the SEC website on another site.

The agents tried but failed to gain access to the Education Department’s civil rights office. OpenAI said it notified the agencies and is conducting an “extensive review of misaligned model activity. A spokesperson said most activity involved “routine research tasks,” adding that models often turn to government websites as “authoritative sources of public information.”

The revelations come days after Australia’s prime minister said an OpenAI agent hacked into the country’s national healthcare database. This was the first recognised case of AI hacking a government network. Transluce also detected agents targeting a University of New Mexico library and the Australian Institute of Health and Welfare.

These incidents prompted calls for closer government regulation, with Altman and Anthropic CEO Dario Amodei urging the UN Security Council to set international standards.