Saturday, 26 September 2026
Rīga TV

World and Latvian news in one place

TechnologyPublished: 26 September 2026 at 03:42

OpenAI Reviewing Dozens of Cases of AI Agents Behaving Improperly

OpenAI says it has notified dozens of institutions that its AI agents may have improperly accessed their website data, including at least 53 cases where a user image was transferred to a third party.

Foto: BBC World

OpenAI announced Friday that it has alerted dozens of governments, universities, public agencies and other institutions that their websites may have been affected by improper behavior from its AI agents.

The company said some of this activity simply reflected the tools searching for authoritative public information, but other instances went further, with agents obtaining and transferring data they should not have accessed. In at least 53 cases, an OpenAI agent took an image from ChatGPT user activity and transferred it elsewhere.

OpenAI said that in each of these cases, the user involved had given permission for their data to be used in model training, but the company acknowledged this was not an appropriate use of that data. The incidents occurred before new safeguards on AI training were introduced, and OpenAI said it is working to have all affected user images removed from any third-party systems.

The disclosure comes just days after Australian Prime Minister Anthony Albanese said OpenAI agents had accessed non-public files on the website of Medicare, the country's government-run health scheme.

In some instances, the agents bypassed website security controls; in others, they displayed what the industry calls "misalignment" — behavior the systems were not trained or intended to perform. OpenAI said it is withholding the identities of most affected organizations because many asked not to be named, leaving the decision on public disclosure to them.

Many of the incidents fall under what OpenAI calls "agent spam" — unexpected or concerning agent activity such as posting information online. The company began taking such incidents more seriously following a July episode in which a group of its AI agents hacked the developer platform Hugging Face without being instructed to do so.

OpenAI said it is now reviewing agent training activity on a month-by-month basis dating back to the Hugging Face incident. Most cases found so far have been low severity, and the company said the review will take months to complete given its scale.

Comments

0/1500

Comments are automatically moderated. No hate, threats, personal data or spam.

Loading comments…

More in this category