Business

OpenAI says its AI agents posted user images online in error


ROGUE AI AGENTS 

According to OpenAI, agents that it uses for its research transmitted the training data to external platforms.

The incidents occurred before OpenAI strengthened the security protocols of its research environment in August following other rogue actions by AI agents.

The company said it is scrutinising the past activity of its AI agents, work that “will take months to complete”.

“Most of the activity we’ve reviewed so far involved routine research tasks, such as accessing public web content to answer questions. Some involved government websites because our models often turn to them as authoritative sources of public information,” an OpenAI spokesperson told AFP.

OpenAI chief executive Sam Altman acknowledged Friday on X that “we have not been as fast as we would have liked” in reviewing and disclosing the incidents.

But he said it was important to “balance our desire for transparency” with assessing the massive volume of data to be analysed.

On Jul 21 OpenAI revealed that during tests it ran that month, two of its models escaped their closed environments, got onto the internet on their own and broke into the internal systems of Hugging Face, a kind of online library for AI software.

The episode drew wide attention and fed worries that the biggest AI companies cannot keep their own models under control.

Altman reiterated Friday that the Hugging Face hack “is still the most severe event we’ve seen”.

That discovery was followed by revelations of several similar incidents at OpenAI and its rivals, such as Anthropic and Meta.

On Wednesday in New York, Australian Prime Minister Anthony Albanese said an OpenAI agent had gained unauthorised access to a government health portal in June, and the leader criticised the company for delaying its notification to the authorities.



Source link

Leave a Reply

Your email address will not be published. Required fields are marked *