Home Technology OpenAI models went rogue. We urgently need a better...
Technology

OpenAI models went rogue. We urgently need a better ‘hugging face’ investigation | Mackenzie Arnold and Stephan Llerena

Key Points

The breach won’t be the last – or the most dangerous – of its kind. We need an agency capable of full investigations into AI incidentsWhen OpenAI first revealed that its AI agents had autonomously hacked a major real-world company, Hugging Face, many assumed only one or two agents were involved. The truth, a new report reveals, is far stranger: the incident involved about 1,200 AI agents, 700 of which directly participated in the attack.

The breach won’t be the last – or the most dangerous – of its kind. We need an agency capable of full investigations into AI incidents

When OpenAI first revealed that its AI agents had autonomously hacked a major real-world company, Hugging Face, many assumed only one or two agents were involved. The truth, a new report reveals, is far stranger: the incident involved about 1,200 AI agents, 700 of which directly participated in the attack.

OpenAI invited researchers from METR, along with an expert from Redwood Research, to produce the new report, alongside the company’s own investigation. The findings shocked the experts.

Continue reading...
Mackenzie Arnold (PERSON) Stephan Llerena (PERSON) AI (ORG) METR (ORG) Redwood Research (ORG)
Originally published by The Guardian UK Read original →