OpenAI halts training of latest models as reports mount of AI agents going rogue

OpenAI halts training of latest models as reports mount of AI agents going rogue
OpenAI CEO Sam Altman. The company said it would resume training its latest models ‘only when we are confident that we have addition safeguards’. Photograph: ZUMA Press, Inc./Alamy
OpenAI CEO Sam Altman. The company said it would resume training its latest models ‘only when we are confident that we have addition safeguards’. Photograph: ZUMA Press, Inc./Alamy
OpenAI halts training of latest models as reports mount of AI agents going rogue

Decision follows disclosures that OpenAI agents searching government websites had acted in unexpected ways

OpenAI said it has paused training of its latest artificial intelligence models as reports of AI agents going rogue mount

The decision to halt development came just hours after the company disclosed Friday that it was reviewing several incidents from the summer in which OpenAI agents searching federal government websites acted in unexpected ways beyond what was asked of them while gathering and distributing information

Separately, the AI evaluator Transluce said agents that appeared to come from OpenAI tried unsuccessfully to hack into a US Department of Education website, a detail that OpenAI has not confirmed

OpenAI said in a statement that it will resume training “only when we are confident that we have additional safeguards” in place, adding that it expects it will have to “hit pause” again as AI develops and other issues emerge

Last week Australia’s prime minister, Anthony Albanese, revealed an OpenAI agent had breached the government’s national healthcare system – but said no sensitive information had been compromised

AI labs are facing pressure from lawmakers and tech experts to slow development so they can build guardrails to stop agents from acting on their own, hacking websites and disclosing nonpublic information. The heads of both OpenAI and rival Anthropic have called for a slowdown too

OpenAI hack on Australian government reveals anxiety at heart of global artificial intelligence dilemma
Read more

It is the second time in three months that OpenAI has halted development of its models. The first came in July after disclosure of a cyber-attack targeting AI startup Hugging Face, a now notorious incident that raised fears the industry was losing control

In a meeting with Chinese president Xi Jinping this week, Donald Trump agreed to share information on AI dangers and coordinate efforts to keep it safe. Trump believes AI fears are overblown, though, and later suggested that he plans no crackdown of his own

The US is not going to be “putting on brakes”, Trump told reporters outside the White House. “They want to stop our progress because we’re leading China by a lot, and we’re going to keep it that way.”

The latest OpenAI incidents did not appear to involve the disclosure of any nonpublic information but were concerning enough for the company to warn the federal agencies involved

In the education department incident, OpenAI agents found API “developer keys” to access government data, though ultimately only publicly available information was gathered

In another case involving the securities and exchange commission, agents found information freely available to all but then posted it elsewhere on the internet, an act that went beyond what they were instructed to do

US Securities and Exchange Commission spokesperson Kurt Hopfenspirger said on Saturday that “no nonpublic information was accessed”

The Department of Education said earlier that it found “no evidence of any impact to our website or databases”

Several other AI companies have disclosed incidents of their models going rogue and even hacking websites

OpenAI’s CEO, Sam Altman, said in a social media post on Friday that the Hugging Face incident “is still the most severe event we’ve seen”

OpenAI previously shared six other reports of “unexpected or concerning” behaviour in AI models and introduced a framework for tracking, probing and disclosing instances

Explorez davantage ces sujets Partager
Réutiliser ce contenu

En rapport:

Formation en automatisation numérique au Bénin : 5 compétences clés recherchées par les employeurs en 2026

<a href="https://yoursite.com/automation-africa/" title="WhatsApp Marketing Automation Africa : 6 erreurs dangereuses commises par les marques au Nigéria”>
Automatisation du marketing WhatsApp en Afrique : 6 erreurs dangereuses commises par les marques au Nigéria

Vous souhaitez apprendre cela de manière pratique ?

Rejoindre Justfine Infotech et développer de véritables compétences numériques en IA, automatisation, développement web, marketing digital, bureautique, e-commerce, travail indépendant et cybersécurité.

Programmes disponibles :
Certificat de 6 semaines • Certificat professionnel de 3 mois • Diplôme de 6 mois • Diplôme professionnel complet

WhatsApp :
+229 01 57 57 99 15
+229 01 66 68 11 60

Inscrivez-vous dès maintenant

Source: www.theguardian.com

Laisser un commentaire

Votre adresse e-mail ne sera pas publiée. Les champs obligatoires sont indiqués avec *

Défiler vers le haut