{"id":4173,"date":"2026-08-07T13:14:23","date_gmt":"2026-08-07T13:14:23","guid":{"rendered":"https:\/\/justfineinfotech.com\/first-openai-now-meta-why-do-ai-hacks-keep-happening\/"},"modified":"2026-08-07T13:14:23","modified_gmt":"2026-08-07T13:14:23","slug":"first-openai-now-meta-why-do-ai-hacks-keep-happening","status":"publish","type":"post","link":"https:\/\/justfineinfotech.com\/fr\/first-openai-now-meta-why-do-ai-hacks-keep-happening\/","title":{"rendered":"First OpenAI, now Meta &#8211; why do AI hacks keep happening?"},"content":{"rendered":"<figure>\n<img decoding=\"async\" src=\"https:\/\/justfineinfotech.com\/wp-content\/uploads\/2026\/08\/grey-placeholder-5.png\" alt=\"\"><br \/>\n<img decoding=\"async\" src=\"https:\/\/justfineinfotech.com\/wp-content\/uploads\/2026\/08\/1e741d30-91aa-11f1-9c12-47c0acd34ba6.jpg.webp\" alt=\"Getty Images A software engineer, wearing a light brown shirt, looks at a computer monitor in front of him with a concerned expression. A larger screen with computer code on it is on a wall behind him.\"><br \/>\nGetty Images<br \/>\n<\/figure>\n<p>Over the last fortnight, reports of AI models going beyond their expected bounds &#8211; be that technically or morally &#8211; have been seemingly unavoidable<\/p>\n<p>What started with a trickle &#8211; ChatGPT-maker OpenAI admitting their AI had hacked the site Hugging Face &#8211; has turned into a flood of groups revealing they had discovered instances of AI going out of control<\/p>\n<p>Claude-maker Anthropic, Meta and the UK&#8217;s AI Security Institute (AISI) have now each reported incidents which seem to paint a worrying picture of a world in which tech going rogue is the norm<\/p>\n<p>In reality, each case offers a window into the risks posed by increasingly capable AI agents &#8211; and the importance of testing their limits before they are released to the world<\/p>\n<p>The OpenAI incident has, as Hugging Face&#8217;s co-founder Thomas Wolf described it, come as a &#8220;wake-up call&#8221; for the tech industry since it happened at the end of July<\/p>\n<p>It was a big moment which caused big companies to reflect on their own systems &#8211; and, in some cases, check they hadn&#8217;t missed something similarly shocking<\/p>\n<p>Anthropic was the first to act. On Friday, the company found three instances out of thousands where its model Claude had managed to gain access to the internet<\/p>\n<p>Then on Tuesday, the AISI, the UK government agency which evaluates cutting-edge models, then said it had detected a &#8220;security incident&#8221; during a routine evaluation<\/p>\n<p>It had been testing models by both OpenAI and Anthropic, and found they too tried to carry out cyber-attacks &#8211; calling for &#8220;scrutiny, transparency, and action&#8221;<\/p>\n<p>Finally followed Meta, which revealed one of its AI models had inadvertently been allowed to access the internet due to a &#8220;misconfiguration&#8221; during a third-party test<\/p>\n<p>In disclosing the incident, it is following in the footsteps of those before it<\/p>\n<p>Before AI models are released to the public, they are put to the test in a series of internal and external evaluations<\/p>\n<p>The aim is to figure out their potential to do good or bad, as well has how they perform in benchmarks measuring their skills<\/p>\n<p>These typically take place in what are known as &#8220;sandboxes&#8221;. These are protected spaces designed to mirror real systems &#8211; but with strict guardrails in place<\/p>\n<p>In the OpenAI-Hugging Face incident, the AI attacked the sandbox itself, finding a vulnerability which let it access the internet and &#8220;go rogue&#8221;<\/p><figcaption>Watch: Why is the OpenAI cyber-attack so alarming?<\/figcaption><p>Meanwhile the AISI said its own incident, which saw two powerful AI tools create fake human profiles to try and trick people in attempted cyber-attacks, was not down to an issue with the sandbox<\/p>\n<p>Instead, it was due to how it went about its tests<\/p>\n<p>The models it tested were granted access to the internet, and the AISI also disabled in-built filters that would usually block dangerous cyber-attacks<\/p>\n<p>&#8220;To some degree, our evaluation design choices and specific configurations enabled the behaviour,&#8221; it said, while noting its unexpected &#8220;signs of novel, potentially deceptive behaviours&#8221;<\/p>\n<p>Prof Alan Woodward, professor of cyber-security at the University of Surrey, said these cases &#8211; while distinct in what happened and why &#8211; tell an important story<\/p>\n<p>&#8220;For 30 years, one rule of software testing held firm: whatever happens in the test environment stays in the test environment,&#8221; he said<\/p>\n<p>&#8220;In the past month, that rule has been broken three times.&#8221;<\/p>\n<p>&#8220;One model broke out. One walked through a door left open by mistake. One was deliberately given the keys so testers could measure what it would do.&#8221;<\/p>\n<p>He said these were different causes, but they had the same lesson &#8211; &#8220;the testing lab is now where the risk lives&#8221;<\/p>\n<p>He told the BBC that as models become more capable, more must be done to secure the environments where they are tested<\/p>\n<p>&#8220;Testing an AI agent is less like checking code and more like handling a hazardous material: sealed rooms, constant monitoring of what leaves the building, a rehearsed containment plan,&#8221; he said<\/p>\n<p>&#8220;AISI contained its incident within an hour. The next organisation may not.&#8221;<\/p>\n<p>For those developing AI tools which are designed to take actions on a person&#8217;s behalf, there is a careful balance to be struck between harnessing their benefits and exposing their risks<\/p>\n<p>The benefit is significant. In theory, we could be able to liberate ourselves of dull, menial tasks, such as replying to emails, going to meetings or managing calendars and diaries, by delegating these to capable bots<\/p>\n<p>The downside is that with great power comes great responsibility, and risk<\/p>\n<p>It&#8217;s something particularly realised when handing power to tools which are not, like us, able to bring a range of values, context and understanding to decisions we made<\/p>\n<p>&#8220;Recent incidents of frontier AI models carrying out unsanctioned actions and, in some cases, human-like deceptive behaviour on the open internet are a serious reminder of the risks AI capabilities pose,&#8221; said Ollie Whitehouse, the National Cyber Security Centre&#8217;s chief technology officer on Tuesday<\/p>\n<p>Some believe the sheer volume of tasks that will be handled by these tools will mean human oversight might not be enough to contain the problem of models going rogue<\/p>\n<p>But in the meantime, many feel strengthening oversight overall is vital if development continues at its same, frenzied pace<\/p>\n<figure>\n<img decoding=\"async\" src=\"https:\/\/justfineinfotech.com\/wp-content\/uploads\/2026\/08\/fb29e250-91aa-11f1-9c12-47c0acd34ba6.jpg.webp\" alt=\"Getty Images ChatGPT, OpenClaw and Claude app widgets shown on a smartphone screen\"><br \/>\nGetty Images<figcaption>AI agents are already becoming a part of our digital lives through services like ChatGPT, OpenClaw and Claude<\/figcaption><\/figure>\n<p>It is unlikely Meta will be the last to emerge with findings of models showing they have, as Prof Woodward puts it, &#8220;gone to school&#8221; &#8211; and learnt our own ways of finding and exploiting gaps in systems<\/p>\n<p>For some, these episodes point to clear security failures on the part of AI companies leading the charge on this game-changing, era-defining tech<\/p>\n<p>For others, they are merely <a href=\"https:\/\/justfineinfotech.com\/fr\/meta-becomes-latest-firm-to-say-its-ai-hacked-another-company\/\" title=\"Meta est la derni\u00e8re entreprise en date \u00e0 affirmer que son IA a pirat\u00e9 une autre soci\u00e9t\u00e9.\">another<\/a> vehicle for tech firms to hype up their powerful models and compete with rivals<\/p>\n<p>For me, both theories hold some grain of truth<\/p>\n<p>But in rearing their head one after another, these events have nonetheless spurred fears about AI&#8217;s capabilities and where these are headed as developers forge ahead<\/p>\n<p>And the question inevitably moves to what regulators can and should do next<\/p>\n<p>Michael Birtwistle, associate director at the Ada Lovelace Institute, makes the point that the UK lacks legal incentives for AI firms to prevent systems from developing capabilities which could pose dangers, and that there are no repercussions if testing protocols fail<\/p>\n<p>More broadly, Dr Imogen Stead, AI policy manager at the Centre for Long-Term Resilience, told the BBC that with opportunities to test frontier AI systems narrowing for many, governments should follow the UK in setting up dedicated institutes for testing<\/p>\n<p>Improving third-party evaluations with initiatives such as a &#8220;trusted tester scheme&#8221; for the most risky types of challenges could also be used to limit adverse impacts, she said<\/p>\n<p>Rather than fear an AI-cyber apocalypse in the meantime, Prof Woodward says, &#8220;it&#8217;s a case of &#8216;keep calm and fix stuff'&#8221;<\/p>\n<figure>\n<img decoding=\"async\" src=\"https:\/\/justfineinfotech.com\/wp-content\/uploads\/2026\/08\/348b21e0-26a8-11f0-8f57-b7237f6a66e6.png.webp\" alt=\"Une banni\u00e8re publicitaire verte compos\u00e9e de carr\u00e9s et de rectangles noirs formant des pixels appara\u00eet en venant de la droite. Le texte indique\u00a0: \u201c\u00a0Tech Decoded\u00a0: Les principales actualit\u00e9s technologiques du monde directement dans votre bo\u00eete mail tous les lundis.\u00a0\u201d\"><br \/>\n<\/figure>\n<p>Cybers\u00e9curit\u00e9<br \/>\nIntelligence artificielle<br \/>\nM\u00e9ta<\/p>\n<div style=\"clear:both;margin:30px 0 15px 0\">\n<p>\n    <strong>En rapport:<\/strong><br \/>\n    <a href=\"https:\/\/yoursite.com\/automation-training-benin\/\" title=\"Formation en automatisation num\u00e9rique au B\u00e9nin\u00a0: 5 comp\u00e9tences cl\u00e9s recherch\u00e9es par les employeurs en 2026\" target=\"_blank\" rel=\"noopener\"><br \/>\n      Formation en automatisation num\u00e9rique au B\u00e9nin\u00a0: 5 comp\u00e9tences cl\u00e9s recherch\u00e9es par les employeurs en 2026<br \/>\n    <\/a>\n  <\/p>\n<p>\n    &lt;a href=&quot;https:\/\/yoursite.com\/automation-africa\/&quot; title=&quot;<a href=\"https:\/\/justfineinfotech.com\/fr\/meta-will-allow-rival-ai-chatbots-on-whatsapp-in-europe-but-for-a-fee-techcrunch\/\" title=\"Meta will allow rival AI chatbots on WhatsApp in Europe, but for a fee | TechCrunch\">WhatsApp<\/a> <a href=\"https:\/\/justfineinfotech.com\/fr\/social-media-marketing-salary-your-2026-guide\/\" title=\"Social Media Marketing Salary: Your 2026 Guide\">Commercialisation<\/a> Automation Africa : 6 erreurs dangereuses commises par les marques au Nig\u00e9ria<br \/>\n      Automatisation du marketing WhatsApp en Afrique\u00a0: 6 erreurs dangereuses commises par les marques au Nig\u00e9ria<br \/>\n    <\/a>\n  <\/p>\n<\/div>\n<div style=\"clear:both;margin:30px 0;padding:25px;background:#f8f9fc;border:1px solid #ddd;border-radius:8px;text-align:center\">\n<h3>Vous souhaitez apprendre cela de mani\u00e8re pratique ?<\/h3>\n<p>Rejoindre <strong>Justfine Infotech<\/strong> et d\u00e9velopper de v\u00e9ritables comp\u00e9tences num\u00e9riques en IA, automatisation, d\u00e9veloppement web, marketing digital, bureautique, e-commerce, travail ind\u00e9pendant et cybers\u00e9curit\u00e9.<\/p>\n<p><strong>Programmes disponibles :<\/strong><br \/>\n  Certificat de 6 semaines \u2022 Certificat professionnel de 3 mois \u2022 Dipl\u00f4me de 6 mois \u2022 Dipl\u00f4me professionnel complet<\/p>\n<p><strong>WhatsApp :<\/strong><br \/>\n  +229 01 57 57 99 15<br \/>\n  +229 01 66 68 11 60<\/p>\n<p><a href=\"https:\/\/api.whatsapp.com\/send?phone=2348132690270&amp;text=Hello\" target=\"_blank\" rel=\"noopener\">Inscrivez-vous d\u00e8s maintenant<\/a><\/p>\n<\/div>\n<p class=\"ani-source\">Source: <a href=\"https:\/\/www.bbc.com\/news\/articles\/cp30989ee1wo\" target=\"_blank\" rel=\"nofollow noopener\">www.bbc.com<\/a><\/p>","protected":false},"excerpt":{"rendered":"<p>Over the last fortnight, reports of AI models going beyond their expected bounds &#8211; be that technically or morally &#8211; have been seemingly unavoidable<\/p>","protected":false},"author":1,"featured_media":4179,"comment_status":"open","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"site-sidebar-layout":"default","site-content-layout":"","ast-site-content-layout":"default","site-content-style":"default","site-sidebar-style":"default","ast-global-header-display":"","ast-banner-title-visibility":"","ast-main-header-display":"","ast-hfb-above-header-display":"","ast-hfb-below-header-display":"","ast-hfb-mobile-header-display":"","site-post-title":"","ast-breadcrumbs-content":"","ast-featured-img":"","footer-sml-layout":"","ast-disable-related-posts":"","theme-transparent-header-meta":"","adv-header-id-meta":"","stick-header-meta":"","header-above-stick-meta":"","header-main-stick-meta":"","header-below-stick-meta":"","astra-migrate-meta-layouts":"default","ast-page-background-enabled":"default","ast-page-background-meta":{"desktop":{"background-color":"var(--ast-global-color-4)","background-image":"","background-repeat":"repeat","background-position":"center center","background-size":"auto","background-attachment":"scroll","background-type":"","background-media":"","overlay-type":"","overlay-color":"","overlay-opacity":"","overlay-gradient":""},"tablet":{"background-color":"","background-image":"","background-repeat":"repeat","background-position":"center center","background-size":"auto","background-attachment":"scroll","background-type":"","background-media":"","overlay-type":"","overlay-color":"","overlay-opacity":"","overlay-gradient":""},"mobile":{"background-color":"","background-image":"","background-repeat":"repeat","background-position":"center center","background-size":"auto","background-attachment":"scroll","background-type":"","background-media":"","overlay-type":"","overlay-color":"","overlay-opacity":"","overlay-gradient":""}},"ast-content-background-meta":{"desktop":{"background-color":"var(--ast-global-color-5)","background-image":"","background-repeat":"repeat","background-position":"center center","background-size":"auto","background-attachment":"scroll","background-type":"","background-media":"","overlay-type":"","overlay-color":"","overlay-opacity":"","overlay-gradient":""},"tablet":{"background-color":"var(--ast-global-color-5)","background-image":"","background-repeat":"repeat","background-position":"center center","background-size":"auto","background-attachment":"scroll","background-type":"","background-media":"","overlay-type":"","overlay-color":"","overlay-opacity":"","overlay-gradient":""},"mobile":{"background-color":"var(--ast-global-color-5)","background-image":"","background-repeat":"repeat","background-position":"center center","background-size":"auto","background-attachment":"scroll","background-type":"","background-media":"","overlay-type":"","overlay-color":"","overlay-opacity":"","overlay-gradient":""}},"footnotes":""},"categories":[63],"tags":[99,1069,349,101,79],"class_list":["post-4173","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-ai-tools-chatgpt-updates","tag-first","tag-hacks","tag-keep","tag-meta","tag-openai"],"_links":{"self":[{"href":"https:\/\/justfineinfotech.com\/fr\/wp-json\/wp\/v2\/posts\/4173","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/justfineinfotech.com\/fr\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/justfineinfotech.com\/fr\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/justfineinfotech.com\/fr\/wp-json\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/justfineinfotech.com\/fr\/wp-json\/wp\/v2\/comments?post=4173"}],"version-history":[{"count":1,"href":"https:\/\/justfineinfotech.com\/fr\/wp-json\/wp\/v2\/posts\/4173\/revisions"}],"predecessor-version":[{"id":4178,"href":"https:\/\/justfineinfotech.com\/fr\/wp-json\/wp\/v2\/posts\/4173\/revisions\/4178"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/justfineinfotech.com\/fr\/wp-json\/wp\/v2\/media\/4179"}],"wp:attachment":[{"href":"https:\/\/justfineinfotech.com\/fr\/wp-json\/wp\/v2\/media?parent=4173"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/justfineinfotech.com\/fr\/wp-json\/wp\/v2\/categories?post=4173"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/justfineinfotech.com\/fr\/wp-json\/wp\/v2\/tags?post=4173"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}