Les chercheurs en cybersécurité sont mécontents des garde-fous de Fable d'Anthropic | TechCrunch

Les chercheurs en cybersécurité sont mécontents des garde-fous de Fable d'Anthropic | TechCrunch

Anthropic a lancé mardi son dernier modèle, Fable, le présentant comme une version publique et limitée de son puissant et très médiatisé modèle de cybersécurité Mythos.

But not everyone is happy with the restrictions, and a number of cybersécuritéresearchers et professionalshave airedcomplaints online. 

“[Fable] rejects any request that could be tangentially cyber related. Even innocuous tasks like reading a blog post,” dit Valentina “Chompie” Palmiotti, a well-known security researcher who works at IBM X-Force. 

When a prompt triggers its guardrails, Fable pauses the chat and says that its “safety measures flagged this message for cybersecurity or biology topics.”

The guardrails were put in place to limit the risk that Fable could be used to develop malware or compromise software — a long-standing concern within Anthropic. The restrictions on biology come from a similar concern around developing biological weapons

When the AI giant released Mythos in April, it restricted the model to a limited number of companies and organizations in what it called Project Glasswing, an effort to deploy the model to secure critical software and infrastructure. Last week, Anthropic expanded access to Mythos to hundreds of organizations in 15 countries. 

But despite the good intentions, many cybersecurity experts are still put off by the haphazard nature of the restrictions. Matt Suiche, a cybersecurity veteran, told TechCrunch that “if you ask it to write secure code, it assumes it is cybersecurity related work instead of software engineering best practices, and you get downgraded.” Fable is programmed to fall back to Claude Opus 4.8 if it hits a guardrail. “It seems to be keyword based, so anything in the lexical field of ‘cybersecurity’ triggers the guardrails.”

Contactez-nous

Do you have more information about how hackers are using AI? Or how cybersecuity companies are using AI? We’d love to hear from you. From a non-work device and network, you can contact Lorenzo Franceschi-Bicchierai securely on Signal at +1 917 257 1382, or via Telegram and Keybase @lorenzofb, or email.

“But it is understandable as we are still in the early days and they are still adapting their guardrails. I am sure they are going to evolve over time as Anthropic and other frontier model companies will collaborate more with the current new generation of cybersecurity companies,” said Suiche, who is a member of the technical staff at Tolmo, an AI cybersecurity startup. “It’s better to catch more people than not enough when you do such a release and to relax the guardrails over time.”

Another researcher griped on X that “even asking for a code review” triggers Fable’s guardrails. 

Anthropic did not immediately respond to a request for comment

Apart from guardrails inside its models, Anthropic requires cybersecurity professionals to apply to the Cyber Verification Program. If they get approved, the applicants have fewer limitations on using Claude for cybersecurity work. OpenAI has a similar program called Trusted Access for Cyber

En rapport:
<a href="https://yoursite.com/automation-training-benin/” title=”Formation en automatisation numérique au Bénin : 5 compétences gagnantes que les employeurs recherchent en 2026″>
Formation en automatisation numérique au Bénin : 5 compétences clés recherchées par les employeurs en 2026

<a href="https://yoursite.com/automation-africa/" title="WhatsApp Marketing Automation Africa : 6 erreurs dangereuses commises par les marques au Nigéria”>
Automatisation du marketing WhatsApp en Afrique : 6 erreurs dangereuses commises par les marques au Nigéria

Vous souhaitez apprendre cela de manière pratique ?

Rejoindre Justfine Infotech et développer de véritables compétences numériques en IA, automatisation, développement web, marketing digital, bureautique, e-commerce, travail indépendant et cybersécurité.

Programmes disponibles :
Certificat de 6 semaines • Certificat professionnel de 3 mois • Diplôme de 6 mois • Diplôme professionnel complet

WhatsApp :
+229 01 57 57 99 15
+229 01 66 68 11 60

Inscrivez-vous dès maintenant

Source: techcrunch.com

Laisser un commentaire

Votre adresse e-mail ne sera pas publiée. Les champs obligatoires sont indiqués avec *

Défiler vers le haut