Dernière minute
CNFormer Chinese premier and economic reformer Zhu Rongji dies at 97JP辺野古沖転覆事故で両親が同志社に要望書 教育活動の検証求めTRLetonya hava sahasına giren İHA düşürüldüCN日本阿蘇中岳火山活動加劇 警戒等級升至第3級KR국민의힘, 청와대 4년 중임제 개헌 언급에 "정국 전환용·탈옥형 개헌" 비판ITCisgiordania, cinque italiani nel villaggio di Qusra assediato dai coloniINBar Council of India Chairman Manan Kumar Mishra at Centre of Controversy Over NALSAR RowKR검경, '선피아 카르텔' 의혹 중앙선관위 등 강제수사 착수DEZweijähriger verschwindet aus Kita in Haldensleben: Polizei ermitteltRUIsraeli defense officials shocked by rapid speed of Iran's military recoveryCNFormer Chinese premier and economic reformer Zhu Rongji dies at 97JP辺野古沖転覆事故で両親が同志社に要望書 教育活動の検証求めTRLetonya hava sahasına giren İHA düşürüldüCN日本阿蘇中岳火山活動加劇 警戒等級升至第3級KR국민의힘, 청와대 4년 중임제 개헌 언급에 "정국 전환용·탈옥형 개헌" 비판ITCisgiordania, cinque italiani nel villaggio di Qusra assediato dai coloniINBar Council of India Chairman Manan Kumar Mishra at Centre of Controversy Over NALSAR RowKR검경, '선피아 카르텔' 의혹 중앙선관위 등 강제수사 착수DEZweijähriger verschwindet aus Kita in Haldensleben: Polizei ermitteltRUIsraeli defense officials shocked by rapid speed of Iran's military recovery
Newsgather
RetourAnthropic's Fable AI faces backlash over cybersecurity restrictions
Anthropic's Fable AI faces backlash over cybersecurity restrictions
En développement
TechCrunch10/06/2026Tech2 min de lectureUnited States

Anthropic's Fable AI faces backlash over cybersecurity restrictions

L'essentiel

Anthropic's new Fable AI model, a public version of its cybersecurity tool Mythos, is facing criticism from researchers for overly strict guardrails that block even benign requests related to cybersecurity or biology.

Résumé généré par IA

Taille de police

Anthropic released its latest model Fable on Tuesday, billing it as a public and limited version of its powerful and much-hyped cybersecurity model Mythos.

But not everyone is happy with the restrictions, and a number of cybersecurity researchers and professionals have aired complaints online.

“[Fable] rejects any request that could be tangentially cyber related. Even innocuous tasks like reading a blog post,” said Valentina “Chompie” Palmiotti, a well-known security researcher who works at IBM X-Force.

When a prompt triggers its guardrails, Fable pauses the chat and says that its “safety measures flagged this message for cybersecurity or biology topics.”

The guardrails were put in place to limit the risk that Fable could be used to develop malware or compromise software — a long-standing concern within Anthropic. The restrictions on biology come from a similar concern around developing biological weapons.

When the AI giant released Mythos in April, it restricted the model to a limited number of companies and organizations in what it called Project Glasswing, an effort to deploy the model to secure critical software and infrastructure. Last week, Anthropic expanded access to Mythos to hundreds of organizations in 15 countries.

But despite the good intentions, many cybersecurity experts are still put off by the haphazard nature of the restrictions. Matt Suiche, a cybersecurity veteran, told TechCrunch that “if you ask it to write secure code, it assumes it is cybersecurity related work instead of software engineering best practices, and you get downgraded.” Fable is programmed to fall back to Claude Opus 4.8 if it hits a guardrail. “It seems to be keyword based, so anything in the lexical field of ‘cybersecurity’ triggers the guardrails.”

“But it is understandable as we are still in the early days and they are still adapting their guardrails. I am sure they are going to evolve over time as Anthropic and other frontier model companies will collaborate more with the current new generation of cybersecurity companies,” said Suiche, who is a member of the technical staff at Tolmo, an AI cybersecurity startup. “It’s better to catch more people than not enough when you do such a release and to relax the guardrails over time.”

Another researcher griped on X that “even asking for a code review” triggers Fable’s guardrails.

Anthropic did not immediately respond to a request for comment.

Sujets liés

This article was originally published by TechCrunch.

Articles liés

Plus sur ce sujetAnthropic