Cybersecurity researchers aren’t happy about the guardrails on Anthropic’s Fable

Cybersecurity researchers aren’t happy about the guardrails on Anthropic’s Fable

Anthropic has unveiled its new AI model, Fable, which is being marketed as a public and constrained version of the highly anticipated cybersecurity tool, Mythos. However, the launch has sparked dissatisfaction among various cybersecurity researchers and professionals who have taken to social media to express their grievances. Valentina “Chompie” Palmiotti, a prominent security researcher at IBM X-Force, stated that Fable's restrictions significantly hinder its usability. "[Fable] rejects any request that could be tangentially cyber-related, including simple tasks like reading a blog post," she noted. When users attempt to engage in topics associated with cybersecurity or biology, Fable halts the conversation, citing its “safety measures” as the reason. These guardrails were implemented to mitigate the potential misuse of Fable for creating malware or compromising software, a priority for Anthropic. Similarly, concerns regarding biological weapon development have led to restrictions in that domain as well. In April, when Mythos was launched, access was limited to a select group of companies and organizations under what was termed Project Glasswing, aimed at safeguarding critical software and infrastructure. Recently, Anthropic expanded Mythos’s availability to numerous organizations across 15 countries, yet the response from cybersecurity experts remains tepid. Matt Suiche, an experienced figure in the cybersecurity field, highlighted that Fable’s programming often misinterprets requests. He remarked, "If you ask it to write secure code, it assumes it is cybersecurity-related work rather than adhering to software engineering best practices, leading to a downgrade in response quality." Fable defaults to Claude Opus 4.8 when it encounters its guardrails, which appear to be triggered by any terminology related to cybersecurity. While some researchers expressed frustration that even simple requests, such as a code review, activate these guardrails, others, like Suiche, acknowledge the necessity of caution in early model releases. He suggested that as Anthropic and similar companies collaborate with evolving cybersecurity firms, the model's guardrails may become more flexible over time. Anthropic has not yet responded to inquiries regarding these issues. In addition to the internal limitations of its models, Anthropic has established a Cyber Verification Program that cybersecurity professionals must apply to for access. Approved applicants face fewer restrictions when utilizing Claude for cybersecurity tasks. OpenAI has implemented a comparable initiative known as Trusted Access for Cyber.

Sources : TechCrunch

Published On : Jun 10, 2026, 16:35

AI
Revolutionizing Office Automation: Prentis Aims to Secure $100 Million in Funding

Prentis, a newly established AI research lab, is making waves in the tech industry as it prepares to raise $100 million ...

TechCrunch | Jul 25, 2026, 24:00
Revolutionizing Office Automation: Prentis Aims to Secure $100 Million in Funding
Streaming
Kalshi Challenges Netflix Over Controversial Documentary Trailer

Kalshi, the prediction market platform, has taken significant legal steps against Netflix, sending a cease-and-desist le...

TechCrunch | Jul 25, 2026, 17:10
Kalshi Challenges Netflix Over Controversial Documentary Trailer
Startups
The Boring Company Eyes $4 Billion Funding Boost Amid Expanding Tunnel Ventures

Elon Musk's tunneling enterprise, The Boring Company, is reportedly negotiating a substantial funding round of $4 billio...

TechCrunch | Jul 25, 2026, 19:50
The Boring Company Eyes $4 Billion Funding Boost Amid Expanding Tunnel Ventures
AI
Shifting Focus: The Cost-Effectiveness of AI Models Takes Center Stage

In recent years, the AI sector has been intensely focused on identifying the most advanced models. While this pursuit re...

Business Insider | Jul 25, 2026, 13:10
Shifting Focus: The Cost-Effectiveness of AI Models Takes Center Stage
Cybersecurity
The Elusive Phineas Fisher: The Hacktivist Who Took Down Spyware Giants

In the realm of cybersecurity, few figures are as intriguing as Phineas Fisher, a hacker who has evaded capture for near...

TechCrunch | Jul 25, 2026, 21:00
The Elusive Phineas Fisher: The Hacktivist Who Took Down Spyware Giants
View All News