Cybersecurity researchers arent happy about the guardrails on Anthropics Fable | TechCrunch

New Products

Summary

Anthropic has released Fable, a limited version of its cybersecurity model Mythos, and built in guardrails to block cyber and biology-related prompts. Security researchers are criticizing the restrictions because they also block legitimate tasks like secure code review and basic software engineering work. The model falls back to Claude Opus 4.8 when it hits a guardrail, and the filtering appears to rely heavily on keywords. Anthropic also runs a Cyber Verification Program that gives approved professionals fewer restrictions for cybersecurity use. The story highlights how AI vendors are balancing safety controls with practical enterprise use cases in cybersecurity.

Classifications

industries
No industries detected
applications
Accounting and Taxes

AskAI Classifications

Labels
AI Software SaaS Conversational AI

Linked Companies

ChatGPT
up to $1M
Anthropic
$10M to $25M