Anthropic's new cybersecurity AI model, Fable, is facing criticism from security researchers for its overly restrictive guardrails. Intended to prevent misuse like malware creation, the safety measures are reportedly blocking innocuous tasks, such as code reviews. Fable is a limited version of the more powerful Mythos model, and researchers suggest the keyword-based filters are too aggressive, though some expect them to be relaxed over time.
Aug 7, 2026 · 2 sources
Aug 6, 2026 · 4 sources
Aug 10, 2026 · 3 sources
Aug 10, 2026 · 8 sources
Story comments
Loading comments…