Anthropic launches Claude Fable 5, its first 'Mythos-class' AI model with new safeguards
After deeming it too risky, Anthropic releases its most capable model yet—Claude Fable 5.
Anthropic has unveiled Claude Fable 5, the first widely available model from its previously unreleased Mythos class. The company had earlier described the Mythos family as too dangerous for public release due to its advanced cybersecurity capabilities. Now, Anthropic says new safeguards make the release possible, blocking responses in specific high-risk areas such as cybersecurity and biology, and falling back to the 'honest' Claude Opus 4.8 when necessary. In testing, 95% of sessions ran entirely on Fable without requiring fallback.
Fable 5 shows exceptional performance across software engineering, knowledge work, and vision tasks, with its advantage expanding as tasks grow longer and more complex. Anthropic is also releasing Claude Mythos 5, which is the same underlying model but with safeguards lifted in some areas, initially available to organizations in the private Project Glasswing initiative. Pricing is $10 per million input tokens and $50 per million output tokens—double Opus 4.8 but half the cost of Mythos Preview. Anthropic plans to expand access to Mythos 5 over time through a systematic trusted-access program.
- Claude Fable 5 is Anthropic's first Mythos-class model, previously deemed too risky for release due to cybersecurity concerns.
- New safeguards block high-risk responses (e.g., biology, cybersecurity) and fall back to Opus 4.8, but 95% of sessions stay on Fable.
- Pricing is $10/M input and $50/M output tokens, double Opus 4.8 but half of Mythos Preview; Mythos 5 offers unfiltered access to trusted users.
Why It Matters
Anthropic shows safe deployment of extremely capable AI is possible, setting a precedent for balancing capability and risk.