Media & Culture

Anthropic adds new guardrails to Claude Fable 5 after Trump admin export controls

Model redirects sensitive queries to Opus 4.8 after Amazon paper exploit found

Deep Dive

Anthropic has resolved an impasse with the Trump administration over its Claude Fable 5 AI model by agreeing to expand security guardrails. The model was previously subject to export controls after the administration learned users could circumvent restrictions by asking Fable 5 to fix code instead of identifying security issues—a technique highlighted in a paper by Amazon. The new safeguard extends an existing barrier to any request related to that specific behavior, automatically redirecting such queries to the less-capable Opus 4.8 model and notifying users the request was blocked.

Commerce Secretary Howard Lutnick confirmed the removal of restrictions in a letter, noting Anthropic committed to proactively detecting and addressing security risks. While the commerce department is satisfied, defense secretary Pete Hegseth still considers Anthropic a supply chain risk under a February 28 order, leaving the company's full reinstatement incomplete. The incident underscores the delicate relationship between frontier AI companies and national security regulations.

Key Points
  • Anthropic extended guardrails on Claude Fable 5 after users exploited a fix-code loophole from an Amazon paper
  • Restricted requests are now redirected to Opus 4.8 with a block notification
  • Commerce Dept lifted export controls, but Defense Dept still labels Anthropic a supply chain risk

Why It Matters

Regulatory divide between commerce and defense creates ongoing uncertainty for AI companies working with sensitive capabilities.

📬 Get the top 10 AI stories daily