Media & Culture

Anthropic adds new guardrails to Claude Fable 5 after Trump admin export controls

⚡Model redirects sensitive queries to Opus 4.8 after Amazon paper exploit found

Deep Dive

Anthropic has resolved an impasse with the Trump administration over its Claude Fable 5 AI model by agreeing to expand security guardrails. The model was previously subject to export controls after the administration learned users could circumvent restrictions by asking Fable 5 to fix code instead of identifying security issues—a technique highlighted in a paper by Amazon. The new safeguard extends an existing barrier to any request related to that specific behavior, automatically redirecting such queries to the less-capable Opus 4.8 model and notifying users the request was blocked.

Commerce Secretary Howard Lutnick confirmed the removal of restrictions in a letter, noting Anthropic committed to proactively detecting and addressing security risks. While the commerce department is satisfied, defense secretary Pete Hegseth still considers Anthropic a supply chain risk under a February 28 order, leaving the company's full reinstatement incomplete. The incident underscores the delicate relationship between frontier AI companies and national security regulations.

Key Points
  • Anthropic extended guardrails on Claude Fable 5 after users exploited a fix-code loophole from an Amazon paper
  • Restricted requests are now redirected to Opus 4.8 with a block notification
  • Commerce Dept lifted export controls, but Defense Dept still labels Anthropic a supply chain risk

Why It Matters

Regulatory divide between commerce and defense creates ongoing uncertainty for AI companies working with sensitive capabilities.

📬 Get the top 10 AI stories daily