Skip to main content
Back to News Hub
Wired AI
June 11, 2026
Society & Culture

Anthropic Walks Back Policy That Could Have 'Sabotaged' AI Researchers Using Claude

Overview

Anthropic reversed a hidden safeguard in its new Claude Fable 5 model that quietly weakened the assistant when people asked it to help build competing AI systems. The restriction, disclosed in a single paragraph inside a 319-page system card, rerouted certain frontier research requests to a weaker model without telling the user. After AI researchers and policy analysts criticized the move within hours of the June 9, 2026 release, the company apologized and said it will make any such limits visible rather than silent.

Stats & Key Facts

  • #The restriction, disclosed in a single paragraph inside a 319-page system card, rerouted certain frontier research requests to a weaker model without telling the user.

Continue Learning

Comments

Comments appear only after moderation. Your email identifies your submission to the moderator and is never displayed here.

No approved comments yet.

Originally published by Wired AI
Read the original