TECHNOLOGYFORTUNE
After backlash, Anthropic says its AI will now tell users when their request is being rejected or downgraded for national security concerns
Anthropic has updated its AI model Fable 5 to transparently inform users when requests are downgraded or rejected due to national security concerns, following backlash over prior silent downgrades. The company cited safety guardrails and industry-standard terms of service as reasons for restricting AI development use, while emphasizing transparency improvements to address researcher criticisms.
Mentioned
Related Signal
Adjacent reporting
- Anthropic: 'We made the wrong tradeoff' in new model guardrails
- Anthropic accused of ‘secret sabotage’ as Claude Fable 5 silently limits capabilities for AI researchers and developers
- Anthropic releases guardrailed version of Mythos for public use
- Anthropic's AI downgrade stings power users
- Anthropic’s new cybersecurity model could get it back in the government’s good graces