TECHNOLOGYSEMAFOR
Anthropic releases guardrailed version of Mythos for public use
Anthropic released a guardrailed version of its Mythos model called Fable 5, which blocks answers to cybersecurity and biology-related questions to ensure public safety. The company claims its safeguards withstood hacker testing, though concerns remain about potential jailbreaking attempts. Anthropic also upgraded Mythos 5 for select customers, highlighting its strong cybersecurity capabilities and adjusted pricing.
Mentioned
Related Signal
Adjacent reporting
- OpenAI expands access to cyber AI as hacking risks grow
- Anthropic limits access to Mythos, its new cybersecurity AI model
- AI guardrails stripped from Meta and Google models in minutes
- OpenAI rolls out new model for cybersecurity teams a month after Anthropic's Mythos debut
- OpenAI makes its rival to Anthropic's Mythos more widely available to cyber defenders