TECHNOLOGYTHE VERGE
Claude’s new model is more ‘honest’ when it messes up
Anthropic is releasing Claude Opus 4.8, which the company claims is more 'honest' by flagging uncertainties and avoiding unsupported claims. Early testers found it significantly less likely than its predecessor to make confident assertions without sufficient evidence.
Related Signal
Adjacent reporting
- Anthropic releases Claude Opus 4.7, concedes it trails unreleased Mythos
- OpenAI claims ChatGPT’s new default model hallucinates way less
- Anthropic releases a new Opus model amid Mythos Preview buzz
- AI will soon be capable of telling convincing lies
- Claude Opus 4.7 Is Here: Anthropic’s Latest Model Delivers, But It’s a Token Eating Machine
- ChatGPT's new default model is more factual and better at personalization