TECHNOLOGYHACKER NEWS
Five frontier LLMs disagree on 67% of 1k real-world fact-check claims
A study by Lenz.io found that five leading large language models (LLMs) disagreed on 67% of 1,000 real-world fact-check claims, highlighting limitations in their consensus. The findings were discussed on Hacker News, with 66 points and 29 comments.
Mentioned
Related Signal
Adjacent reporting
- Evaluating large language models for accuracy incentivizes hallucinations
- I’m a Professional Fact-Checker. AI Is Wrong More Often Than You Think
- Friendlier LLMs tell users what they want to hear — even when it is wrong
- Prompt Politeness Affects LLM Accuracy (2025)
- Study of 1,700 languages reveals surprising hidden patterns
- See through local AI lies with Irish eyes