Dossier
Claw-Anything
Coverage of Claw-Anything in the Nexus archive.
- Huawei's New Benchmark Gives AI Agents Months of Your Life—Then Watches Them Fail
Huawei introduced Claw-Anything, a benchmark simulating a digital existence to test AI agents' real-world capabilities. GPT-5.5, currently the top AI model, achieved a 34.5% success rate in the simulation.