HumanCLAW: Vision-Language Models Lose Track of Their Own Bodies

HumanCLAW tests whether vision-language models can control a simulated human body without motor errors muddying the result. The best of nine models completed only 16.8% of 1,218 tasks.
artificial-intelligence
Author

Kabui, Charles

Published

2026-07-31

Keywords

humanclaw, embodied-ai, vision-language-models, robot-evaluation, body-awareness