Machine Heart
Sep 4, 2026 · Artificial Intelligence
HumanCLAW Benchmark Shows VLMs Achieve Only 16.8% Success in Embodied Action Tasks
Meta's HumanCLAW benchmark evaluates nine vision-language models on embodied action intelligence, separating high-level decisions from low-level control; the best model completes full interactions at just 16.8% success, revealing critical gaps in embodied self-awareness and closed-loop reasoning.
Action IntelligenceHumanCLAWMeta
0 likes · 12 min read
