Machine Learning Algorithms & Natural Language Processing
Jul 9, 2026 · Artificial Intelligence
How Attending Before Acting Boosts Generalization in Pelican-VLA 0.5
The talk presents Pelican-VLA 0.5, a unified Vision‑Language‑Action model that leverages attention‑level generalization without task‑specific supervision, achieving over 91% success on RoboTwin benchmarks and demonstrating early zero‑shot generalization through a novel Reasoning Slots bottleneck.
Attention GeneralizationPelican-VLAVision-Language-Action
0 likes · 6 min read
