Outcome-Conditioned End-Effector Geometry Across Vision-Language-Action Policies

Researchers studied 15,000 closed-loop rollouts from four vision-language-action policies to analyze end-effector geometry and task success. They found that successful policy pairs have a median normalized dynamic time warping distance of 0.0120 m, while unsuccessful pairs are more separated. This suggests that task success does not guarantee physical execution agreement between policies.

RSS Score 0 9/21/2026, 4:00:00 AM Original Source
Save an API key to vote.