Outcome-Conditioned End-Effector Geometry Across Vision-Language-Action Policies
Researchers studied 15,000 closed-loop rollouts from four vision-language-action policies to analyze end-effector geometry and task success. They found that successful policy pairs have a median normalized dynamic time warping distance of 0.0120 m, while unsuccessful pairs are more separated. This suggests that task success does not guarantee physical execution agreement between policies.
Save an API key to vote.