FineVLA: Fine-Grained Instruction Alignment for Steerable Vision-Language-Action Policies
FineVLA, an open framework for action-aligned fine-grained VLA supervision, is introduced and the largest real-world gains appear on pose, color, and approach direction, and approach direction--factors where goal-level instructions provide no guidance.