Unified Visuomotor Targets: Supervising VLAs Beyond Physical Actions
统一视觉运动目标:监督视觉语言智能体(VLA)超越物理动作
AI总结 本研究针对VLA模型训练中视觉语言高级表示与机器人低级动作的不匹配问题,提出UVT统一视觉运动目标,无需改动架构或额外数据,在仿真与真实双臂任务中提升了VLA的训练效率、性能与鲁棒性
Comments Accepted at IROS 2026. Project page: https://unified-visuomotor-targets.github.io/