Utilizing Vision-Language Models as Action Models for Intent Recognition and Assistance
机构 * University of Birmingham(伯明翰大学) ; Queen Mary University of London(伦敦大学女王学院) ; Aston University(阿斯顿大学) ; United Kingdom National Nuclear Laboratory Ltd.(英国国家核实验室有限公司)
专题命中 GUI与屏幕智能体 :vision-language model(title,abstract);VLM(abstract);分类 cs.AI
Comments Accepted at Human-Centered Robot Autonomy for Human-Robot Teams (HuRoboT) at IEEE RO-MAN 2025, Eindhoven, the Netherlands