UAD: Unsupervised Affordance Distillation for Generalization in Robotic Manipulation
机构 * Stanford University(斯坦福大学)
专题命中 其他VLM :vision-language model(abstract);分类 cs.CV、cs.AI
AI 大模型
视觉语言模型、视觉推理、视觉问答、图文理解和视觉 grounding。
机构 * Stanford University(斯坦福大学)
专题命中 其他VLM :vision-language model(abstract);分类 cs.CV、cs.AI
机构 * State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences(多模态人工智能系统国家重点实验室,自动化研究所,中国科学院) ; School of Artificial Intelligence, University of Chinese Academy of Sciences(人工智能学院,中国科学院大学) ; Department of Physics and Technology, UiT The Arctic University of Norway(物理与技术系,UiT 北极大学)
专题命中 其他VLM :vision-language model(abstract);分类 cs.CV
Comments Accepted by EMNLP 2025 Findings