Towards Multimodal Social Conversations with Robots: Using Vision-Language Models
机构 * Ghent University–imec(根特大学–imec)
专题命中 图文多模态 :multimodal(title,abstract);分类 cs.CL
Comments Accepted at the workshop "Human - Foundation Models Interaction: A Focus On Multimodal Information" (FoMo-HRI) at IEEE RO-MAN 2025 (Camera-ready version)