arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

多模态大模型

跨文本、图像、视频、音频等模态的大模型与学习方法。

2025-08-06 至 2025-08-06 共收录 64 信号源:cs.CV, cs.CL, cs.AI, cs.MM, eess.AS

1. 其他多模态 6 篇

2508.02982 2025-08-06 cs.RO 71%

Multimodal Human-Intent Modeling for Contextual Robot-to-Human Handovers of Arbitrary Objects

Lucas Chen, Guna Avula, Hanwen Ren, Zixing Wang, Ahmed H. Qureshi

机构 * Departmant of Computer Science, Purdue University(计算机科学系,普渡大学)

专题命中 其他多模态 :multimodal(title)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.03094 2025-08-06 cs.CV 70%

Augmenting Continual Learning of Diseases with LLM-Generated Visual Concepts

Jiantao Tan, Peixian Ma, Kanghao Chen, Zhiming Dai, Ruixuan Wang

机构 * Sun Yat-sen University(中山大学) The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) Peng Cheng Laboratory(鹏城实验室) Key Laboratory of Machine Intelligence and Advanced Computing, MOE(教育部机器智能与先进计算重点实验室)

专题命中 其他多模态 :multimodal(abstract);cross-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.03562 2025-08-06 cs.CV cs.CL 62%

Beyond Meme Templates: Limitations of Visual Similarity Measures in Meme Matching

Muzhaffar Hazman, Susan McKeever, Josephine Griffith

机构 * School of Computer Science University of Galway(计算机科学学院 Galway大学) School of Computer Science Technological University Dublin(计算机科学学院 技术大学都柏林)

专题命中 其他多模态 :multimodal(abstract);分类 cs.CV、cs.CL

Comments Accepted for publication at IEEE International Conference on Image Processing Theory, Tools and Applications (IPTA) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.13730 2025-08-06 cs.NE 50%

Cascading CMA-ES Instances for Generating Input-diverse Solution Batches

Maria Laura Santoni, Christoph Dürr, Carola Doerr, Mike Preuss, Elena Raponi

专题命中 其他多模态 :multi-modal(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏