arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

多模态大模型

跨文本、图像、视频、音频等模态的大模型与学习方法。

2025-11-04 至 2025-11-04 共收录 9 信号源:cs.CV, cs.CL, cs.AI, cs.MM, eess.AS

1. 其他多模态 9 篇

2505.23118 2025-11-04 cs.CL cs.AI 81%

Elicit and Enhance: Advancing Multimodal Reasoning in Medical Scenarios

Zhongzhen Huang, Linjie Mu, Yakun Zhu, Xiangyu Zhao, Shaoting Zhang, Xiaofan Zhang

机构 * Shanghai Jiao Tong University(上海交通大学)

专题命中 其他多模态 :multimodal(title,abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00997 2025-11-04 cs.CV 79%

MID: A Self-supervised Multimodal Iterative Denoising Framework

Chang Nie, Tianchen Deng, Zhe Liu, Hesheng Wang

机构 * School of Automation and Intelligent Sensing, Shanghai Jiao Tong University(自动化与智能感知学院,上海交通大学) Key Laboratory of System Control and Information Processing, Ministry of Education of China(系统控制与信息处理重点实验室,中华人民共和国教育部)

专题命中 其他多模态 :multimodal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.11807 2025-11-04 cs.CL 79%

Are Multimodal Large Language Models Pragmatically Competent Listeners in Simple Reference Resolution Tasks?

Simeon Junker, Manar Ali, Larissa Koch, Sina Zarrieß, Hendrik Buschmeier

机构 * Bielefeld University(比勒菲尔德大学)

专题命中 其他多模态 :multimodal(title,abstract);分类 cs.CL

Comments To appear in ACL Findings 2025

Journal ref Findings of the Association for Computational Linguistics: ACL 2025, pp. 24101-24109

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.25381 2025-11-04 cs.HC 78%

CGM-Led Multimodal Tracking with Chatbot Support: An Autoethnography in Sub-Health

Dongyijie Primo Pan, Lan Luo, Yike Wang, Pan Hui

专题命中 其他多模态 :multimodal(title,abstract)

Comments International Conference on Human-Engaged Computing (ICHEC 2025), Singapore

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00411 2025-11-04 cs.LG cs.AI cs.CV 62%

Enhancing Adversarial Transferability by Balancing Exploration and Exploitation with Gradient-Guided Sampling

Zenghao Niu, Weicheng Xie, Siyang Song, Zitong Yu, Feng Liu, Linlin Shen

机构 * School of Computer Science & Software Engineering, Shenzhen University, China(深圳大学计算机科学与软件工程学院) Guangdong Laboratory of Artificial Intelligence and Digital Economy (SZ), Shenzhen, China(广东省人工智能与数字经济发展实验室(深圳)) Guangdong Provincial Key Laboratory of Intelligent Information Processing, Shenzhen University, China(广东省智能信息处理省级重点实验室) School of Computer Science, University of Exeter, U.K.(埃克塞特大学计算机科学学院) Department of Computing and Information Technology, Great Bay University, China(大鹏大学计算与信息技术系) Computer Vision Institute, School of Artificial Intelligence, Shenzhen University, China(人工智能学院计算机视觉研究所)

专题命中 其他多模态 :multimodal(abstract);分类 cs.CV、cs.AI

Comments accepted by iccv 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01379 2025-11-04 cs.RO 50%

CM-LIUW-Odometry: Robust and High-Precision LiDAR-Inertial-UWB-Wheel Odometry for Extreme Degradation Coal Mine Tunnels

Kun Hu, Menggang Li, Zhiwen Jin, Chaoquan Tang, Eryi Hu, Gongbo Zhou

机构 * School of Mechatronic Engineering, China University of Mining and Technology(机械电子工程学院,中国矿业大学)

专题命中 其他多模态 :multimodal(abstract)

Comments Accepted by IROS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.19660 2025-11-04 cs.ET q-bio.BM 50%

Machine Olfaction and Embedded AI Are Shaping the New Global Sensing Industry

Andreas Mershin, Nikolas Stefanou, Adan Rotteveel, Matthew Kung, George Kung, Alexandru Dan, Howard Kivell, Zoia Okulova, Zoi Kountouri, Paul Pu Liang

专题命中 其他多模态 :multimodal(abstract)

Comments 23 pages, 116 citations, combination tech review/industry roadmap/white paper on the rise of machine olfaction as an essential AI modality

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13066 2025-11-04 astro-ph.HE 50%

Double-Peaked Optical Afterglow in GRB 110213A Inferring a Magnetized Thick Shell Ejecta

Yo Kusafuka, Kaori Obayashi, Katsuaki Asano, Ryo Yamazaki

专题命中 其他多模态 :multimodal(abstract)

Comments 9 pages, 5 figures, accepted for publication in MNRAS, Magglow is available from https://github.com/yo3-sun/Magglow

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00309 2025-11-04 math.OC 50%

Transit-MP: Transit-Prioritized Max-Pressure Control in Sparse Connected Vehicle Environments

Chaopeng Tan, Hao Liu, Dingshan Sun, Marco Rinaldi, Hans van Lint

专题命中 其他多模态 :multi-modal(abstract)

Comments 36 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏