arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

多模态大模型

跨文本、图像、视频、音频等模态的大模型与学习方法。

2025-08-28 至 2025-08-28 共收录 7 信号源:cs.CV, cs.CL, cs.AI, cs.MM, eess.AS

1. 多模态生成 7 篇

2502.09242 2025-08-28 cs.AI 79%

From large language models to multimodal AI: A scoping review on the potential of generative AI in medicine

Lukas Buess, Matthias Keicher, Nassir Navab, Andreas Maier, Soroosh Tayebi Arasteh

专题命中 多模态生成 :multimodal(title,abstract);分类 cs.AI

Journal ref Biomed. Eng. Lett. 15 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.20020 2025-08-28 cs.CV 57%

GS: Generative Segmentation via Label Diffusion

Yuhao Chen, Shubin Chen, Liang Lin, Guangrun Wang

专题命中 多模态生成 :multimodal(abstract);分类 cs.CV

Comments 12 pages, 7 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.19508 2025-08-28 cs.RO cs.CV 57%

DATR: Diffusion-based 3D Apple Tree Reconstruction Framework with Sparse-View

Tian Qiu, Alan Zoubi, Yiyuan Lin, Ruiming Du, Lailiang Cheng, Yu Jiang

机构 * School of Electrical and Computer Engineering, Cornell University(电气与计算机工程系,康奈尔大学) Sibley School of Mechanical and Aerospace Engineering, Cornell University(机械与航空航天工程系,康奈尔大学) School of Biological and Environmental Engineering, Cornell University(生物与环境工程系,康奈尔大学) School of Integrative Plant Science, Cornell University(整合植物科学系,康奈尔大学)

专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.15842 2025-08-28 cs.CV cs.GR 57%

DiffArtist: Towards Structure and Appearance Controllable Image Stylization

Ruixiang Jiang, Changwen Chen

机构 * The Hong Kong Polytechnic University(香港理工大学)

专题命中 多模态生成 :multimodal(abstract);分类 cs.CV

Comments Accepted to ACM MM 2025, Homepage: https://DiffusionArtist.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.14874 2025-08-28 cs.CV 57%

TraceNet: Segment one thing efficiently

Mingyuan Wu, Zichuan Liu, Haozhen Zheng, Hongpeng Guo, Bo Chen, Xin Lu, Klara Nahrstedt

机构 * Coordinated Science Laboratory, University of Illinois at Urbana-Champaign, Champaign, USA(伊利诺伊大学厄巴纳-香槟分校协调科学实验室)

专题命中 多模态生成 :multimodal(abstract);分类 cs.CV

Comments Best Student Paper in IEEE MIPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.12166 2025-08-28 cs.RO cs.LG cs.SY eess.SY 50%

Belief-Conditioned One-Step Diffusion: Real-Time Trajectory Planning with Just-Enough Sensing

Gokul Puthumanaillam, Aditya Penumarti, Manav Vora, Paulo Padrao, Jose Fuentes, Leonardo Bobadilla, Jane Shin, Melkior Ornik

机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) University of Florida(佛罗里达大学) Providence College(普罗维登斯学院) Florida International University(佛罗里达国际大学)

专题命中 多模态生成 :multi-modal(abstract)

Comments Accepted to CoRL 2025 (Conference on Robot Learning)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.14207 2025-08-28 cs.RO 50%

A Comprehensive Review on Traffic Datasets and Simulators for Autonomous Vehicles

Supriya Sarker, Brent Maples, Iftekharul Islam, Muyang Fan, Christos Papadopoulos, Weizi Li

专题命中 多模态生成 :multimodal(abstract)

Comments This manuscript has been withdrawn due to the need for substantial updates and revisions

详情

展开后加载摘要…

URL PDF HTML 收藏