arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

多模态大模型

跨文本、图像、视频、音频等模态的大模型与学习方法。

2025-08-04 至 2025-08-04 共收录 3 信号源:cs.CV, cs.CL, cs.AI, cs.MM, eess.AS

1. 多模态生成 3 篇

2508.00303 2025-08-04 cs.RO 78%

TopoDiffuser: A Diffusion-Based Multimodal Trajectory Prediction Model with Topometric Maps

Zehui Xu, Junhui Wang, Yongliang Shi, Chao Gao, Guyue Zhou

机构 * School of Astronautics, Harbin Institute of Technology(哈尔滨工业大学航天学院) Institute for AI Industry Research (AIR), Tsinghua University(清华大学人工智能产业研究院) Institute of Systems Engineering and Collaborative Laboratory for Intelligent Science and Systems, Macau University of Science and Technology(澳门科学大学系统工程研究所) School of Vehicle and Mobility, Tsinghua University(清华大学车辆与移动系统学院)

专题命中 多模态生成 :multimodal(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.18004 2025-08-04 cs.AI 57%

E.A.R.T.H.: Structuring Creative Evolution through Model Error in Generative AI

Yusen Peng, Shuhua Mao

机构 * University of Warwick(沃里克大学) Wuhan University of Technology(武汉理工大学)

专题命中 多模态生成 :cross-modal(abstract);分类 cs.AI

Comments 44 pages,11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.00583 2025-08-04 cs.NI 50%

Enhancing Wireless Networks for IoT with Large Vision Models: Foundations and Applications

Yunting Xu, Jiacheng Wang, Ruichen Zhang, Dusit Niyato, Deepu Rajan, Liang Yu, Haibo Zhou, Abbas Jamalipour, Xianbin Wang

专题命中 多模态生成 :multimodal(abstract)

Comments 7 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏