Part-X-MLLM: Part-aware 3D Multimodal Large Language Model
机构 * Zhejiang University(浙江大学) ; Tencent Hunyuan(腾讯文言) ; Tsinghua University(清华大学) ; The University of Hong Kong(香港大学)
专题命中 多模态生成 :multimodal(title,abstract);MLLM(title,abstract);分类 cs.CV
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
机构 * Zhejiang University(浙江大学) ; Tencent Hunyuan(腾讯文言) ; Tsinghua University(清华大学) ; The University of Hong Kong(香港大学)
专题命中 多模态生成 :multimodal(title,abstract);MLLM(title,abstract);分类 cs.CV
专题命中 多模态生成 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CV、cs.AI
Comments Accepted at AAAI'26
机构 * Computer Engineering, Texas A\&M University 3 Texas A\&M University Emergency Medical Services (EMS), 4 Biomedical Data Science, Stanford University, 5 Texas A\&M School of Public Health
专题命中 多模态生成 :multimodal(title,abstract);分类 eess.AS
机构 * Verily Life Sciences(Verily生命科学公司) ; Nvidia(英伟达公司) ; Google(谷歌公司)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.AI
机构 * Project leader(项目负责人)
专题命中 多模态生成 :MLLM(abstract,comments);multimodal(abstract);分类 cs.CV、cs.CL
Comments Accepted by AAAI Conference on Artificial Intelligence (AAAI) 2026. Code available at https://github.com/bcmi/D3ToM-Diffusion-MLLM
机构 * Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) ; Beihang University(北航大学)
专题命中 多模态生成 :multimodal(abstract);cross-modal(abstract);分类 cs.AI、cs.MM
机构 * The University of Texas at Dallas(德克萨斯大学达拉斯分校) ; University of Toronto(多伦多大学) ; University of Notre Dame(诺特大学) ; Stony Brook University(石溪大学)
专题命中 多模态生成 :multimodal(abstract);MLLM(abstract);分类 cs.CV
机构 * Dept. of Artificial Intelligence Korea University Seoul, Republic of Korea(人工智能系韩国大学首尔共和国韩国)
专题命中 多模态生成 :multimodal(abstract);分类 cs.CL、cs.AI
专题命中 多模态生成 :audio-visual(abstract);分类 cs.CV、eess.AS
Comments Project page: https://readportrait.github.io/READ/
机构 * University of California, Berkeley(加州大学伯克利分校) ; Adobe Research(Adobe研究)
专题命中 多模态生成 :cross-modal(abstract);分类 cs.AI、cs.MM
Comments Accepted to TMLR
机构 * Wangxuan Institute of Computer Technology, Peking University(王轩计算机技术研究所,北京大学)
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV、cs.AI
Comments Accepted conference paper of ACM MM 2024
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
Comments AAAI2026
机构 * University of Southern California(南加州大学) ; Intel Labs(英特尔实验室)
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
Comments AAAI2026-Oral. Project Page: https://xinyuanhu66.github.io/SRSplat/
机构 * Department of Electrical and Computer Engineering, University of Waterloo(滑铁卢大学电气与计算机工程系)
专题命中 多模态生成 :multimodal(abstract)
Comments Initial results presented at the IJCAI 2025 Workshop on User-Aligned Assessment of Adaptive AI Systems. Project page: https://aku02.github.io/projects/difffp/