MIDI-LLaMA: An Instruction-Following Multimodal LLM for Symbolic Music Understanding
MIDI-LLaMA:一种用于符号音乐理解的指令遵循多模态大语言模型
Meng Yang, Jon McCormack, Maria Teresa Llano, Wanchao Su, Chao Lei
机构
*
SensiLab, Monash University, Australia(Monash大学)
;
University of Sussex, Brighton, United Kingdom(Sussex大学)
;
School of Computing and Information Systems, The University of Melbourne, Australia(墨尔本大学计算机与信息系统学院)
机构
*
Graduate Institute of Electronics Engineering(电子工程研究所)
;
the Department of Electrical Engineering, National Taiwan University, Taipei 106319, Taiwan(国立台湾大学电子工程系)
;
Research Center for Information Technology Innovation Technology, Academia Sinica, Taipei 115201, Taiwan(中科院资讯科技创新研究中心)
;
Department of Electrical Engineering, Chung Yuan Christian University, Taoyuan 320314, Taiwan(Chung Yuan Christian University 电子工程系)
机构
*
SuZhou Automotive Research Institute of Tsinghua University(清华大学苏州汽车研究院)
;
Department of Electrical and Electronic Engineering, The University of Hong Kong(香港大学电子与电气工程系)
;
Hyundai Motor Advanced Tech. R&D Center
;
School of Vehicle and Mobility, Tsinghua University(清华大学车辆与移动系统学院)
机构
*
Department of Computer Science, University of Exeter(埃克塞特大学计算机科学系)
;
Institute for Analytics and Data Science, University of Essex(埃塞克斯大学分析与数据科学研究所)
;
School of Computer Science, University of Birmingham(伯明翰大学计算机科学学院)
Understanding Multimodal Complementarity for Single-Frame Action Anticipation
理解单帧动作预测中的多模态互补性
Manuel Benavent-Lledo, Konstantinos Bacharidis, Konstantinos Papoutsakis, Antonis Argyros, Jose Garcia-Rodriguez
机构
*
Department of Computer Technology, University of Alicante(阿尔瓦雷斯大学计算机技术系)
;
Institute of Computer Science, FORTH(福蒂研究所)
;
Computer Science Department, University of Crete(克里特大学计算机科学系)
;
Department of Management, Science and Technology, Hellenic Mediterranean University(希腊地中海大学管理、科学与技术系)
A Three-Level Alignment Framework for Large-Scale 3D Retrieval and Controlled 4D Generation
一种用于大规模3D检索和受控4D生成的三级对齐框架
Philip Xu
专题命中
跨模态检索
:multimodal(abstract);分类 cs.CV
AI总结
Uni4D通过三级对齐框架实现大规模3D检索与可控4D生成,提升多模态动态理解与应用
CommentsarXiv admin note: Author list truncated. This submission has been withdrawn by arXiv administrators as authors were added without their knowledge or consent
机构
*
Organizational Management Department, School of Management, Xi’an Jiaotong University(管理学院组织管理部,西安交通大学)
;
West China Longquan Hospital, Sichuan University(四川大学西部临床医学院)
;
School of Electronic Science and Engineering, Xi’an Jiaotong University(西安交通大学电子科学与工程学院)
;
Systems Engineering Institute, Xi’an Jiaotong University(西安交通大学系统工程研究院)
;
Institute of Medical Artificial Intelligence, the Second Affiliated Hospital of Xi’an Jiaotong University(西安交通大学第二附属医院医学人工智能研究所)
;
School of Human Settlements and Civil Engineering, Xi’an Jiaotong University(西安交通大学人居环境与土木工程学院)
;
School of Life Science and Technology, Xi’an Jiaotong University(西安交通大学生命科学与技术学院)
Multi-modal Imputation for Alzheimer's Disease Classification
多模态缺失数据填补用于阿尔茨海默病分类
Abhijith Shaji, Tamoghna Chattopadhyay, Sophia I. Thomopoulos, Greg Ver Steeg, Paul M. Thompson, Jose-Luis Ambite
机构
*
Information Sciences Institute(信息科学研究所)
;
University of Southern California(美国南加州大学)
;
University of California(加州大学)
;
Stevens Neuroimaging and Informatics Institute(史蒂文斯神经影像与信息学研究所)
机构
*
HKUST (GZ)(香港科技大学)
;
Tongji University(同济大学)
;
University of Rochester(罗切斯特大学)
;
Meta
;
Fudan University(复旦大学)
;
The University of British Columbia(不列颠哥伦比亚大学)
;
University of California, Davis(加州大学戴维斯分校)
Guangping Liu, Tipu Sultan, Vittorio Di Giorgio, Nick Hawkins, Flavio Esposito, Madi Babaiasl
机构
*
Aerospace and Mechanical Engineering Department, Saint Louis University(圣路易斯大学航空航天与机械工程系)
;
Computer Science Department, Saint Louis University(圣路易斯大学计算机科学系)