机构
*
Gaoling School of Artificial Intelligence, Renmin University of China(中国人民大学北京校区人工智能学院)
;
Xiamen University(厦门大学)
;
The Hong Kong University of Science and Technology(香港理工大学)
;
Nanyang Technological University(南洋理工大学)
MetaCaptioner: Towards Generalist Visual Captioning with Open-source Suites
Zhenxin Lei, Zhangwei Gao, Changyao Tian, Erfei Cui, Guanzhou Chen, Danni Yang, Yuchen Duan, Zhaokai Wang, Wenhao Li, Weiyun Wang, Xiangyu Zhao, Jiayi Ji, Yu Qiao, Wenhai Wang, Gen Luo
机构
*
Shanghai AI Laboratory(上海人工智能实验室)
;
Shanghai Jiao Tong University(上海交通大学)
;
Fudan University(复旦大学)
;
University of Chinese Academy of Science(中国科学院大学)
;
Xiamen University(厦门大学)
;
The Chinese University of Hong Kong(香港中文大学)
Generalist vs Specialist Time Series Foundation Models: Investigating Potential Emergent Behaviors in Assessing Human Health Using PPG Signals
Saurabh Kataria, Yi Wu, Zhaoliang Chen, Hyunjung Gloria Kwak, Yuhao Xu, Lovely Yeswanth Panchumarthi, Ran Xiao, Jiaying Lu, Ayca Ermis, Anni Zhao, Runze Yan, Alex Federov, Zewen Liu, Xu Wu, Wei Jin, Carl Yang, Jocelyn Grunwell, Stephanie R. Brown, Amit Shah, Craig Jabaley, Tim Buchman, Sivasubramanium V Bhavani, Randall J. Lee, Xiao Hu
机构
*
Nell Hodgson Woodruff School of Nursing(Nell Hodgson Woodruff护理学院)
;
School of Computer Science(计算机科学学院)
;
Department of Pediatrics(儿科系)
;
Department of Computer Science(计算机科学系)
;
Department of Epidemiology(流行病学系)
;
Department of Anesthesiology(麻醉学系)
;
Department of Surgery(外科系)
;
Department of Medicine(医学系)
;
School of Medicine(医学院)
NExT-OMNI: Towards Any-to-Any Omnimodal Foundation Models with Discrete Flow Matching
Run Luo, Xiaobo Xia, Lu Wang, Longze Chen, Renke Shan, Jing Luo, Min Yang, Tat-Seng Chua
机构
*
Shenzhen Institutes of Advanced Technology, Chinese Academy of Sciences(中国科学院深圳先进技术研究院)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
NExT++ Research Center(NExT++研究中心)
;
National University of Singapore(新加坡国立大学)
专题命中
多模态生成
:any-to-any(title,abstract);multimodal(abstract);cross-modal(abstract);multimodal foundation model(abstract)
Zekun Wang, King Zhu, Chunpu Xu, Wangchunshu Zhou, Jiaheng Liu, Yibo Zhang, Jiashuo Wang, Ning Shi, Siyu Li, Yizhi Li, Haoran Que, Zhaoxiang Zhang, Yuanxing Zhang, Ge Zhang, Ke Xu, Jie Fu, Wenhao Huang
机构
*
Beihang University(北航)
;
M-A-P
;
The Hong Kong Polytechnic University(香港理工大学)
;
AIWaves
;
University of Alberta(阿尔伯塔大学)
;
University of Waterloo(滑铁卢大学)
;
University of Manchester(曼彻斯特大学)
;
Chinese Academy of Sciences(中国科学院)
;
Peking University(北京大学)
;
Shanghai AI Lab(上海AI实验室)
;
Nanjing University(南京大学)
;
Kuaishou Technology(快手科技)
CommentsFor a published version refer to the Information Fusion, DOI. 10.1016/j.inffus.2025.103572
Journal refSui, W., Lichau, D., Lefèvre, J., & Phelippeau, H. (2026). Incomplete multimodal industrial anomaly detection via cross-modal distillation. Information Fusion, 126, Article 103572