机构
*
Northeastern University(东北大学)
;
Microsoft Research(微软研究院)
;
University of Southern California(南加州大学)
;
University of California, Santa Cruz(加州大学圣克鲁兹分校)
Hulu-Med: A Transparent Generalist Model towards Holistic Medical Vision-Language Understanding
Songtao Jiang, Yuan Wang, Sibo Song, Tianxiang Hu, Chenyi Zhou, Bin Pu, Yan Zhang, Zhibo Yang, Yang Feng, Joey Tianyi Zhou, Jin Hao, Zijian Chen, Ruijia Wu, Tao Tang, Junhui Lv, Hongxia Xu, Hongwei Wang, Jun Xiao, Bin Feng, Fudong Zhu, Kenli Li, Weidi Xie, Jimeng Sun, Jian Wu, Zuozhu Liu
机构
*
College of Computer Science and Technology, Zhejiang University-University of Illinois Urbana-Champaign Institute(浙江大学计算机科学与技术学院)
;
Stomatology Hospital, School of Stomatology, Zhejiang University School of Medicine(浙江大学口腔医院)
;
Alibaba Inc(阿里巴巴集团)
;
College of Computer Science and Electronic Engineering, Hunan University(湖南大学计算机科学与电子工程学院)
;
Angelalign Technology Inc.(Angelalign技术有限公司)
;
CFAR & IHPC, Agency for Science, Technology and Research(CFAR与IHPC,新加坡科技研究局)
;
Department of Orthodontics, Shanghai Ninth People’s Hospital, College of Stomatology, Shanghai Jiao Tong University(上海第九人民医院正畸科,上海交通大学口腔医学院)
Unifying Symbolic Music Arrangement: Track-Aware Reconstruction and Structured Tokenization
Longshen Ou, Jingwei Zhao, Ziyu Wang, Gus Xia, Qihao Liang, Torin Hopkins Ye Wang
机构
*
Sound and Music Computing Lab, School of Computing, NUS(新加坡国立大学计算机学院声音与音乐计算实验室)
;
Courant Institute of Mathematical Sciences, New York University(纽约大学应用数学科学学院)
;
Music X Lab, MBZUAI(MBZUAI音乐X实验室)
Breaking the Encoder Barrier for Seamless Video-Language Understanding
Handong Li, Yiyuan Zhang, Longteng Guo, Xiangyu Yue, Jing Liu
机构
*
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
MMLab, CUHK(CUHK MMLab)
;
Institute of Automation, Chinese Academy of Science(中国科学院自动化研究所)
;
Shanghai AI Lab(上海人工智能实验室)
专题命中
视频多模态
:multimodal(abstract);分类 cs.CV
Comments12 pages
Journal refProceedings of the IEEE/CVF International Conference on Computer Vision (ICCV), 2025, pp. 23167-23176
机构
*
The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州))
;
University of Oxford(牛津大学)
;
Xi’an Jiaotong University(西安交通大学)
;
The Hong Kong University of Science and Technology(香港科学与技术大学)
;
City University of Hong Kong(香港城市大学)
;
Beijing University of Technology(北京理工大学)
;
Duke University(杜克大学)
;
X-Humanoid Project(X-Humanoid 项目)
Using Multi-modal Large Language Model to Boost Fireworks Algorithm's Ability in Settling Challenging Optimization Tasks
Shipeng Cen, Ying Tan
机构
*
School of Intelligence Science
;
Technology, Institute for Artificial Intellignce, Peking University, Beijing, China
;
Technology, Institute for Artificial Intellignce, National Key Laboratory of General Artificial Intelligence, Peking University, Beijing, China
DiffSpectra: Molecular Structure Elucidation from Spectra using Diffusion Models
Liang Wang, Yu Rong, Tingyang Xu, Zhenyi Zhong, Zhiyuan Liu, Pengju Wang, Deli Zhao, Qiang Liu, Shu Wu, Liang Wang, Yang Zhang
机构
*
NLPR, MAIS, Institute of Automation(NLPR、MAIS、自动化研究所)
;
School of Artificial Intelligence(人工智能学院)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Alibaba Group(阿里巴巴集团)
;
Hupan Lab(华普实验室)
;
College of Intelligence and Computing(智能与计算学院)
;
National University of Singapore(新加坡国立大学)
;
Cancer Science Institute of Singapore(新加坡癌症科学研究所)
机构
*
School of Computer Science and Technology, Shandong University(山东大学计算机科学与技术学院)
;
Institute of Artificial Intelligence, Beihang University(北京航空航天大学人工智能研究院)
Individualizing Glioma Radiotherapy Planning by Optimization of Data and Physics-Informed Discrete Loss
Michal Balcerak, Jonas Weidner, Petr Karnakov, Ivan Ezhov, Sergey Litvinov, Petros Koumoutsakos, Tamaz Amiranashvili, Ray Zirui Zhang, John S. Lowengrub, Bene Wiestler, Bjoern Menze
专题命中
多模态Agent
:multi-modal(abstract)
CommentsAccepted for publication in Nature Communications (DOI: 10.1038/s41467-025-60366-4)