机构
*
Bytedance(字节跳动)
;
Peking University(北京大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
The University of Hong Kong(香港大学)
;
Mohamed bin Zayed University of Artificial Intelligence(马尔代夫人工智能大学)
;
Stanford University(斯坦福大学)
;
University of Michigan(密歇根大学)
专题命中
视觉问答
:multimodal large language model(title,abstract);分类 cs.CV、cs.AI、cs.LG
Hulu-Med: A Transparent Generalist Model towards Holistic Medical Vision-Language Understanding
Songtao Jiang, Yuan Wang, Sibo Song, Tianxiang Hu, Chenyi Zhou, Bin Pu, Yan Zhang, Zhibo Yang, Yang Feng, Joey Tianyi Zhou, Jin Hao, Zijian Chen, Ruijia Wu, Tao Tang, Junhui Lv, Hongxia Xu, Hongwei Wang, Jun Xiao, Bin Feng, Fudong Zhu, Kenli Li, Weidi Xie, Jimeng Sun, Jian Wu, Zuozhu Liu
机构
*
College of Computer Science and Technology, Zhejiang University-University of Illinois Urbana-Champaign Institute(浙江大学计算机科学与技术学院)
;
Stomatology Hospital, School of Stomatology, Zhejiang University School of Medicine(浙江大学口腔医院)
;
Alibaba Inc(阿里巴巴集团)
;
College of Computer Science and Electronic Engineering, Hunan University(湖南大学计算机科学与电子工程学院)
;
Angelalign Technology Inc.(Angelalign技术有限公司)
;
CFAR & IHPC, Agency for Science, Technology and Research(CFAR与IHPC,新加坡科技研究局)
;
Department of Orthodontics, Shanghai Ninth People’s Hospital, College of Stomatology, Shanghai Jiao Tong University(上海第九人民医院正畸科,上海交通大学口腔医学院)
SurgAnt-ViVQA: Learning to Anticipate Surgical Events through GRU-Driven Temporal Cross-Attention
Shreyas C. Dhake, Jiayuan Huang, Runlong He, Danyal Z. Khan, Evangelos B. Mazomenos, Sophia Bano, Hani J. Marcus, Danail Stoyanov, Matthew J. Clarkson, Mobarak I. Hoque
机构
*
UCL Hawkes Institute(UCL哈维斯研究所)
;
University College London(伦敦大学学院)
;
Dept of Medical Physics & Biomedical Engineering(医学物理与生物医学工程系)
;
UCL(伦敦大学学院)
;
Dept of Computer Science(计算机科学系)
;
National Hospital for Neurology and Neurosurgery(神经病学与神经外科国家医院)
;
Division of Informatics, Imaging and Data Science(信息学、成像与数据科学 division)