UniSurgSAM: A Unified Promptable Model for Reliable Surgical Video Segmentation
UniSurgSAM: 一种用于可靠外科视频分割的统一提示模型
Haofeng Liu, Ziyue Wang, Alex Y. W. Kong, Guanyi Qin, Yunqiu Xu, Chang Han Low, Mingqi Gao, Lap Yan Lennon Chan, Yueming Jin
机构
*
Department of Biomedical Engineering, National University of Singapore(新加坡国立大学生物医学工程系)
;
Department of Electrical and Computer Engineering, National University of Singapore(新加坡国立大学电气与计算机工程系)
;
School of Computer Science, The University of Sheffield(谢菲尔德大学计算机科学学院)
;
Department of Computer Science and Engineering, The Chinese University of Hong Kong(香港中文大学计算机科学与工程系)
TTA-Vid: Generalized Test-Time Adaptation for Video Reasoning
TTA-Vid: 视频推理的通用测试时间适应
Soumya Shamarao Jahagirdar, Edson Araujo, Anna Kukleva, M. Jehanzeb Mirza, Saurabhchand Bhati, Samuel Thomas, Brian Kingsbury, Rogerio Feris, James R. Glass, Hilde Kuehne
机构
*
University of Tübingen(蒂宾根大学)
;
MIT(麻省理工学院)
;
IBM Research(IBM研究院)
;
MIT-IBM Watson AI Lab(MIT-IBM沃森人工智能实验室)
;
Tuebingen AI Center(蒂宾根人工智能中心)
;
Max Planck Institute for Informatics(马克斯·普朗克信息学研究所)
Know-Show: Benchmarking Video-Language Models on Spatio-Temporal Grounded Reasoning
Know-Show:在时空 grounded 推理上评估视频语言模型的基准测试
Chinthani Sugandhika, Chen Li, Deepu Rajan, Basura Fernando
机构
*
College of Computing and Data Science, Nanyang Technological University(南洋理工大学计算与数据科学学院)
;
Institute of High-Performance Computing, Agency for Science, Technology and Research(新加坡科技研究局高性能计算研究所)
;
Centre for Frontier AI Research, Agency for Science, Technology and Research(新加坡科技研究局前沿人工智能研究中心)
MA-Bench: Towards Fine-grained Micro-Action Understanding
MA-Bench:迈向细粒度微动作理解
Kun Li, Jihao Gu, Fei Wang, Zhiliang Wu, Hehe Fan, Dan Guo
机构
*
CVLab, College of Information Technology, United Arab Emirates University(阿拉伯联合酋长国大学信息技术学院CVLab)
;
University College London(伦敦大学学院)
;
Hefei University of Technology(合肥工业大学)
;
Institute of Artificial Intelligence, Hefei Comprehensive National Science Center(合肥综合性国家科学中心人工智能研究院)
;
CCAI, Zhejiang University(浙江大学计算机辅助设计与图形学国家重点实验室)