MedVL-SAM2: A unified 3D medical vision-language model for multimodal reasoning and prompt-driven segmentation
MedVL-SAM2:一种统一的3D医学视觉-语言模型,用于多模态推理和基于提示的分割
Yang Xing, Jiong Wu, Savas Ozdemir, Ying Zhang, Yang Yang, Wei Shao, Kuang Gong
机构
*
Department of Biomedical Engineering, University of Florida(佛罗里达大学生物医学工程系)
;
Department of Radiology, University of Florida(佛罗里达大学放射学系)
;
Research Computing, University of Florida(佛罗里达大学研究计算中心)
;
Department of Medicine, University of Florida(佛罗里达大学医学系)
;
Department of Radiology, UC San Francisco(旧金山大学放射学系)
SVII-3D: Advancing Roadside Infrastructure Inventory with Decimeter-level 3D Localization and Comprehension from Sparse Street Imagery
SVII-3D:利用厘米级3D定位与稀疏街道影像的综合理解,推进道路基础设施库存建设
Chong Liu, Luxuan Fu, Yang Jia, Zhen Dong, Bisheng Yang
机构
*
State Key Laboratory of Information Engineering in Surveying, Mapping and Remote Sensing (LIESMARS), Wuhan University, Wuhan 430079, China(信息工程测绘遥感国家重点实验室(LIESMARS),武汉大学)
;
Research Institute Ltd, Chengdu 610000, China(四川省公路规划设计研究有限公司)
MoST: Mixing Speech and Text with Modality-Aware Mixture of Experts
MoST:通过模态感知混合专家混合语音和文本
Yuxuan Lou, Kai Yang, Yang You
机构
*
School of Computer Science, National University of Singapore(新加坡国立大学计算机科学学院)
;
School of Computer Science, Shanghai Jiao Tong University(上海交通大学计算机科学学院)
The State-of-the-Art in Lifelog Retrieval: A Review of Progress at the ACM Lifelog Search Challenge Workshop 2022-24
生命周期检索的最新进展:ACM生命周期检索挑战工作坊2022-24年进展综述
Allie Tran, Werner Bailer, Duc-Tien Dang-Nguyen, Graham Healy, Steve Hodges, Björn Þór Jónsson, Luca Rossetto, Klaus Schoeffmann, Minh-Triet Tran, Lucia Vadicamo, Cathal Gurrin
机构
*
College of Artificial Intelligence, Nankai University, Tianjin, China(人工智能学院,南开大学,天津,中国)
;
School of Data Science, Fudan University, Shanghai, China(数据科学学院,复旦大学,上海,中国)
AEQ-Bench: Measuring Empathy of Omni-Modal Large Models
AEQ-Bench:衡量多模态大模型的共情能力
Xuan Luo, Lewei Yao, Libo Zhao, Lanqing Hong, Kai Chen, Dehua Tao, Daxin Tan, Ruifeng Xu, Jing Li
机构
*
The Hong Kong Polytechnic University(香港理工大学)
;
The Harbin Institute of Technology(哈尔滨工业大学)
;
Huawei(华为)
;
Hong Kong University of Science and Technology(香港理工大学)
;
Shenzhen Loop Area Institute(深圳河套学院)
The CASTLE 2024 Dataset: Advancing the Art of Multimodal Understanding
CASTLE 2024数据集:推动多模态理解艺术的发展
Luca Rossetto, Werner Bailer, Duc-Tien Dang-Nguyen, Graham Healy, Björn Þór Jónsson, Onanong Kongmeesub, Hoang-Bao Le, Stevan Rudinac, Klaus Schöffmann, Florian Spiess, Allie Tran, Minh-Triet Tran, Quang-Linh Tran, Cathal Gurrin
机构
*
Dublin City University(都柏林城市大学)
;
JOANNEUM RESEARCH(JOANNEUM研究机构)
;
University of Bergen(卑尔根大学)
;
Reykjavik University(雷克雅未克大学)
;
University of Amsterdam(阿姆斯特丹大学)
;
Klagenfurt University(克雷克夫特大学)
;
University of Basel(巴塞尔大学)
;
VNU Ho Chi Minh University of Science(越南胡志明国家科学大学)
机构
*
CAS Key Laboratory of AI Safety, Institute of Computing Technology, CAS, Beijing, China(中国科学院人工智能安全重点实验室,计算技术研究所,中国科学院,北京,中国)
;
University of Chinese Academy of Sciences, Beijing, China(中国科学院大学,北京,中国)
;
Tsinghua University, Beijing, China(清华大学,北京,中国)
Advancing Adaptive Multi-Stage Video Anomaly Reasoning: A Benchmark Dataset and Method
推动自适应多阶段视频异常推理:一个基准数据集和方法
Chao Huang, Benfeng Wang, Wei Wang, Jie Wen, Li Shen, Wenqi Ren, Yong Xu, Xiaochun Cao
机构
*
School of Cyber Science and Technology, Shenzhen Campus of Sun Yat-sen University(中山大学计算机科学与技术学院(深圳校区))
;
Shenzhen Key Laboratory of Visual Object Detection and Recognition, Harbin Institute of Technology(哈尔滨工业大学深圳视觉目标检测与识别重点实验室)
A continental-scale dataset of ground beetles with high-resolution images and validated morphological trait measurements
大陆尺度的地面甲虫高分辨率图像数据集及经验证的形态学特征测量
S M Rayeed, Mridul Khurana, Alyson East, Isadora E. Fluck, Elizabeth G. Campolongo, Samuel Stevens, Iuliia Zarubiieva, Scott C. Lowe, Michael W. Denslow, Evan D. Donoso, Jiaman Wu, Michelle Ramirez, Benjamin Baiser, Charles V. Stewart, Paula Mabee, Tanya Berger-Wolf, Anuj Karpatne, Hilmar Lapp, Robert P. Guralnick, Graham W. Taylor, Sydne Record
机构
*
Rensselaer Polytechnic Institute(伦斯勒理工学院)
;
Virginia Tech(弗吉尼亚理工大学)
;
The University of Maine(缅因大学)
;
University of Florida(佛罗里达大学)
;
The Ohio State University(俄亥俄州立大学)
;
Vector Institute(向量研究所)
;
University of Guelph(圭尔夫大学)
;
National Ecological Observatory Network(国家生态观测网络)
OT-Drive: Out-of-Distribution Off-Road Traversable Area Segmentation via Optimal Transport
OT-Drive: 基于最优传输的离群分布越野可通行区域分割
Zhihua Zhao, Guoqiang Li, Chen Min, Kangping Lu
机构
*
School of Mechanical Engineering, Beijing Institute of Technology(北京理工大学机械工程学院)
;
Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所)
;
Shandong Pengxiang Automobile Co., Ltd(山东鹏翔汽车有限公司)
机构
*
School of Biomedical Engineering, Southern Medical University(生物医学工程学院,南方医科大学)
;
School of Biomedical Engineering, Shanghai Jiaotong University(生物医学工程学院,上海交通大学)
;
Department of Electronic Engineering, Chinese University of Hong Kong(电子工程系,中国香港大学)
;
Faculty of Dentistry, The University of Hong Kong(牙科学院,香港大学)
;
Department of Nuclear Medicine, The Second Affiliated Hospital of Guangzhou University of Chinese Medicine(核医学科,广州中医药大学第二附属医院)
;
PET Center, Department of Nuclear Medicine, Guangdong Provincial People’s Hospital, Southern Medical University(PET中心,核医学科,广东省人民医院,南方医科大学)
;
Department of Nuclear Medicine, Nanfang Hospital, Southern Medical University(核医学科,南芳医院,南方医科大学)
;
Division of Nuclear Medicine and Molecular Imaging, Geneva University Hospitals(核医学与分子影像学部,日内瓦大学医院)
;
Departments of Radiology, Physics, and Biomedical Engineering, The University of British Columbia(放射学、物理和生物医学工程系,不列颠哥伦比亚大学)
;
Medical Artificial Intelligence Laboratory, Westlake University(医学人工智能实验室,西湖大学)
Terrain-Adaptive Mobile 3D Printing with Hierarchical Control
地形自适应移动三维打印与分层控制
Shuangshan Nors Li, J. Nathan Kutz
机构
*
Department of Electrical and Computer Engineering, University of Washington, USA(电气与计算机工程系,华盛顿大学)
;
Department of Applied Mathematics, University of Washington, USA(应用数学系,华盛顿大学)