HalluCXR: Benchmarking and Mitigating Hallucinations in Medical Vision-Language Models for Chest Radiograph Interpretation
HalluCXR: 评估和缓解医疗视觉-语言模型在胸部X光解读中的幻觉
Haoyu Wang, Zitong Li
机构
*
Department of Biostatistics & Health Informatics, Institute of Psychiatry, Psychology & Neuroscience, King’s College London(生物统计学与健康信息学系,精神病学、心理学与神经科学研究所,伦敦国王学院)
RECIPE: Procedural Planning via Grounding in Instructional Video
RECIPE: 通过指令视频中的 grounding 实现过程规划
Luigi Seminara, Antonino Furnari, Lorenzo Torresani
机构
*
Khoury College of Computer Sciences, Northeastern University, Boston(东北大学北斯托顿学院计算机科学学院)
;
Department of Mathematics and Computer Science, University of Catania, Italy(卡塔尼亚大学数学与计算机科学系)
Rapid patient-specific neural networks for intraoperative X-ray to volume registration
快速的患者特异性神经网络用于术中X射线到体积的配准
Vivek Gopalakrishnan, David-Dimitris Chlorogiannis, Andrew Abumoussa, Anna M. Larson, Nazim Haouchine, Darren B. Orbach, Sarah Frisken, Neel Dey, Polina Golland
机构
*
Harvard-MIT Health Sciences and Technology, Massachusetts Institute of Technology(哈佛-麻省理工健康科学与技术, 麻省理工学院)
;
Computer Science and Artificial Intelligence Laboratory, Massachusetts Institute of Technology(计算机科学与人工智能实验室, 麻省理工学院)
;
Department of Radiology, Harvard Medical School(哈佛医学院放射科)
;
Saint Luke’s Marion Bloch Neuroscience Institute(圣路易斯马里恩布洛克神经科学研究所)
;
Department of Critical Care Medicine, Shriners Children’s Hospital(谢尔曼儿童医院重症医学科)
;
Department of Interventional Neuroradiology, Boston Children’s Hospital(波士顿儿童医院介入神经放射科)
;
Athinoula A. Martinos Center for Biomedical Imaging, Massachusetts General Hospital(阿提努拉A·马丁诺斯生物医学成像中心, 麻省总医院)
机构
*
School of Computer Science and Engineering, Macau University of Science and Technology, Macao SAR(澳门科学技术大学计算机科学与工程学院)
;
Hubei Key Laboratory of Transportation Internet of Things, School of Computer Science and Artificial Intelligence, Wuhan University of Technology(湖北省交通物联网重点实验室,武汉理工大学)
机构
*
School of Geosciences and Info-Physics, Central South University(地质科学与信息物理学院,中南大学)
;
School of Earth Sciences and Spatial Information Engineering, Hunan University of Science and Technology(地球科学与空间信息工程学院,湖南科技大学)
EduVQA: Towards Concept-Aware Assessment of Educational AI-Generated Videos
EduVQA: 向概念感知的教育AI生成视频评估迈进
Baoliang Chen, Xinlong Bu, Hanwei Zhu, Lingyu Zhu, Jieyu Zhan
机构
*
College of Computing and Data Science, Nanyang Technological University(南洋理工大学计算与数据科学学院)
;
Department of Computer Science, South China Normal University, China(华南师范大学计算机学院)
;
School of Computer Science, City University of Hong Kong(香港城市大学计算机科学学院)
Comments34 pages, 3 figures, 12 tables. Submitted to ACM Computing Surveys. v2: title shortened to "A Survey"; restructured taxonomy section; captions and acronym handling aligned with ACM CSUR style; bibliography updated to 174 entries
DisasterVQA: A Visual Question Answering Benchmark Dataset for Disaster Scenes
DisasterVQA: 一个用于灾难场景的视觉问答基准数据集
Aisha Al-Mohannadi, Ayisha Firoz, Yin Yang, Muhammad Imran, Ferda Ofli
机构
*
Qatar Computing Research Institute(卡塔尔计算研究所)
;
Hamad Bin Khalifa University(哈马德·本·卡伊夫大学)
;
College of Science & Engineering(科学与工程学院)
;
Qatar University(卡塔尔大学)
EPIC-Bench: A Perception-Centric Benchmark for Fine-Grained Embodied Visual Grounding in Vision-Language Models
EPIC-Bench: 一种以感知为中心的细粒度具身视觉 grounding 的基准
Haozhe Shan, Xiancong Ren, Han Dong, Haoyuan Shi, Yingji Zhang, Jiayu Hu, Yi Zhang, Yong Dai, Bin Shen, Lizhen Qu, Zenglin Xu, Xiaozhu Ju
机构
*
X-Humanoid
;
Fudan University(复旦大学)
;
University of Science and Technology of China(中国科学技术大学)
;
University of Manchester(曼彻斯特大学)
;
Monash University(墨尔本大学)
;
Celonis AI
;
University of New South Wales(新南威尔士大学)
ClickSeg3D: Few-Click Interactive Segmentation via Semantic Embeddings
ClickSeg3D: 通过语义嵌入实现少点击交互分割
Xueyang Kang, Zijian Yu, Kourosh Khoshelham, Liangliang Nan
机构
*
School of Electrical and Electronic Engineering, Nanyang Technological University(南洋理工大学电子与电气工程学院)
;
University of Science and Technology of China(中国科学技术大学)
;
Faculty of Architecture and the Built Environment, Delft University of Technology(代尔夫特理工大学建筑与环境学院)