MedCLM: Learning to Localize and Reason via a CoT-Curriculum in Medical Vision-Language Models
Soo Yong Kim, Suin Cho, Vincent-Daniel Yun, Gyeongyeon Hwang
机构
*
A.I.MATICS Inc(A.I.MATICS公司)
;
Boston University(波士顿大学)
;
University of Southern California(南加州大学)
;
Heuron(Heuron公司)
;
MODULABS, Open Neural Networks Research Lab(MODULABS,开放神经网络研究实验室)
机构
*
School of Computer Science and Engineering, Southeast University, China(东南大学计算机科学与工程学院)
;
Key Laboratory of New Generation Artificial Intelligence Technology and its Interdisciplinary Applications (Southeast University), Ministry of Education, China(新一代人工智能技术及 interdisciplinary 应用关键实验室(东南大学))
;
China Mobile Research Institute(中国移动研究院)
;
Department of Data Science & AI, Monash University, Australia(莫纳什大学数据科学与人工智能系)
机构
*
XLANG Lab, The University of Hong Kong(香港大学XLANG实验室)
;
Moonshot AI
;
Stanford University(斯坦福大学)
;
University of Waterloo(滑铁卢大学)
;
Carnegie Mellon University(卡内基梅隆大学)
CommentsThe extended full version of the accepted paper in 2025 IEEE BHI conference with title: Evaluating Large Multimodal Models for Nutrition Analysis: A New Benchmark Enriched with Contextual Metadata. Dataset is available at: https://skynet.ecn.purdue.edu/~coburn6/ACETADA/
GRACE: Generative Representation Learning via Contrastive Policy Optimization
Jiashuo Sun, Shixuan Liu, Zhaochen Su, Xianrui Zhong, Pengcheng Jiang, Bowen Jin, Peiran Li, Weijia Shi, Jiawei Han
机构
*
University of Illinois Urbana–Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
Australian National University(澳大利亚国立大学)
;
Hong Kong University of Science and Technology(香港科学与技术大学)
;
University of Wisconsin–Madison(威斯康星大学麦迪逊分校)
;
University of Washington(华盛顿大学)
GDPval: Evaluating AI Model Performance on Real-World Economically Valuable Tasks
Tejal Patwardhan, Rachel Dias, Elizabeth Proehl, Grace Kim, Michele Wang, Olivia Watkins, Simón Posada Fishman, Marwan Aljubeh, Phoebe Thacker, Laurance Fauconnet, Natalie S. Kim, Patrick Chao, Samuel Miserendino, Gildas Chabot, David Li, Michael Sharman, Alexandra Barr, Amelia Glaese, Jerry Tworek
Small Language Models for Emergency Departments Decision Support: A Benchmark Study
Zirui Wang, Jiajun Wu, Braden Teitge, Jessalyn Holodinsky, Steve Drew
机构
*
Department of Electrical and Software Engineering, University of Calgary, Calgary, AB, Canada(电气与软件工程系,卡尔加里大学)
;
Department of Emergency Medicine, University of Calgary, Calgary, AB, Canada(急诊医学系,卡尔加里大学)
;
Rockview General Hospital, Calgary, AB, Canada(罗克维尔医院)
专题命中
推理评测
:reasoning(abstract);分类 cs.CL、cs.AI
CommentsAccepted to 2025 IEEE International Conference on Autonomous and Trusted Computing (ATC 2025)
机构
*
Department of Data Science and AI, IIT Madras, India(数据科学与人工智能系,印度理工学院马德拉斯学院)
;
Department of Engineering Design, IIT Madras, India(工程设计系,印度理工学院马德拉斯学院)
;
LoveForm Health Technologies, India(LoveForm健康科技公司,印度)
;
Department of Radiology and Imaging Sciences, Sri Ramachandra Institute of Higher Education and Research, India(放射学与成像科学系, Sri Ramachandra高等教育与研究学院,印度)
;
Department of Neuro and Interventional Radiology, Sri Ramachandra Institute of Higher Education and Research, India(神经放射学与介入放射学系,Sri Ramachandra高等教育与研究学院,印度)
专题命中
推理评测
:reasoning(abstract);分类 cs.AI
CommentsPaper published at "Agentic AI for Medicine" Workshop, MICCAI 2025
MMLongBench: Benchmarking Long-Context Vision-Language Models Effectively and Thoroughly
Zhaowei Wang, Wenhao Yu, Xiyu Ren, Jipeng Zhang, Yu Zhao, Rohit Saxena, Liang Cheng, Ginny Wong, Simon See, Pasquale Minervini, Yangqiu Song, Mark Steedman
机构
*
CSE Department, HKUST(香港科技大学计算机科学与工程系)
;
Tencent AI Seattle Lab(腾讯AI西雅图实验室)
;
University of Edinburgh(爱丁堡大学)
;
NVIDIA AI Technology Center (NVAITC), NVIDIA, Santa Clara, USA(英伟达圣克拉拉人工智能技术中心)
Social Good or Scientific Curiosity? Uncovering the Research Framing Behind NLP Artefacts
Eric Chamoun, Nedjma Ousidhoum, Michael Schlichtkrull, Andreas Vlachos
机构
*
Department of Computer Science and Technology, University of Cambridge(计算机科学与技术系,剑桥大学)
;
Cardiff University(卡迪夫大学)
;
Queen Mary University of London(伦敦女王学院)