Standardization of Neuromuscular Reflex Analysis -- Role of Fine-Tuned Vision-Language Model Consortium and OpenAI gpt-oss Reasoning LLM Enabled Decision Support System
Eranga Bandara, Ross Gore, Sachin Shetty, Ravi Mukkamala, Christopher Rhea, Atmaram Yarlagadda, Shaifali Kaushik, L. H. M. P. De Silva, Andriy Maznychenko, Inna Sokolowska, Amin Hass, Kasun De Zoysa
机构
*
Old Dominion University(旧 Dominion 大学)
;
AnaletIQ
;
McDonald Army Health Center(麦克唐纳陆军医疗中心)
;
University of Sri Jayewardenepura(Sri Jayewardenepura 大学)
;
University of Gdańsk(盖尔范大学)
;
University of Colombo(科伦坡大学)
MoMA: A Mixture-of-Multimodal-Agents Architecture for Enhancing Clinical Prediction Modelling
Jifan Gao, Mahmudur Rahman, John Caskey, Madeline Oguss, Ann O'Rourke, Randy Brown, Anne Stey, Anoop Mayampurath, Matthew M. Churpek, Guanhua Chen, Majid Afshar
机构
*
University of Wisconsin-Madison(威斯康星大学麦迪逊分校)
;
Northwestern University(西北大学)
Knowledge to Sight: Reasoning over Visual Attributes via Knowledge Decomposition for Abnormality Grounding
Jun Li, Che Liu, Wenjia Bai, Mingxuan Liu, Rossella Arcucci, Cosmin I. Bercea, Julia A. Schnabel
机构
*
Technical University of Munich(慕尼黑技术大学)
;
Munich Center for Machine Learning(慕尼黑机器学习中心)
;
Imperial College London(伦敦帝国学院)
;
University of Trento(特伦托大学)
;
Helmholtz AI and Helmholtz Munich(海德堡人工智能与海德堡慕尼黑)
;
King’s College London(伦敦国王学院)
机构
*
School of Computer Science, Peking University(北京大学计算机科学学院)
;
School of Information and Software Engineering, University of Electronic Science and Technology of China(电子科技大学信息与软件工程学院)
;
National Engineering Research Center for Software Engineering, Peking University(软件工程国家工程研究中心)
;
Institute of Software, Chinese Academy of Science(中国科学院软件研究所)
;
School of Software and Microelectronics, Peking University(北京大学软件与微电子学院)
;
Institute of Basic Theory of Chinese Medicine, China Academy of Chinese Medical Sciences(中国中医科学院中医基础理论研究所)
Distribution-Based Masked Medical Vision-Language Model Using Structured Reports
Shreyank N Gowda, Ruichi Zhang, Xiao Gu, Ying Weng, Lu Yang
机构
*
School of Computer Science, University of Nottingham(计算机科学学院,诺丁汉大学)
;
Department of Computer Science and Technology, School of Informatics, Xiamen University(计算机科学与技术系,信息学院,厦门大学)
;
CHI Lab, University of Oxford(CHI实验室,牛津大学)
;
School of Computer Science, University of Nottingham Ningbo China(计算机科学学院,宁波大学中国)
机构
*
School of Mechanical and Electrical Engineering, University of Electronic Science and Technology of China(电子科技大学机械与电子工程学院)
;
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
Department of Pathology, Sichuan Clinical Research Center for Cancer, Sichuan Cancer Hospital & Institute, Affiliated Cancer Hospital of University of Electronic Science and Technology of China(pathology department, 四川省癌症临床研究中心, 四川省肿瘤医院及研究所, 电子科技大学附属肿瘤医院)
;
Department of Radiation Oncology, Sichuan Cancer Hospital and Institute, University of Electronic Science and Technology of China(放射肿瘤科, 四川省肿瘤医院及研究所, 电子科技大学)
Mind the Modality Gap: Towards a Remote Sensing Vision-Language Model via Cross-modal Alignment
Angelos Zavras, Dimitrios Michail, Begüm Demir, Ioannis Papoutsis
机构
*
organization= Orion Lab, National Observatory of Athens \& National Technical University of Athens , country= Greece
;
organization= Department of Informatics \& Telematics, Harokopio University of Athens , country= Greece
;
organization= Faculty of Electrical Engineering
;
organization= BIFOLD - Berlin Institute for the Foundations of Learning
专题命中
医疗多模态
:medical image(abstract);分类 cs.CV
CommentsAccepted at the ISPRS Journal of Photogrammetry and Remote Sensing. Our code implementation and weights for all experiments are publicly available at https://github.com/Orion-AI-Lab/MindTheModalityGap
机构
*
National Key Laboratory of Space Integrated Information System(国家空间信息集成系统重点实验室)
;
Institute of Software Chinese Academy of Sciences(中国科学院软件研究所)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Wangxuan Institute of Computer Technology(计算机技术王轩研究所)
;
Peking University(北京大学)
;
Department of Computer Science and Technology(计算机科学与技术系)
;
Tsinghua University(清华大学)
;
Thrust of Artificial Intelligence, the Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)人工智能方向)
;
Department of Computer Science & Engineering, the Hong Kong University of Science and Technology(香港科技大学计算机科学与工程系)
Hyperbolic Kernel Graph Neural Networks for Neurocognitive Decline Analysis from Multimodal Brain Imaging
Meimei Yang, Yongheng Sun, Qianqian Wang, Andrea Bozoki, Maureen Kohi, Mingxia Liu
机构
*
Department of Radiology and Biomedical Research Imaging Center (BRIC), University of North Carolina at Chapel Hill(放射科与生物医学研究成像中心(BRIC)、北卡罗来纳大学教堂山分校)
CommentsExtended version of EMNLP 2024 paper arXiv:2411.04118. Includes additional results on clinical note QA tasks and supervised fine-tuning evaluations
机构
*
MBZUAI, United Arab Emirates(MBZUAI,阿联酋)
;
Monash University, Australia(墨尔本大学,澳大利亚)
;
Liverpool University, United Kingdom(利物浦大学,英国)
;
China Unicom (Shanghai) Industrial Internet Co., Ltd., China(中国联合(上海)工业互联网有限公司,中国)
;
Shanghai AI Lab, China(上海人工智能实验室,中国)
Memory-Augmented Incomplete Multimodal Survival Prediction via Cross-Slide and Gene-Attentive Hypergraph Learning
Mingcheng Qu, Guang Yang, Donglin Di, Yue Gao, Tonghua Su, Yang Song, Lei Fan
机构
*
Faculty of Computing, Harbin Institute of Technology(哈尔滨工业大学计算机学院)
;
School of Software, Tsinghua University(清华大学软件学院)
;
School of Computer Science and Engineering, UNSW Sydney(新南威尔士大学计算机科学与工程学院)
VideoMathQA: Benchmarking Mathematical Reasoning via Multimodal Understanding in Videos
Hanoona Rasheed, Abdelrahman Shaker, Anqi Tang, Muhammad Maaz, Ming-Hsuan Yang, Salman Khan, Fahad Shahbaz Khan
机构
*
MBZUAI(穆桑大学人工智能研究所)
;
University of California Merced(加州大学默塞德分校)
;
Google Research(谷歌研究)
;
Australian National University(澳大利亚国立大学)
;
Linköping University(林肯大学)
机构
*
Hong Kong Baptist University(香港 Baptist 大学)
;
Shanxi University(山西大学)
;
Shanghai Institute for Advanced Study of Zhejiang University(浙江大学上海研究院)
;
Tencent AI Lab(腾讯人工智能实验室)
;
Manchester Metropolitan University(曼彻斯特 Metropolitan 大学)
;
The University of Manchester(曼彻斯特大学)
专题命中
医疗多模态
:diagnosis(abstract);分类 cs.LG
CommentsThis paper is accepted by ACL2025(Findings)
Prescribing the Right Remedy: Mitigating Hallucinations in Large Vision-Language Models via Targeted Instruction Tuning
Rui Hu, Yahan Tu, Shuyu Wei, Dongyuan Lu, Jitao Sang
机构
*
Beijing Key Lab of Traffic Data Analysis and Mining(北京交通大数据分析与挖掘重点实验室)
;
Beijing Jiaotong University(北京交通大学)
;
School of Information Technology and Management(信息科学技术学院)