Jailbreaking Large Vision Language Models in Intelligent Transportation Systems
Badhan Chandra Das, Md Tasnim Jawad, Md Jueal Mia, M. Hadi Amini, Yanzhao Wu
机构
*
KFSCIS, Florida International University(凯斯-西储大学信息科学学院)
;
Knight Foundation School of Computing and Information Sciences(骑士基金会计算与信息科学学院)
;
Learning for InterDependent Networks Laboratory (solid lab)(依赖网络学习实验室)
专题命中
视觉问答
:vision language model(title,abstract);visual question answering(abstract);分类 cs.AI
机构
*
South China University of Technology(华南理工大学)
;
Sun Yat-sen University(中山大学)
;
Hangzhou Dianzi University(杭州电子科技大学)
;
Zhejiang University of Finance & Economics(浙江财经大学)
;
National University of Singapore(新加坡国立大学)
;
Shenzhen Research Institute of Big Data(深圳大数据研究院)
;
City University of Hong Kong(香港城市大学)
VLMs Guided Interpretable Decision Making for Autonomous Driving
Xin Hu, Taotao Jing, Renran Tian, Zhengming Ding
机构
*
Department of Computer Science, Tulane University(路易斯安那大学计算机科学系)
;
Qualcomm(高通公司)
;
Department of Industrial and Systems Engineering, North Carolina State University(北卡罗来纳州立大学工业与系统工程系)
Enhancing Agentic Autonomous Scientific Discovery with Vision-Language Model Capabilities
Kahaan Gandhi, Boris Bolliet, Inigo Zubeldia
机构
*
Department of Physics, University of Cambridge, Cambridge, United Kingdom(剑桥大学物理系)
;
Kavli Institute for Cosmology, University of Cambridge, Cambridge, United Kingdom(剑桥大学卡弗利天文研究所)
;
Department of Physics and Astronomy, Haverford College, 370 Lancaster Avenue, Haverford, PA 19041, USA(哈弗福德学院物理与天文学系)
;
Division of Physics, Mathematics and Astronomy, California Institute of Technology, Pasadena, CA 91125, USA(加州理工学院物理、数学与天文学系)
;
Institute of Astronomy, University of Cambridge, Cambridge, United Kingdom(剑桥大学天文研究所)
Scene Graph-Guided Generative AI Framework for Synthesizing and Evaluating Industrial Hazard Scenarios
Sanjay Acharjee, Abir Khan Ratul, Diego Patino, Md Nazmus Sakib
机构
*
Ph.D. Student, Dept. of Civil Eng., University of Texas at Arlington. E-mail
;
Assistant Professor, Dept. of Computer Sci. \& Eng., University of Texas at Arlington. E-mail
;
Assistant Professor, Dept. of Civil Eng., University of Texas at Arlington. E-mail
MiniGPT-Pancreas: Multimodal Large Language Model for Pancreas Cancer Classification and Detection
Andrea Moglia, Elia Clement Nastasio, Luca Mainardi, Pietro Cerveri
机构
*
Department of Electronics, Information, and Bioengineering(电子、信息与生物工程系)
;
Polytechnic University of Milan(米兰理工学院)
;
Department of Industrial, and Information Engineering(工业与信息工程系)
;
University of Pavia(帕维亚大学)
专题命中
视觉定位与Grounding
:multimodal large language model(title,abstract);MLLM(abstract);分类 cs.CV、cs.AI
Journal refMoglia, A., Nastasio, E.C., Mainardi, L. et al. MiniGPT-Pancreas: Multimodal Large Language Model for Pancreas Cancer Observation and Localization in CT Images. J Healthc Inform Res (2025)
HiEAG: Evidence-Augmented Generation for Out-of-Context Misinformation Detection
Junjie Wu, Yumeng Fu, Nan Yu, Guohong Fu
机构
*
School of Computer Science and Technology, Soochow University(苏州大学计算机科学与技术学院)
;
Institute of Artificial Intelligence, Soochow University(苏州大学人工智能研究院)
;
School of Computer Science and Technology, Harbin Institute of Technology(哈尔滨工业大学计算机科学与技术学院)
专题命中
视觉定位与Grounding
:multimodal large language model(abstract);MLLM(abstract)
DepthVision: Enabling Robust Vision-Language Models with GAN-Based LiDAR-to-RGB Synthesis for Autonomous Driving
Sven Kirchner, Nils Purschke, Ross Greer, Alois C. Knoll
机构
*
Chair of Robotics, Artificial Intelligence and Real-time Systems, Technical University of Munich(机器人学、人工智能与实时系统教授会,慕尼黑技术大学)
;
Computer Science and Engineering Department, University of California Merced(计算机科学与工程系,加州大学默塞德分校)
Governance-Ready Small Language Models for Medical Imaging: Prompting, Abstention, and PACS Integration
Yiting Wang, Ziwei Wang, Di Zhu, Jiachen Zhong, Weiyi Li
机构
*
Department of Data Science, University of Southern California(数据科学系,南加州大学)
;
Department of Electrical and Computer Engineering, Carnegie Mellon University(电气与计算机工程系,卡内基梅隆大学)
;
Department of Computer Science and Engineering, Santa Clara University(计算机科学与工程系,圣克拉拉大学)
;
Department of Applied Mathematics, University of Washington(应用数学系,华盛顿大学)
;
School of Computer Science, Georgia Institute of Technology(计算机科学学院,佐治亚理工学院)
StyleDrive: Towards Driving-Style Aware Benchmarking of End-To-End Autonomous Driving
Ruiyang Hao, Bowen Jing, Haibao Yu, Zaiqing Nie
机构
*
AIR, Tsinghua University(空气动力学研究所,清华大学)
;
King’s College London(伦敦国王学院)
;
The University of Manchester(曼彻斯特大学)
;
The University of Hong Kong(香港大学)