VAGU & GtS: LLM-Based Benchmark and Framework for Joint Video Anomaly Grounding and Understanding
Shibo Gao, Peipei Yang, Yangyang Liu, Yi Chen, Han Zhu, Xuyao Zhang, Linlin Huang
机构
*
Beijing Jiaotong University(北京交通大学)
;
State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences(多模态人工智能系统国家重点实验室,自动化研究所,中国科学院)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
Knowledge Regularized Negative Feature Tuning of Vision-Language Models for Out-of-Distribution Detection
Wenjie Zhu, Yabin Zhang, Xin Jin, Wenjun Zeng, Lei Zhang
机构
*
Hong Kong Polytechnic University(香港理工大学)
;
Eastern Institute of Technology(东部技术研究所)
;
Stanford University(斯坦福大学)
;
Ningbo Institute of Digital Twin, Eastern Institute of Technology(宁波数字孪生研究所,东部技术研究所)
;
Institute for Clarity in Documentation(文档清晰研究所)
;
Inria Paris-Rocquencourt(巴黎-罗克琴堡研究所)
;
Rajiv Gandhi University(拉贾·甘地大学)
;
Tsinghua University(清华大学)
;
Palmer Research Laboratories(帕勒研究实验室)
The Evolution of Video Anomaly Detection: A Unified Framework from DNN to MLLM
Shibo Gao, Peipei Yang, Haiyang Guo, Yangyang Liu, Yi Chen, Shuai Li, Han Zhu, Jian Xu, Xu-Yao Zhang, Linlin Huang
机构
*
Beijing Jiaotong University(北京交通大学)
;
State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences(多模态人工智能系统国家重点实验室,自动化研究所,中国科学院)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(人工智能学院,中国科学院大学)
;
Zhongguancun Academy, Beijing, China(中关村学院,北京,中国)
机构
*
School of Mechanical and Electrical Engineering, University of Electronic Science and Technology of China(电子科技大学机械与电子工程学院)
;
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
Department of Pathology, Sichuan Clinical Research Center for Cancer, Sichuan Cancer Hospital & Institute, Affiliated Cancer Hospital of University of Electronic Science and Technology of China(pathology department, 四川省癌症临床研究中心, 四川省肿瘤医院及研究所, 电子科技大学附属肿瘤医院)
;
Department of Radiation Oncology, Sichuan Cancer Hospital and Institute, University of Electronic Science and Technology of China(放射肿瘤科, 四川省肿瘤医院及研究所, 电子科技大学)
Self-Aware Safety Augmentation: Leveraging Internal Semantic Understanding to Enhance Safety in Vision-Language Models
Wanying Wang, Zeyu Ma, Han Zheng, Xin Tan, Mingang Chen
机构
*
Shanghai Key Laboratory of Computer Software Testing and Evaluating(上海软件测试与评估 key laboratory)
;
Shanghai Normal University(上海Normal University)
;
TrustAI Pte. Ltd.
;
East China Normal University(东华师范大学)
Distribution-Based Masked Medical Vision-Language Model Using Structured Reports
Shreyank N Gowda, Ruichi Zhang, Xiao Gu, Ying Weng, Lu Yang
机构
*
School of Computer Science, University of Nottingham(计算机科学学院,诺丁汉大学)
;
Department of Computer Science and Technology, School of Informatics, Xiamen University(计算机科学与技术系,信息学院,厦门大学)
;
CHI Lab, University of Oxford(CHI实验室,牛津大学)
;
School of Computer Science, University of Nottingham Ningbo China(计算机科学学院,宁波大学中国)
机构
*
Institute of Artificial Intelligence (TeleAI), China Telecom, China(人工智能研究院(TeleAI),中国电信,中国)
;
Northwestern Polytechnical University(西北工业大学)
;
China Telecom, China(中国电信,中国)
专题命中
幻觉与鲁棒性
:multimodal large language model(abstract);分类 cs.AI
Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval
Zhichuan Wang, Yang Zhou, Zhe Liu, Rui Yu, Song Bai, Yulong Wang, Xinwei He, Xiang Bai
机构
*
Huazhong Agricultural University(华中农业大学)
;
Shenzhen University(深圳大学)
;
The University of Hong Kong(香港大学)
;
University of Louisville(路易斯安那大学)
;
ByteDance(字节跳动)
;
Huazhong University of Science and Technology(华中科技大学)