MLLM-HWSI: A Multimodal Large Language Model for Hierarchical Whole Slide Image Understanding
MLLM-HWSI: 一种用于分层全滑动图像理解的多模态大语言模型
Basit Alawode, Arif Mahmood, Muaz Khalifa Al-Radi, Shahad Albastaki, Asim Khan, Muhammad Bilal, Moshira Ali Abdalla, Mohammed Bennamoun, Sajid Javed
机构
*
Department of Computer Science, Khalifa University of Science and Technology(卡利法科技大学计算机科学系)
;
Information Technology University(信息技术大学)
;
KAU(卡乌大学)
;
University of the Western Australia(西澳大学)
Traffic Sign Recognition in Autonomous Driving: Dataset, Benchmark, and Field Experiment
自动驾驶中的交通标志识别:数据集、基准测试与实地实验
Guoyang Zhao, Weiqing Qi, Kai Zhang, Chenguang Zhang, Zeying Gong, Zhihai Bi, Kai Chen, Benshan Ma, Ming Liu, Jun Ma
机构
*
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
Lingnan University(岭大)
;
Shenzhen Unity Drive Innovation Technology Co., Ltd.(深圳Unity Drive创新技术有限公司)
;
The Hong Kong University of Science and Technology(香港科技大学)
机构
*
Tsinghua University(清华大学)
;
Beijing University of Posts and Telecommunications(北京邮电大学)
;
Tianjin University(天津大学)
;
Institute of Microelectronics of the Chinese Academy of Sciences(中国科学院微电子研究所)
;
HKUST (Guangzhou)(香港科技大学(广州))
;
National University of Defense Technology(国防科技大学)
;
Beihang University(北航)
;
Beijing Information Science and Technology University(北京信息科技大学)
;
Artificial Intelligence Institute of China Electronics Technology Group Corporation(中国电子科技集团人工智能研究院)
Mamba Learns in Context: Structure-Aware Domain Generalization for Multi-Task Point Cloud Understanding
Mamba在上下文中学习:面向多任务点云理解的结构感知领域泛化
Jincen Jiang, Qianyu Zhou, Yuhang Li, Kui Su, Meili Wang, Jian Chang, Jian Jun Zhang, Xuequan Lu
机构
*
Bournemouth University(伯恩茅斯大学)
;
Jilin University(吉林大学)
;
The University of Western Australia(西澳大学)
;
Hangzhou City University(杭州城市大学)
;
Northwest A&F University(西北农林科技大学)
Cross-modal Fuzzy Alignment Network for Text-Aerial Person Retrieval and A Large-scale Benchmark
跨模态模糊对齐网络用于文本-空中人检索及一个大规模基准
Yifei Deng, Chenglong Li, Yuyang Zhang, Guyue Hu, Jin Tang
机构
*
State Key Laboratory of Opto-Electronic Information Acquisition and Protection Technology(光电信息采集与防护技术国家重点实验室)
;
School of Computer Science and Technology, Anhui University(安徽大学计算机科学与技术学院)
;
School of Artificial Intelligence, Anhui University(安徽大学人工智能学院)
;
The University of Hong Kong(香港大学)
机构
*
Department of Management Science and Technology(管理科学与技术系)
;
Hellenic Mediterranean University(希伯伦地中海大学)
;
Institute of Computer Science, FORTH(信息科学研究所,FORTH)
GenVideoLens: Where LVLMs Fall Short in AI-Generated Video Detection?
GenVideoLens:在AI生成视频检测中LVLMs的局限性
Yueying Zou, Pei Pei Li, Zekun Li, Xinyu Guo, Xing Cui, Huaibo Huang, Ran He
机构
*
Beijing University of Posts and Telecommunications(北京邮电大学)
;
University of California, Santa Barbara(加州大学圣巴巴拉分校)
;
Center for Research on Intelligent Perception and Computing, NLPR, Institute of Automation, Chinese Academy of Sciences(智能感知与计算中心、国家智能感知与信息处理实验室、中国科学院自动化研究所)
Evidence Packing for Cross-Domain Image Deepfake Detection with LVLMs
跨领域图像深度伪造检测中的证据打包与大视觉语言模型
Yuxin Liu, Fei Wang, Kun Li, Yiqi Nie, Junjie Chen, Zhangling Duan, Zhaohong Jia
机构
*
Anhui University(安徽大学)
;
Hefei University of Technology(合肥工业大学)
;
IAI, Hefei Comprehensive National Science Center(合肥综合国家科学中心人工智能研究院)
;
United Arab Emirates University(阿拉伯联合酋长国大学)