UASTrack: A Unified Adaptive Selection Framework with Modality-Customization in Single Object Tracking
UASTrack: 一种具有模态定制化的统一自适应选择框架在单目标跟踪中
He Wang, Tianyang Xu, Zhangyong Tang, Xiao-Jun Wu, Josef Kittler
机构
*
School of Artificial Intelligence and Computer Science, Jiangnan University(江南大学人工智能与计算机科学学院)
;
Centre for Vision, Speech and Signal Processing, University of Surrey(萨里大学视觉、语音和信号处理中心)
PFM-VEPAR: Prompting Foundation Models for RGB-Event Camera based Pedestrian Attribute Recognition
基于RGB事件相机的行人属性识别的PFM-VEPAR:用于基础模型的提示方法
Minghe Xu, Rouying Wu, ChiaWei Chu, Xiao Wang, Yu Li
机构
*
City University of Macau(澳门城市大学)
;
Zhuhai College of Science and Technology(珠海科学技术学院)
;
Macau University of Science and Technology(澳门科学大学)
;
School of Computer Science and Technology, Anhui University(安徽大学计算机科学与技术学院)
S3T-Former: A Purely Spike-Driven State-Space Topology Transformer for Skeleton Action Recognition
S3T-Former:一种纯粹由脉冲驱动的状态空间拓扑变换器用于骨骼动作识别
Naichuan Zheng, Hailun Xia, Zepeng Sun, Weiyi Li, Yujia Wang
机构
*
School of Information and Communication Engineering(信息与通信工程学院)
;
Beijing University of Posts and Telecommunications(北京邮电大学)
;
International School(国际学校)
;
School of Economics and Management(经济管理学院)
VideoAtlas: Navigating Long-Form Video in Logarithmic Compute
VideoAtlas: 在对数计算中导航长视频
Mohamed Eltahir, Ali Habibullah, Yazan Alshoibi, Lama Ayash, Tanveer Hussain, Naeemullah Khan
机构
*
King Abdullah University of Science and Technology (KAUST)(卡斯特大学)
;
Department of Computer Science, King Khalid University (KKU)(国王 Khalid 大学计算机科学系)
;
Department of Computer Science, Edge Hill University(Edge Hill 大学计算机科学系)
Symphony: A Cognitively-Inspired Multi-Agent System for Long-Video Understanding
Symphony:一种受认知启发的多智能体系统用于长视频理解
Haiyang Yan, Hongyun Zhou, Peng Xu, Xiaoxue Feng, Mengyi Liu
机构
*
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
Kuaishou Technology(快手科技)
;
School of Future Technology, University of Chinese Academy of Sciences(中国科学院大学未来技术学院)
Geometry-Aware Semantic Reasoning for Training Free Video Anomaly Detection
基于几何的语义推理用于无训练视频异常检测
Ali Zia, Usman Ali, Muhammad Umer Ramzan, Hamza Abid, Abdul Rehman, Wei Xiang
机构
*
School of Computing, Engineering \& Mathematical Sciences, La Trobe University, Melbourne, Australia School of Engineering
;
Applied Sciences, GIFT University, Gujranwala 52250, Pakistan
Aura: Universal Multi-dimensional Exogenous Integration for Aviation Time Series
Aura:通用多维外源整合用于航空时间序列
Jiafeng Lin, Mengren Zheng, Simeng Ye, Yuxuan Wang, Huan Zhang, Yuhui Liu, Zhongyi Pei, Jianmin Wang
机构
*
School of Software, BNRist, Tsinghua University(软件学院、BNRist、清华大学)
;
Hongshen College, Chongqing University(弘深学院、重庆大学)
;
School of Software, Tsinghua University(软件学院、清华大学)
;
China Southern Airlines(中国南方航空)
Zikang Liu, Longteng Guo, Handong Li, Ru Zhen, Xingjian He, Ruyi Ji, Xiaoming Ren, Yanhao Zhang, Haonan Lu, Jing Liu
机构
*
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
OPPO AI Center, OPPO Inc.(OPPO人工智能中心,OPPO公司)
机构
*
School of Mechanical Engineering, University of Science and Technology Beijing(北京科技大学机械工程学院)
;
Laboratory for Computational Sensing and Robotics, Johns Hopkins University(约翰霍普金斯大学计算感知与机器人实验室)
MEGC2026: Micro-Expression Grand Challenge on Visual Question Answering
MEGC2026:面向视觉问答的微表情大奖挑战
Xinqi Fan, Jingting Li, John See, Moi Hoon Yap, Su-Jing Wang, Adrian K. Davison
机构
*
Department of Computing and Mathematics, Manchester Metropolitan University(曼彻斯特 Metropolitan 大学计算与数学系)
;
State Key Laboratory of Cognitive Science and Mental Health, Institute of Psychology, CAS & Department of Psychology, University of the Chinese Academy of Sciences(中国科学院心理研究所认知科学与心理健康国家重点实验室及中国科学院大学心理学系)
;
School of Mathematical and Computer Sciences, Heriot-Watt University Malaysia(赫瑞-沃森大学马来西亚分校数学与计算机科学学院)
机构
*
Hong Kong University of Science and Technology (HKUST)(香港科技大学)
;
Tongyi Fun Team, Alibaba Group(通义Fun团队,阿里巴巴集团)
;
The Chinese University of Hong Kong (CUHK)(香港中文大学)
机构
*
School of Intelligence Science and Technology, Peking University(北京大学智能科学与技术学院)
;
State Key Laboratory of General Artificial Intelligence, BIGAI(通用人工智能国家重点实验室)
机构
*
Shenzhen International Graduate School Tsinghua University Shenzhen China(深圳国际研究生院清华大学深圳中国)
;
Independent Researcher London United Kingdom(独立研究者伦敦英国)
;
Queen Mary University of London London United Kingdom(女王玛丽大学伦敦英国)
;
Imperial College London London United Kingdom(帝国理工学院伦敦英国)
;
University Of Surrey Guildford United Kingdom(Surrey大学Guildford英国)
Prompt When the Animal is: Temporal Animal Behavior Grounding with Positional Recovery Training
提示当动物是:基于位置恢复训练的时序动物行为 grounding
Sheng Yan, Xin Du, Zongying Li, Yi Wang, Hongcang Jin, Mengyuan Liu
机构
*
School of Artificial Intelligence, Chongqing University of Technology(重庆理工大学人工智能学院)
;
National Key Laboratory of General Artificial Intelligence, Shenzhen Graduate School, Peking University(北京大学深圳研究生院国家通用人工智能实验室)
Ctrl-GenAug: Controllable Generative Augmentation for Medical Sequence Classification
Ctrl-GenAug: 可控生成增强用于医学序列分类
Xinrui Zhou, Yuhao Huang, Haoran Dou, Shijing Chen, Ao Chang, Jia Liu, Weiran Long, Jian Zheng, Erjiao Xu, Jie Ren, Alejandro F. Frangi, Ruobing Huang, Jun Cheng, Xiaomeng Li, Wufeng Xue, Dong Ni
机构
*
National-Regional Key Technology Engineering Laboratory for Medical Ultrasound(国家级区域医疗超声关键技术工程实验室)
;
School of Biomedical Engineering(生物医学工程学院)
;
Medical School(医学院)
;
Shenzhen University(深圳大学)
;
Medical UltraSound Image Computing (MUSIC) Lab(医学超声图像计算(MUSIC)实验室)
;
School of Artificial Intelligence(人工智能学院)
;
The Hong Kong University of Science and Technology(香港科学与技术大学)
;
Department of Electronic and Computer Engineering(电子与计算机工程系)
;
The University of Manchester(曼彻斯特大学)
;
The Third Affiliated Hospital of Sun Yat-sen University(中山大学第三附属医院)
;
Longgang District People’s Hospital of Shenzhen(深圳龙岗区人民医院)
;
The Second Affiliated Hospital of The Chinese University of Hong Kong(香港中文大学第二附属医院)
;
School of Biomedical Engineering and Informatics(生物医学工程与信息学学院)
;
NIHR Manchester Biomedical Research Centre(NIHR曼彻斯特生物医学研究中心)