No Place to Hide: Benchmarking Video Hallucination with Background-Controlled Pairs
无处可藏:基于背景控制对的视频幻觉基准测试
Haojian Huang, Harold Haodong Chen, Meng Luo, Junjia Du, Shanqing Xu, Ziheng Chen, Yanxiang Huang, Yinchuan Li, Ying-Cong Chen
机构
*
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
Knowin AI(考恩人工智能)
;
The Hong Kong Polytechnic University(香港理工大学)
;
National University of Singapore(新加坡国立大学)
;
Nanyang Technological University(南洋理工大学)
;
Huazhong University of Science and Technology(华中科技大学)
;
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
RESOLVE: A Multi-Resolution and Multi-Modal Dataset for Roadside Cooperative Perception
RESOLVE:用于路边协同感知的多分辨率多模态数据集
Shaozu Ding, Linan Song, Marco De Vincenzi, Dajiang Suo
机构
*
The Polytechnic School, Arizona State University(亚利桑那州立大学理工学院)
;
Department of Computer Science and Engineering, New York University(纽约大学计算机科学与工程系)
MuSViT: A Foundation Vision Model for Sheet Music Representation
MuSViT: 一种用于乐谱表示的基础视觉模型
Carlos Penarrubia, Antonio Rios-Vila, Eliseo Fuentes-Martinez, Juan C. Martinez-Sevilla, Francisco J. Castellanos, María Alfaro-Contreras, Jorge Calvo-Zaragoza
机构
*
Pattern Recognition and Artificial Intelligence Group, University of Alicante, Spain(西班牙阿利坎特大学模式识别与人工智能组)
Intrinsically Stable Spiking Neural Networks: Overcoming the Performance Barrier in the Absence of Batch Normalization
内在稳定的脉冲神经网络:克服无批归一化下的性能障碍
Ruichen Ma, Xiaoyang Zhang, Jian Bai, Guanchao Qiao, Liwei Meng, Ning Ning, Yang Liu, Shaogang Hu
机构
*
University of Electronic Science and Technology of China (UESTC)(电子科技大学)
;
Beijing Institute of Remote-Sensing Equipment(北京遥感设备研究所)
;
Shenzhen Institute for Advanced Study, UESTC(电子科技大学深圳高等研究院)
Sparsity-Inducing Divergence Losses for Biometric Verification
用于生物特征验证的稀疏诱导散度损失
Dimitrios Koutsianos, Ladislav Mošner, Yannis Panagakis, Themos Stafylakis
机构
*
Athens University of Economics and Business(雅典经济与商业大学)
;
Archimedes/Athena Research Center(阿基米德/雅典娜研究中心)
;
Brno University of Technology(布尔诺理工大学)
;
Omilia
;
National and Kapodistrian University of Athens(雅典大学)
机构
*
Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院)
;
SparcAI Inc(SparcAI公司)
;
University of Science and Technology of China(中国科学技术大学)
;
The Hong Kong Polytechnic University(香港理工大学)
;
University of Electronic Science and Technology of China(电子科技大学)
;
Nanyang Technological University(南洋理工大学)
;
Institute for AI Industry Research (AIR), Tsinghua University(清华大学人工智能产业研究院)
Revisiting Parameter Redundancy in Vision-Language-Action Models: Insights from VLM-to-VLA Adaptation
重新审视视觉-语言-动作模型中的参数冗余:从VLM到VLA适配的见解
Fengnian Zhang, Tao Huang, Siyu Xu, Zhong Jin, Chang Xu
机构
*
Computer Network Information Center, Chinese Academy of Sciences(中国科学院计算机网络信息中心)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
School of Computer Science, Shanghai Jiao Tong University(上海交通大学计算机科学与工程学院)
;
School of Computer Science, The University of Sydney(悉尼大学计算机科学学院)
UHD-MFF: Shattering Barriers in Multi-Focus Ultra-High-Definition Image Fusion via Learnable Lookup Tables
UHD-MFF:通过可学习查找表打破多焦点超高清图像融合的障碍
Yibing Zhang, Xunpeng Yi, Qinglong Yan, Yeda Wang, Han Xu, Jiayi Ma
机构
*
Electronic Information School, Wuhan University(武汉大学电子信息学院)
;
School of Robotics, Wuhan University(武汉大学机器人学院)
;
School of Automation, Southeast University(东南大学自动化学院)
AC3S: Adaptive Conditioning for 3D-Aware Synthetic Data Generation
AC3S: 面向3D感知合成数据生成的自适应条件控制
Eric Ji, Qiran Hu, Wufei Ma, Sarthak Jain, Yingying Li, Minh N. Do, Yaoyao Liu
机构
*
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
Johns Hopkins University(约翰霍普金斯大学)
;
College of Engineering and Computer Science, VinUniversity(VinUniversity工程与计算机科学学院)
机构
*
School of Electrical Engineering, Korea University(韩国大学电气工程学院)
;
Department of Hospital Pathology, The Catholic University of Korea College of Medicine(韩国天主教大学医学院医院病理学系)