Barycentric alignment for instance-level comparison of neural representations
实例层面神经表示的重心对齐
Shreya Saha, Zoe Wanying He, Meenakshi Khosla
机构
*
Department of XXX, University of YYY, Location, Country(YYY大学XXX系)
;
School of ZZZ, Institute of WWW, Location, Country(WWW研究所ZZZ学院)
;
Department of Electrical and Computer Engineering, UCSD(UCSD电子与计算机工程系)
;
Cognitive Science Department, UCSD(UCSD认知科学系)
;
Department of Computer Science(计算机科学系)
机构
*
Center for Data Science, Peking University(北京大学数据科学中心)
;
Alibaba Group(阿里巴巴集团)
;
CASIA
;
Center for Machine Learning Research, Peking University(北京大学机器学习研究中心)
;
State Key Laboratory of General Artificial Intelligence, Peking University(北京大学通用人工智能国家重点实验室)
专题命中
VLM训练与架构
:multimodal large language model(abstract);分类 cs.CV
机构
*
Institute for Artificial Intelligence, Peking University
;
School of Software \& Microelectronics, Peking University
;
School of Computer Science, Peking University
;
School of Life Sciences, Peking University
;
Department of Comprehensive Oncology, National Cancer Center/National Clinical Research Center for Cancer/Cancer Hospital, Chinese Academy of Medical Sciences
;
Peking Union Medical College Beijing, China
LoVR: A Benchmark for Long Video Retrieval in Multimodal Contexts
LoVR:一种多模态背景下长视频检索的基准
Qifeng Cai, Hao Liang, Zhaoyang Han, Hejun Dong, Meiyi Qiang, Ruichuan An, Quanqing Xu, Bin Cui, Wentao Zhang
机构
*
East China Normal University Shanghai China
;
Peking University \& Zhongguancun Academy Beijing China
;
Huazhong University of Science
;
Beihang University Beijing China
;
Peking University Beijing China
;
East China Normal University
;
Peking University \& Zhongguancun Academy
;
Beihang University
;
Peking University
FastPhysGS: Accelerating Physics-based Dynamic 3DGS Simulation via Interior Completion and Adaptive Optimization
FastPhysGS: 通过内部补全与自适应优化加速基于物理的动态3DGS仿真
Yikun Ma, Yiqing Li, Jingwen Ye, Zhongkai Wu, Weidong Zhang, Lin Gao, Zhi Jin
机构
*
School of Intelligent Systems Engineering, Shenzhen Campus of Sun Yat-sen University(中山大学智能系统工程学院)
;
Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Guangdong Provincial Key Laboratory of Fire Science(广东省消防科学与智能应急技术重点实验室)
;
Guangdong Provincial Key Laboratory of Robotics(广东省机器人与数字智能制造技术重点实验室)
机构
*
Peking University(北京大学)
;
Central Conservatory of Music(中央音乐学院)
;
The Chinese University of Hong Kong(香港中文大学)
;
University of Electronic Science and Technology of China(电子科技大学)
专题命中
VLM训练与架构
:multimodal large language model(abstract);分类 cs.AI
机构
*
Department of Computer Science and Engineering, The Chinese University of Hong Kong(计算机科学与工程系,香港中文大学)
;
Institute of Medical Intelligence and XR, The Chinese University of Hong Kong(医学智能与XR研究院,香港中文大学)
;
Department of Computer Science, Khalifa University(计算机科学系,哈利法大学)
Knowledge-enhanced Pretraining for Vision-language Pathology Foundation Model on Cancer Diagnosis
基于知识增强的视觉语言病理基础模型用于癌症诊断
Xiao Zhou, Luoyi Sun, Dexuan He, Wenbin Guan, Ge Wang, Ruifen Wang, Lifeng Wang, Xiaojun Yuan, Xin Sun, Ya Zhang, Kun Sun, Yanfeng Wang, Weidi Xie
机构
*
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
Shanghai Jiao Tong University(上海交通大学)
;
Xinhua Hospital Affiliated to Shanghai Jiao Tong University School of Medicine(上海交通大学医学院附属新华医院)
;
School of Artificial Intelligence(人工智能学院)
;
Department of Pathology(病理学部)
;
Department of Oral Pathology(口腔病理学部)
;
Department of Pediatric Hematology/Oncology(儿科血液肿瘤科)
;
Clinical Research and Innovation Unit(临床研究与创新单元)
;
Department of Pediatric Cardiology(儿童心脏病科)
DTP: A Simple yet Effective Distracting Token Pruning Framework for Vision-Language Action Models
DTP: 一种简单而有效的干扰令牌修剪框架用于视觉-语言动作模型
Chenyang Li, Jieyuan Liu, Bin Li, Bo Gao, Yilin Yuan, Yangfan He, Yuchen Li, Jingqun Tang
机构
*
Australian National University(澳大利亚国立大学)
;
University of California, San Diego(加州大学圣地亚哥分校)
;
Chinese Academy of Sciences(中国科学院)
;
Beijing Institute of Graphic Communication(北京印刷学院)
;
University of North Carolina at Chapel Hill(北卡罗来纳大学教堂山分校)
;
Baidu Search(百度搜索)
;
Bytedance(字节跳动)
Can Synthetic Images Serve as Effective and Efficient Class Prototypes?
合成图像能否作为有效的高效类别原型?
Dianxing Shi, Dingjie Fu, Yuqiao Liu, Jun Wang
机构
*
Beijing Research Institute of Uranium Geology(铀矿北京研究机构)
;
Huazhong University of Science and Technology(华中科技大学)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
Great Bay University(大湾大学)
机构
*
Orion Lab, School of Rural, Surveying and Geoinformatics Engineering, National Technical University of Athens(奥里昂实验室,农村测绘与地理信息工程学院,希腊雅典国家技术大学)
;
Institute of Astronomy, Astrophysics, Space Applications and Remote Sensing, National Observatory of Athens(天文、天体物理、空间应用与遥感研究所,希腊雅典国家天文台)
;
Department of Informatics and Telematics, Harokopio University of Athens(信息与电信技术系,雅典霍罗科皮欧大学)
;
Chair of Data Science in Earth Observation, Technical University of Munich(地球观测数据科学教授职位,慕尼黑技术大学)
;
Munich Center for Machine Learning(慕尼黑机器学习中心)
;
Faculty of Electrical Engineering and Computer Science, Technische Universität Berlin(电气工程与计算机科学系,柏林技术大学)
;
BIFOLD - Berlin Institute for the Foundations of Learning and Data(柏林学习与数据基础研究所)