Comments17 pages; Previously this version appeared as arXiv:2510.15430 (https://arxiv.org/abs/2510.15430) which was submitted as a new work by accident
机构
*
State Key Laboratory of Novel Software Technology, Nanjing University, China(南京大学新型软件技术国家重点实验室)
;
School of Artificial Intelligence, Nanjing University, China(南京大学人工智能学院)
专题命中
VLM训练与架构
:grounding(abstract);multimodal large language model(abstract);分类 cs.AI、cs.LG
Label-Free Foundational Model Selection for Medical Image Classification under Distribution Shift via Pseudo Label Discrepancy
分布偏移下基于伪标签差异的医学图像分类无基准模型选择
Juan Iñaki Larrea, Lucas Mansilla, Enzo Ferrante
机构
*
Universidad de Buenos Aires(布宜诺斯艾利斯大学)
;
CONICET(阿根廷国家科学技术研究委员会)
;
Universidad Nacional del Litoral(国立 littoral 大学)
;
Research Institute for Signals, Systems and Computational Intelligence(信号、系统与计算智能研究所)
Training-Free Interaction-Aligned Visual Token Pruning for Efficient Embodied Manipulation
VLA-IAP: 通过交互对齐实现无训练视觉令牌剪枝用于视觉-语言-动作模型
Jintao Cheng, Weibin Li, Haozhe Wang, Gang Wang, Yipu Zhang, Xiaoyu Tang, Jin Wu, Xieyuanli Chen, Yunhui Liu, Wei Zhang
机构
*
Hong Kong University of Science and Technology(香港科学与技术大学)
;
South China Normal University(华南师范大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
University of Science and Technology Beijing(北京科技大学)
;
National University of Defense Technology(国防科技大学)
机构
*
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
Zhongguancun Academy(中关村学院)
专题命中
其他VLM
:multimodal large language model(abstract);分类 cs.AI、cs.LG