机构
*
The Hong Kong Polytechnic University(香港理工大学)
;
Tsinghua University(清华大学)
;
InspireOmni AI
;
Alibaba Group(阿里巴巴集团)
;
Case Western Reserve University(凯斯西储大学)
专题命中
VLM训练与架构
:VLM(title,abstract);vision language model(abstract);分类 cs.CV
No Tokens Wasted: Leveraging Long Context in Biomedical Vision-Language Models
Min Woo Sun, Alejandro Lozano, Javier Gamazo Tejero, Vishwesh Nath, Xiao Xiao Sun, James Burgess, Yuhui Zhang, Kun Yuan, Robert Tibshirani, Sean Huver, Serena Yeung-Levy
The Security Threat of Compressed Projectors in Large Vision-Language Models
Yudong Zhang, Ruobing Xie, Xingwu Sun, Jiansheng Chen, Zhanhui Kang, Di Wang, Yu Wang
机构
*
Department of Electronic Engineering, Tsinghua University(清华大学电子工程系)
;
Large Language Model Department, Tencent(腾讯大语言模型部门)
;
School of Computer and Communication Engineering, University of Science and Technology Beijing(北京科技大学计算机与通信工程学院)
;
Faculty of Science and Technology, University of Macau(澳门大学科技学院)
专题命中
VLM训练与架构
:vision-language model(title);visual language model(abstract);分类 cs.AI
AutoPrune: Each Complexity Deserves a Pruning Policy
Hanshi Wang, Yuhao Xu, Zekun Xu, Jin Gao, Yufan Liu, Weiming Hu, Ke Wang, Zhipeng Zhang
机构
*
State Key Laboratory of Multimodal Artificial Intelligence Systems (MAIS), CASIA(多模态人工智能系统国家重点实验室(MAIS),中国科学院自动化所)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
AutoLab, School of Artificial Intelligence, Shanghai Jiao Tong University(上海交通大学人工智能学院AutoLab)
;
Anyverse Intelligence
;
Beijing Key Laboratory of Super Intelligent Security of Multi-Modal Information(北京多模态信息超智能安全重点实验室)
;
School of Information Science and Technology, ShanghaiTech University(上海科技大学信息科学与技术学院)