SAB-LVLM: Significance-Aware Binarization for Large Vision-Language Models
SAB-LVLM: 面向大型视觉-语言模型的重要性感知二值化
Qi Lyu, Jiahua Dong, Baichen Liu, Xudong Wang, Mingfei Han, Yulun Zhang, Fahad Shahbaz Khan, Salman Khan, Lianqing Liu, Zhi Han
机构
*
State Key Laboratory of Robotics and Intelligent Systems(机器人学国家重点实验室)
;
Shenyang Institute of Automation, Chinese Academy of Sciences(中国科学院沈阳自动化研究所)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)
;
Shanghai Jiao Tong University(上海交通大学)
Minimizing Quantized Semantic Age of Information (QSAoI) in Foundation Model-Based Semantic Communications
最小化基于基础模型的语义通信中的量化语义信息年龄(QSAoI)
Huanyu Zhang, Yulin Hu, Xiaopeng Yuan, Aydin Sezgin, Anke Schmeink
机构
*
INDA Chair, RWTH Aachen University(亚琛工业大学INDA教席)
;
School of Electronic Information, Wuhan University(武汉大学电子信息学院)
;
Department of Digital Communication Systems, Ruhr University Bochum(波鸿鲁尔大学数字通信系统系)
Efficient and Trainable Language Model Test-Time Scaling via Local Branch Routing
通过局部分支路由实现高效且可训练的语言模型测试时扩展
Yutong Yin, Mingyu Jin, Jin Pan, Changyi Yang, Zijie Xia, Dhruv Pai, Shuming Hu, Zhen Zhang, Chenyang Zhao, Jinman Zhao, Wujiang Xu, Raymond Li, Xin Eric Wang, Julian McAuley, Zhaoran Wang
机构
*
Northwestern University(西北大学)
;
Rutgers University(罗格斯大学)
;
University of Wisconsin–Madison(威斯康星大学麦迪逊分校)
;
Carnegie Mellon University(卡内基梅隆大学)
;
LMSYS Org(LMSYS组织)
;
Tilde Research(Tilde研究)
;
University of California, Santa Barbara(加州大学圣塔芭芭拉分校)
;
University of Toronto(多伦多大学)
;
University of British Columbia(不列颠哥伦比亚大学)
;
University of California, San Diego(加州大学圣迭戈分校)
机构
*
School of Computer Science, University of Nottingham, UK(英国诺丁汉大学计算机科学学院)
;
School of Computer Science, University of Nottingham Ningbo China, China(中国宁波诺丁汉大学计算机科学学院)
;
School of Engineering and Physical Science, University of Lincoln, UK(英国林肯大学工程与物理科学学院)
;
Department of Computer Science, University of Rochester, USA(美国罗切斯特大学计算机科学系)
Translating Inference-Time Control to Radiology Vision-Language Models: Activation Steering for Pneumonia Classification on Chest X-rays
将推理时控制转化为放射学视觉语言模型:针对胸部X光片肺炎分类的激活引导
Eduardo Moreno Judice de Mattos Farina, Mateus A. Esmeraldo, Felipe Akio Matsuoka, Paulo Eduardo de Aguiar Kuriki, Felipe Campos Kitamura
机构
*
Universidade Federal de São Paulo (UNIFESP)(圣保罗联邦大学)
;
Hospital Israelita Albert Einstein(以色列阿尔伯特·爱因斯坦医院)
;
Stanford University School of Medicine(斯坦福大学医学院)
;
DASA
;
University of Texas Southwestern Medical Center (UTSW)(德克萨斯大学西南医学中心)
;
Eden
Mutual Distillation of Dual-Foundation Models for Semi-Supervised PET/CT Segmentation
双基础模型的相互蒸馏用于半监督PET/CT分割
Fuyou Mao, Beining Wu, Yanfeng Jiang, Bohan Xu, Lixin Lin, Naye Ji, Hao Zhang, Yan Tang
机构
*
Central South University(中南大学)
;
Hangzhou Dianzi University(杭州电子科技大学)
;
Communication University of Zhejiang(浙江传媒学院)
;
Northeastern University(东北大学)
Vision-Assisted Foundation Model for Solving Multi-Task Vehicle Routing Problems
视觉辅助的基础模型解决多任务车辆路径问题
Shuangchun Gui, Zhiguang Cao, Wen Song, Yew-Soon Ong
机构
*
School of Computing and Information Systems, Singapore Management University(新加坡管理大学计算与信息系统学院)
;
Institute of Marine Science and Technology, Shandong University(山东大学海洋科学与技术研究院)
;
College of Computing and Data Science, Nanyang Technological University(南洋理工大学计算与数据科学学院)
;
Centre for Frontier AI Research, Institute of High Performance Computing, Agency for Science, Technology and Research(新加坡科技研究局高性能计算研究所前沿人工智能研究中心)
机构
*
University of Science and Technology of China(中国科学技术大学)
;
Institute of Artificial Intelligence, Hefei Comprehensive National Science Center(合肥综合性国家科学中心人工智能研究院)
;
Shanxi University(山西大学)
Towards Understanding On-Policy Distillation through the Lens of Test-Time Scaling
从测试时缩放的视角理解在线策略蒸馏
Xinmu Ge, Zizhuo Zhang, Yu Huang, Jianing Zhu, Lin Yuan, Wanli Gu, Weichang Wu, Weiran Huang, Xiaolu Zhang, Bo Han, Jun Zhou, Jiangchao Yao
机构
*
Hong Kong Baptist University(香港浸会大学)
;
University of Texas at Austin(德克萨斯大学奥斯汀分校)
;
Shanghai Jiao Tong University(上海交通大学)
;
Shanghai Innovation Institute(上海创新研究院)
;
Ant Group(蚂蚁集团)
Comments21 pages, 8 figures, 4 tables. Open-source harness, scenario bank, proof texts, and full evaluation (model responses and both judges' verdicts): https://github.com/iaser-ai/jaleesbench . Interactive browser for inspecting scenarios, responses, and judge verdicts: https://s.iaser.ai/jb