机构
*
State Key Laboratory for Multimedia Information Processing, School of Computer Science, PKU-Anker LLM Lab, Peking University(信息处理国家重点实验室,计算机学院,PKU-Anker LLM实验室,北京大学)
;
University of California, Los Angeles(加州大学洛杉矶分校)
;
University of Washington(华盛顿大学)
;
Georgia Institute of Technology(佐治亚理工学院)
;
Nanyang Technological University(南洋理工大学)
;
HKUST(香港科技大学)
;
University of International Business and Economics(国际商务经济大学)
专题命中
效率与部署
:large language model(title,abstract);language model(title,abstract);LLM(abstract);post-training(abstract)
ALMGuard: Safety Shortcuts and Where to Find Them as Guardrails for Audio-Language Models
Weifei Jin, Yuxin Cao, Junjie Su, Minhui Xue, Jie Hao, Ke Xu, Jin Song Dong, Derui Wang
机构
*
Beijing University of Posts and Telecommunications(北京邮电大学)
;
National University of Singapore(新加坡国立大学)
;
CSIRO’s Data61(CSIRO数据61)
;
Responsible AI Research (RAIR) Centre, The University of Adelaide(负责任人工智能研究(RAIR)中心,阿德莱德大学)
;
Tsinghua University(清华大学)
专题命中
效率与部署
:language model(title,abstract);LLM(abstract);large language model(abstract);分类 cs.LG
HyGen: Efficient LLM Serving via Elastic Online-Offline Request Co-location
Ting Sun, Penghan Wang, Fan Lai
机构
*
Siebel School of Computing and Data Science, University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校计算机与数据科学学院)
;
Department of Computer Science, Purdue University(普渡大学计算机科学系)
专题命中
效率与部署
:LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.LG
Comments7 pages, 7 figures and tables, Published in Proceedings of the BabyLM Challenge 2025
Journal refIn Proceedings of the 2nd BabyLM Challenge at the 28th Conference on Computational Natural Language Learning, CoNLL 2024, pages 82 to 94, Miami, FL, USA. Association for Computational Linguistics
专题命中
效率与部署
:LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI
CommentsAccepted for presentation at the IEEE BigData 2025 Workshop (Special Session on Intelligent Data Mining). This v2 updates formatting and adds IEEE copyright notice
RelP: Faithful and Efficient Circuit Discovery in Language Models via Relevance Patching
Farnoush Rezaei Jafari, Oliver Eberle, Ashkan Khakzar, Neel Nanda
机构
*
Machine Learning Group, Technische Universität Berlin(柏林技术大学机器学习组)
;
BIFOLD – Berlin Institute for the Foundations of Learning and Data(柏林学习与数据基础研究院)