From Bits to Chips: An LLM-based Hardware-Aware Quantization Agent for Streamlined Deployment of LLMs
从比特到芯片:一种基于LLM的硬件感知量化代理,用于简化LLM的部署
Kaiyuan Deng, Hangyu Zheng, Minghai Qing, Kunxiong Zhu, Gen Li, Yang Xiao, Lan Emily Zhang, Linke Guo, Bo Hui, Yanzhi Wang, Geng Yuan, Gagan Agrawal, Wei Niu, Xiaolong Ma
机构
*
The University of Arizona(亚利桑那大学)
;
Clemson University(克莱姆森大学)
;
University of Georgia(佐治亚大学)
;
The University of Tulsa(塔尔萨大学)
;
Northeastern University(东北大学)
;
Western Digital Corporation(西部数据公司)
专题命中
效率与部署
:LLM(title);large language model(abstract);language model(abstract);分类 cs.LG
Safer by Diffusion, Broken by Context: Diffusion LLM's Safety Blessing and Its Failure Mode
扩散模型更安全,但受上下文影响而失效:扩散大语言模型的安全祝福及其失效模式
Zeyuan He, Yupeng Chen, Lang Lin, Yihan Wang, Shenxu Chang, Eric Sommerlade, Philip Torr, Junchi Yu, Adel Bibi, Jialin Yu
机构
*
Torr Vision Group, University of Oxford(牛津大学Torr视觉组)
;
School of Data Science, The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳)数据科学学院)
;
The University of Texas at Austin(德克萨斯大学奥斯汀分校)
;
Microsoft(微软)
专题命中
效率与部署
:LLM(title);large language model(abstract);language model(abstract);分类 cs.LG
FEMBA on the Edge: Physiologically-Aware Pre-Training, Quantization, and Deployment of a Bidirectional Mamba EEG Foundation Model on an Ultra-low Power Microcontroller
DualEdit: Mitigating Safety Fallback in LLM Backdoor Editing via Affirmation-Refusal Regulation
DualEdit: 通过肯定-拒绝调节缓解LLM后门编辑中的安全回退
Houcheng Jiang, Zetong Zhao, Junfeng Fang, Haokai Ma, Ruipeng Wang, Xiang Wang, Xiangnan He, Yang Deng
机构
*
University of Science and Technology of China(中国科学技术大学)
;
National University of Singapore(新加坡国立大学)
;
Singapore Management University(新加坡管理大学)
;
MoE Key Lab of BIPC, University of Science and Technology of China(中国科学技术大学信息与通信技术国家重点实验室)
专题命中
效率与部署
:LLM(title);large language model(abstract);language model(abstract);分类 cs.CL
LLM-Enhanced Rumor Detection via Virtual Node Induced Edge Prediction
通过虚拟节点诱导边预测增强的谣言检测
Jiran Tao, Cheng Wang, Binyan Jiang
机构
*
Department of Data Science and Artificial Intelligence(数据科学与人工智能系)
;
Hong Kong Polytechnic University(香港理工大学)
;
School of Mathematical Sciences(数学科学学院)
;
Shanghai Jiao Tong University(上海交通大学)
专题命中
效率与部署
:LLM(title);large language model(abstract);language model(abstract);分类 cs.AI
Semantic Self-Distillation for Language Model Uncertainty
语义自蒸馏用于语言模型不确定性
Edward Phillips, Sean Wu, Fredrik K. Gustafsson, Boyan Gao, David A. Clifton
机构
*
Department of Engineering Science University of Oxford(工程科学系牛津大学)
;
University of Oxford(牛津大学)
;
Oxford Suzhou Centre for Advanced Research University of Oxford Suzhou(牛津苏州先进研究中心)
专题命中
效率与部署
:language model(title,abstract);large language model(abstract);分类 cs.CL
Learn from Foundation Model: Fruit Detection Model without Manual Annotation
从基础模型学习:无需人工标注的水果检测模型
Yanan Wang, Zhenghao Fei, Ruichen Li, Yibin Ying
机构
*
College of Biosystems Engineering and Food Science, Zhejiang University(浙江大学生物系统工程与食品科学学院)
;
ZJU-Hangzhou Global Scientific and Technological Innovation Center, Zhejiang University(浙江大学杭州全球科技创新中心)
;
Faculty of Science, National University of Singapore(新加坡国立大学科学学院)
MAD: Microenvironment-Aware Distillation -- A Pretraining Strategy for Virtual Spatial Omics from Microscopy
MAD:微环境感知蒸馏——一种从显微镜获取虚拟空间组学的预训练策略
Jiashu Han, Kunzan Liu, Yeojin Kim, Saurabh Sinha, Sixian You
机构
*
Research Laboratory of Electronics, Massachusetts Institute of Technology(麻省理工学院电子研究实验室)
;
Department of Electrical Engineering and Computer Science, Massachusetts Institute of Technology(麻省理工学院电气工程与计算机科学系)
;
The Wallace H. Coulter Department of Biomedical Engineering, Georgia Institute of Technology(佐治亚理工学院Wallace H. Coulter生物医学工程系)
;
H. Milton Stewart School of Industrial & Systems Engineering, Georgia Institute of Technology(佐治亚理工学院H. Milton Stewart工业与系统工程学院)