Distilling Safe LLM Systems via Soft Prompts for On Device Settings
通过软提示蒸馏安全的设备端LLM系统
机构 * Qualcomm AI Research(高通人工智能研究院)
AI总结 针对资源受限设备上部署安全大语言模型(LLM)的挑战,提出基于软提示与蒸馏训练的安全对齐方法,在最小化额外计算开销的同时实现优越的安全-有用性权衡。
Comments Accepted to UAI 2026
Journal ref 42nd Conference on Uncertainty in Artificial Intelligence 2026