arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-10-06 至 2025-10-06 共收录 7 信号源:cs.CL, cs.AI, cs.LG

1. 后训练与偏好优化 7 篇

2412.07192 2025-10-06 cs.CR cs.CL cs.LG 91%

PrisonBreak: Jailbreaking Large Language Models with at Most Twenty-Five Targeted Bit-flips

Zachary Coalson, Jeonghyun Woo, Chris S. Lin, Joyce Qu, Yu Sun, Shiyang Chen, Lishan Yang, Gururaj Saileshwar, Prashant Nair, Bo Fang, Sanghyun Hong

机构 * Oregon State University(俄勒冈州立大学) University of British Columbia(不列颠哥伦比亚大学) University of Toronto(多伦多大学) George Mason University(乔治·梅森大学) Rutgers University(罗格斯大学) University of Texas at Arlington(德克萨斯大学阿灵顿分校)

专题命中 后训练与偏好优化 :large language model(title,abstract);language model(title,abstract);LLM(abstract);post-training(abstract)

Comments Pre-print

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.02645 2025-10-06 cs.CL 87%

Mind the Gap: Linguistic Divergence and Adaptation Strategies in Human-LLM Assistant vs. Human-Human Interactions

Fulei Zhang, Zhou Yu

机构 * Amazon.com Inc.(亚马逊公司)

专题命中 后训练与偏好优化 :LLM(title,abstract);large language model(abstract);language model(abstract);post-training(abstract)

Comments Accepted to The Second Workshop on Generative AI for E-commerce (GenAIECommerce '25), held September 22, 2025, in Prague, Czech Republic

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.02611 2025-10-06 cs.AI cs.CL cs.LG 80%

On the Role of Temperature Sampling in Test-Time Scaling

Yuheng Wu, Azalia Mirhoseini, Thierry Tambe

机构 * Stanford University(斯坦福大学)

专题命中 后训练与偏好优化 :large language model(abstract);language model(abstract);post-training(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.03231 2025-10-06 cs.CL cs.AI 79%

Reward Models are Metrics in a Trench Coat

Sebastian Gehrmann

机构 * Bloomberg(高盛)

专题命中 后训练与偏好优化 :large language model(abstract);language model(abstract);post-training(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.02850 2025-10-06 cs.AI 77%

Reward Model Routing in Alignment

Xinle Wu, Yao Lu

机构 * National University of Singapore(新加坡国立大学)

专题命中 后训练与偏好优化 :large language model(abstract);language model(abstract);RLHF(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.18314 2025-10-06 cs.CL 57%

Exploiting Tree Structure for Credit Assignment in RL Training of LLMs

Hieu Tran, Zonghai Yao, Hong Yu

机构 * Center for Healthcare Organization and Implementation Research, VA Bedford Health Care(VA Bedford Health Care 机构) Manning College of Information and Computer Sciences, University of Massachusetts Amherst(马萨诸塞大学阿默斯特分校信息与计算机科学学院) Miner School of Computer and Information Sciences, University of Massachusetts Lowell(马萨诸塞大学洛厄尔分校计算机与信息科学学院)

专题命中 后训练与偏好优化 :LLM(abstract);分类 cs.CL

Comments 15 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.10439 2025-10-06 cs.CV 50%

EFC++: Elastic Feature Consolidation with Prototype Re-balancing for Cold Start Exemplar-free Incremental Learning

Simone Magistri, Tomaso Trinci, Albin Soutif-Cormerais, Joost van de Weijer, Andrew D. Bagdanov

机构 * Communication Center (MICC), University of Florence, Italy Tomaso Trinci. ORCID: 0000-0002-4052-1930 Global Optimization Laboratory, University of Florence, Italy Albin Soutif - Cormerais. ORCID: 0009-0008-1564-2029 LAMP Team, Computer Vision Center, Barcelona, Spain Joost van de Weijer. ORCID: 0000-0002-9656-9706 LAMP Team, Computer Vision Center, Universitat Aut\` o noma de Barcelona, Spain Andrew D. Bagdanov. ORCID: 0000-0001-6408-7043 Media Integration Communication Center (MICC), University of Florence, Italy Equal contribution

专题命中 后训练与偏好优化 :post-training(abstract)

Comments Extension of our previous conference paper https://openreview.net/forum?id=7D9X2cFnt1

详情

展开后加载摘要…

URL PDF HTML 收藏