arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-10-24 至 2025-10-24 共收录 13 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 13 篇

2510.20039 2025-10-24 cs.HC cs.AI cs.CL cs.CY 86%

Beyond One-Way Influence: Bidirectional Opinion Dynamics in Multi-Turn Human-LLM Interactions

Yuyang Jiang, Longjie Guo, Yuchen Wu, Aylin Caliskan, Tanu Mitra, Hua Shen

机构 * University of Chicago(芝加哥大学) New York University(纽约大学) University of Washington(华盛顿大学) New York University Shanghai(纽约大学上海)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments 26 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.04510 2025-10-24 cs.CL 83%

Heterogeneous Swarms: Jointly Optimizing Model Roles and Weights for Multi-LLM Systems

Shangbin Feng, Zifeng Wang, Palash Goyal, Yike Wang, Weijia Shi, Huang Xia, Hamid Palangi, Luke Zettlemoyer, Yulia Tsvetkov, Chen-Yu Lee, Tomas Pfister

机构 * University of Washington(华盛顿大学) Google Cloud AI Research(谷歌云人工智能研究)

专题命中 其他LLM :LLM(title,abstract);language model(abstract);分类 cs.CL

Comments NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.11564 2025-10-24 cs.LG 79%

Continuous Diffusion Model for Language Modeling

Jaehyeong Jo, Sung Ju Hwang

机构 * KAIST(韩国科学技术院)

专题命中 其他LLM :language model(title,abstract);分类 cs.LG

Comments NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.20612 2025-10-24 cs.CY cs.AI cs.CR cs.LG econ.GN q-fin.EC 79%

Black Box Absorption: LLMs Undermining Innovative Ideas

Wenjun Cao

机构 * Independent Researcher(独立研究者)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.20187 2025-10-24 cs.LG cs.CL 79%

Every Question Has Its Own Value: Reinforcement Learning with Explicit Human Values

Dian Yu, Yulai Zhao, Kishan Panaganti, Linfeng Song, Haitao Mi, Dong Yu

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

Comments 15 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.01002 2025-10-24 cs.CL cs.AI 79%

Token embeddings violate the manifold hypothesis

Michael Robinson, Sourya Dey, Tony Chiang

机构 * Mathematics and Statistics(数学与统计学) American University(美国大学) Galois, Inc.(Galois公司) Department of Mathematics, University of Washington(华盛顿大学数学系)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments 30 pages, 9 figures, 10 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.20287 2025-10-24 cs.CV cs.AI cs.LG 73%

Breakdance Video classification in the age of Generative AI

Sauptik Dhar, Naveen Ramakrishnan, Michelle Munson

机构 * Eluvio AI Labs(Eluvio AI实验室)

专题命中 其他LLM :language model(abstract);foundation model(abstract);分类 cs.AI、cs.LG

Comments 11 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.20817 2025-10-24 cs.LG 70%

KL-Regularized Reinforcement Learning is Designed to Mode Collapse

Anthony GX-Chen, Jatin Prakash, Jeff Guo, Rob Fergus, Rajesh Ranganath

机构 * New York University(纽约大学) École Polytechnique Fédérale de Lausanne(瑞士联邦理工学院)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.20190 2025-10-24 cs.AI cs.IT math.IT 70%

The Lock-In Phase Hypothesis: Identity Consolidation as a Precursor to AGI

Marcelo Maciel Amaral, Raymond Aschheim

机构 * Gauge Freedom, Inc. (Public Benefit Corporation)(Gauge Freedom公司)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.20733 2025-10-24 cs.LG cs.AI cs.MA 62%

Thought Communication in Multiagent Collaboration

Yujia Zheng, Zhuokai Zhao, Zijian Li, Yaqi Xie, Mingze Gao, Lizhu Zhang, Kun Zhang

机构 * CMU(卡内基梅隆大学) Meta AI MBZUAI(马克斯·普朗克人工智能研究所)

专题命中 其他LLM :LLM(abstract);分类 cs.AI、cs.LG

Comments NeurIPS 2025 Spotlight

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.23946 2025-10-24 cs.AI cs.LG cs.MA cs.SE 62%

Lessons Learned: A Multi-Agent Framework for Code LLMs to Learn and Improve

Yuanzhe Liu, Ryan Deng, Tim Kaler, Xuhao Chen, Charles E. Leiserson, Yao Ma, Jie Chen

机构 * Rensselaer Polytechnic Institute(拉特兰理工学院) Massachusetts Institute of Technology(麻省理工学院) Michigan State University(密歇根州立大学) MIT-IBM Watson AI Lab, IBM Research(MIT-IBM沃森人工智能实验室,IBM研究院)

专题命中 其他LLM :LLM(abstract);分类 cs.AI、cs.LG

Comments NeurIPS 2025. Code is available at https://github.com/MITIBM-FastCoder/LessonL

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.20244 2025-10-24 cs.CV cs.LG 57%

Empower Words: DualGround for Structured Phrase and Sentence-Level Temporal Grounding

Minseok Kang, Minhyeok Lee, Minjung Kim, Donghyeong Kim, Sangyoun Lee

机构 * Yonsei University(延世大学) LG Electronics(LG电子)

专题命中 其他LLM :language model(abstract);分类 cs.LG

Comments Comments: 28 pages, including appendix. 5 figures. Full version of the NeurIPS 2025 paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.19335 2025-10-24 cs.LG 57%

Gatekeeper: Improving Model Cascades Through Confidence Tuning

Stephan Rabanser, Nathalie Rauschmayr, Achin Kulshrestha, Petra Poklukar, Wittawat Jitkrittum, Sean Augenstein, Congchao Wang, Federico Tombari

机构 * Princeton University(普林斯顿大学) Google(谷歌)

专题命中 其他LLM :language model(abstract);分类 cs.LG

Comments Presented at the TTODLer-FM workshop at the International Conference on Machine Learning (ICML) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏