arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12287 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12287 篇

2409.17213 2025-03-25 cs.CL cs.AI cs.CY cs.HC cs.MA 62%

Plurals: A System for Guiding LLMs Via Simulated Social Ensembles

Joshua Ashkinaze, Emily Fry, Narendra Edara, Eric Gilbert, Ceren Budak

机构 * University of Michigan(密歇根大学) Oakland Community College(奥克兰社区学院)

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.AI

Comments CHI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.15451 2025-03-21 cs.LG cs.CL 62%

Binary-Integer-Programming Based Algorithm for Expert Load Balancing in Mixture-of-Experts Models

Yuan Sun

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.16635 2025-03-19 cs.CL cs.AI 62%

Why Do We Laugh? Annotation and Taxonomy Generation for Laughable Contexts in Spontaneous Text Conversation

Koji Inoue, Mikey Elmers, Divesh Lala, Tatsuya Kawahara

机构 * Graduate School of Informatics, Kyoto University(京都大学情报学研究科)

专题命中 其他LLM :LLM(abstract);分类 cs.CL、cs.AI

Comments This paper has been accepted for presentation at International Workshop on Spoken Dialogue Systems Technology 2025 (IWSDS 2025) and represents the author's version of the work

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.08041 2025-03-18 cs.LG cs.AI stat.ML 62%

The Clever Hans Effect in Unsupervised Learning

Jacob Kauffmann, Jonas Dippel, Lukas Ruff, Wojciech Samek, Klaus-Robert Müller, Grégoire Montavon

机构 * Technische Universität Berlin(柏林工业大学) BIFOLD – Berlin Institute for the Foundations of Learning and Data(柏林学习与数据基础研究所) Aignostics(艾格诺斯蒂克斯公司) Freie Universität Berlin(柏林自由大学) Fraunhofer HHI(弗劳恩霍夫海因里希·赫兹研究所) Korea University(高丽大学) Max-Planck Institute for Informatics(马克斯·普朗克信息学研究所) Google Deepmind(谷歌DeepMind)

专题命中 其他LLM :foundation model(abstract);分类 cs.AI、cs.LG

Comments 12 pages + supplement

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.05864 2025-03-18 cs.CL cs.CR cs.LG 62%

Permute-and-Flip: An optimally stable and watermarkable decoder for LLMs

Xuandong Zhao, Lei Li, Yu-Xiang Wang

机构 * UC Berkeley(加州大学伯克利分校) Carnegie Mellon University(卡内基梅隆大学) UC San Diego(加州大学圣迭戈分校)

专题命中 其他LLM :LLM(abstract);分类 cs.CL、cs.LG

Comments ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.10191 2025-03-14 cs.LG cs.AI cs.CV 62%

Robustness Tokens: Towards Adversarial Robustness of Transformers

Brian Pulfer, Yury Belousov, Slava Voloshynovskiy

机构 * University of Geneva(日内瓦大学)

专题命中 其他LLM :foundation model(abstract);分类 cs.AI、cs.LG

Comments This paper has been accepted for publication at the European Conference on Computer Vision (ECCV), 2024

Journal ref Computer Vision, ECCV 2024 pp 110 to 127, Springer Nature Switzerland

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06784 2025-03-11 cs.GR cs.AI cs.CV cs.LG cs.RO 62%

Infinite Leagues Under the Sea: Photorealistic 3D Underwater Terrain Generation by Latent Fractal Diffusion Models

Tianyi Zhang, Weiming Zhi, Joshua Mangelson, Matthew Johnson-Roberson

机构 * Carnegie Mellon University(卡内基梅隆大学) Brigham Young University(杨百翰大学)

专题命中 其他LLM :foundation model(abstract);分类 cs.AI、cs.LG

Comments 10 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.05870 2025-03-11 cs.CR cs.CL cs.LG 62%

Machine Against the RAG: Jamming Retrieval-Augmented Generation with Blocker Documents

Avital Shafran, Roei Schuster, Vitaly Shmatikov

机构 * The Hebrew University(希伯来大学) Wild Moose(野驼鹿(Wild Moose)) Cornell Tech(康奈尔科技学院)

专题命中 其他LLM :LLM(abstract);分类 cs.CL、cs.LG

Comments To appear in USENIX Security Symposium 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.15018 2025-03-11 cs.CL cs.AI 62%

Answer, Assemble, Ace: Understanding How LMs Answer Multiple Choice Questions

Sarah Wiegreffe, Oyvind Tafjord, Yonatan Belinkov, Hannaneh Hajishirzi, Ashish Sabharwal

机构 * Allen Institute for AI(艾伦人工智能研究所) University of Washington(华盛顿大学) Technion(以色列理工学院)

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.AI

Comments ICLR 2025 (spotlight). Substantially updated from previous preprint to contain experiments on 4-way multiple-choice with various answer choice symbols, 3 open model families, and extensive activation patching results, including on individual attention heads

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.19723 2025-03-10 cs.CL cs.AI 62%

CNsum:Automatic Summarization for Chinese News Text

Yu Zhao, Songping Huang, Dongsheng Zhou, Zhaoyun Ding, Fei Wang, Aixin Nian

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.AI

Comments This withdrawal is due to the lack of authorization from all co-authors for the publication of this version

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.01214 2025-03-06 cs.LG cs.AI 62%

Revisiting Random Walks for Learning on Graphs

Jinwoo Kim, Olga Zaghen, Ayhan Suleymanzade, Youngmin Ryou, Seunghoon Hong

机构 * KAIST(韩国科学技术院) University of Amsterdam(阿姆斯特丹大学)

专题命中 其他LLM :language model(abstract);分类 cs.AI、cs.LG

Comments 51 pages, 14 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.01669 2025-03-06 cs.CY cs.AI cs.LG 62%

Improved Performances and Motivation in Intelligent Tutoring Systems: Combining Machine Learning and Learner Choice

Benjamin Clément, Hélène Sauzéon, Didier Roy, Pierre-Yves Oudeyer

机构 * Inria(法国国家信息与自动化研究所) Université de Bordeaux(波尔多大学)

专题命中 其他LLM :prompting(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.16399 2025-02-25 cs.IR cs.AI cs.CL 62%

Ensemble ToT of LLMs and Its Application to Automatic Grading System for Supporting Self-Learning

Yuki Ito, Qiang Ma

机构 * Kyoto University(京都大学) Kyoto Institute of Technology(京都工艺纤维大学)

专题命中 其他LLM :LLM(abstract);分类 cs.CL、cs.AI

Comments 33 pages, 25 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.04185 2025-02-25 cs.LG cs.CL 62%

Residual Stream Analysis with Multi-Layer SAEs

Tim Lawson, Lucy Farnik, Conor Houghton, Laurence Aitchison

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.LG

Comments ICLR 2025 Camera Ready. 45 pages, 41 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.02306 2025-02-25 cs.LG cs.AI 62%

On Targeted Manipulation and Deception when Optimizing LLMs for User Feedback

Marcus Williams, Micah Carroll, Adhyyan Narang, Constantin Weisser, Brendan Murphy, Anca Dragan

专题命中 其他LLM :LLM(abstract);分类 cs.AI、cs.LG

Comments Accepted to ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.14048 2025-02-21 cs.CL cs.AI cs.HC 62%

Semantic Decomposition and Selective Context Filtering -- Text Processing Techniques for Context-Aware NLP-Based Systems

Karl John Villardar

机构 * Cebu Institute of Technology(宿务理工学院)

专题命中 其他LLM :LLM(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.12959 2025-02-19 cs.CL cs.AI 62%

AlignFreeze: Navigating the Impact of Realignment on the Layers of Multilingual Models Across Diverse Languages

Steve Bakos, Félix Gaschi, David Guzmán, Riddhi More, Kelly Chutong Li, En-Shiun Annie Lee

机构 * Ontario Tech University(安大略理工大学) University of Toronto(多伦多大学) SAS Posos(SAS Posos公司)

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.AI

Comments 24 pages, 2 figures, to be published in Proceedings of NAACL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.10966 2025-02-18 cs.CL cs.AI 62%

Neural Networks Remember More: The Power of Parameter Isolation and Combination

Biqing Zeng, Zehan Li, Aladdin Ayesh

机构 * Aberdeen Institute of Data Science and Artificial Intelligence(阿伯丁数据科学与人工智能学院) South China Normal University(华南师范大学) School of Artificial Intelligence(人工智能学院)

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.14630 2025-02-14 cs.LG cs.AI 62%

On the Regularization of Learnable Embeddings for Time Series Forecasting

Luca Butera, Giovanni De Felice, Andrea Cini, Cesare Alippi

机构 * Università della Svizzera Italiana(瑞士意大利语区大学) IDSIA(IDSIA研究院) University of Liverpool(利物浦大学) Politecnico di Milano(米兰理工大学)

专题命中 其他LLM :foundation model(abstract);分类 cs.AI、cs.LG

Comments Accepted at TMLR

Journal ref L. Butera, G. D. Felice, A. Cini, and C. Alippi. On the regularization of learnable embeddings for time series forecasting. Transactions on Machine Learning Research, 2025. ISSN 2835-8856. URL https://openreview.net/forum?id=F5ALCh3GWG

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.04263 2025-02-07 cs.CV cs.AI cs.LG 62%

Cross the Gap: Exposing the Intra-modal Misalignment in CLIP via Modality Inversion

Marco Mistretta, Alberto Baldrati, Lorenzo Agnolucci, Marco Bertini, Andrew D. Bagdanov

机构 * University of Florence(佛罗伦萨大学) University of Pisa(比萨大学)

专题命中 其他LLM :language model(abstract);分类 cs.AI、cs.LG

Comments Accepted for publication at ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.13554 2025-02-06 cs.CV cs.AI cs.LG 62%

One-Prompt-One-Story: Free-Lunch Consistent Text-to-Image Generation Using a Single Prompt

Tao Liu, Kai Wang, Senmao Li, Joost van de Weijer, Fahad Shahbaz Khan, Shiqi Yang, Yaxing Wang, Jian Yang, Ming-Ming Cheng

机构 * Nankai University(南开大学) Computer Vision Center(计算机视觉中心) Universitat Autònoma de Barcelona(巴塞罗那自治大学) Mohamed bin Zayed University of AI(穆罕默德·本·扎耶德人工智能大学) Linkoping University(林雪平大学) SB Intuitions(SB直觉公司)

专题命中 其他LLM :language model(abstract);分类 cs.AI、cs.LG

Comments 28 pages, 22 figures, ICLR2025 conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.01432 2025-02-04 cs.CL cs.LG 62%

Emergent Stack Representations in Modeling Counter Languages Using Transformers

Utkarsh Tiwari, Aviral Gupta, Michael Hahn

机构 * Birla Institute of Technology and Science, Pilani(比拉理工学院皮拉尼分校) Saarland University(萨尔大学)

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.20294 2025-01-31 cs.LG cs.AI physics.chem-ph 62%

A Bayesian Flow Network Framework for Chemistry Tasks

Nianze Tao, Minori Abe

机构 * Hiroshima University(广岛大学)

专题命中 其他LLM :language model(abstract);分类 cs.AI、cs.LG

Comments 7 figures, 12 tables, 27 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.03865 2025-01-30 cs.MA cs.AI cs.GT cs.LG cs.SI 62%

AdaSociety: An Adaptive Environment with Social Structures for Multi-Agent Decision-Making

Yizhe Huang, Xingbo Wang, Hao Liu, Fanqi Kong, Aoyang Qin, Min Tang, Song-Chun Zhu, Mingjie Bi, Siyuan Qi, Xue Feng

机构 * State Key Laboratory of General Artificial Intelligence, BIGAI(通用人工智能国家重点实验室,北京通用人工智能研究院) Peking University(北京大学) New York University(纽约大学) Tsinghua University(清华大学) University of Science and Technology of China(中国科学技术大学)

专题命中 其他LLM :LLM(abstract);分类 cs.AI、cs.LG

Comments Accepted at NeurIPS D&B 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.15304 2025-01-28 cs.SD cs.AI cs.HC cs.LG eess.AS 62%

Music Generation using Human-In-The-Loop Reinforcement Learning

Aju Ani Justus

专题命中 其他LLM :language model(abstract);分类 cs.AI、cs.LG

Comments This is a preprint of a paper presented at the 2023 IEEE International Conference on Big Data (BigData). It has been made public for the benefit of the community and should be considered a preprint rather than a formally reviewed paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.13533 2025-01-24 cs.AI cs.LG 62%

Towards a Theory of AI Personhood

Francis Rhys Ward

专题命中 其他LLM :language model(abstract);分类 cs.AI、cs.LG

Comments AAAI-25 AI Alignment Track

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.05663 2025-01-13 quant-ph cs.AI cs.ET cs.LG cs.NE 62%

Learning to Measure Quantum Neural Networks

Samuel Yen-Chi Chen, Huan-Hsin Tseng, Hsin-Yi Lin, Shinjae Yoo

机构 * Wells Fargo(富国银行) Brookhaven National Laboratory(布鲁克海文国家实验室) Seton Hall University(薛顿贺尔大学)

专题命中 其他LLM :prompting(abstract);分类 cs.AI、cs.LG

Comments Accepted by ICASSP 2025 Workshop: Quantum Machine Learning in Signal Processing and Artificial Intelligence

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.00334 2025-01-03 cs.CL cs.AI 62%

Loss-Aware Curriculum Learning for Chinese Grammatical Error Correction

Ding Zhang, Yangning Li, Lichen Bai, Hao Zhang, Yinghui Li, Haiye Lin, Hai-Tao Zheng, Xin Su, Zifei Shan

机构 * Shenzhen International Graduate School Tsinghua University(清华大学深圳国际研究生院) Peng Cheng Laboratory(鹏城实验室) WeChat Tencent(腾讯微信)

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.AI

Comments ICASSP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.00136 2025-01-03 cs.CV cs.AI cs.LG 62%

Detection-Fusion for Knowledge Graph Extraction from Videos

Taniya Das, Louis Mahon, Thomas Lukasiewicz

专题命中 其他LLM :language model(abstract);分类 cs.AI、cs.LG

Comments 12 pages, To be submitted to a conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.19814 2024-12-31 q-bio.NC cs.AI cs.LG 62%

Predicting Human Brain States with Transformer

Yifei Sun, Mariano Cabezas, Jiah Lee, Chenyu Wang, Wei Zhang, Fernando Calamante, Jinglei Lv

专题命中 其他LLM :language model(abstract);分类 cs.AI、cs.LG

Comments 11 pages, 4 figures, MICCAI MMMI workshop in press

详情

展开后加载摘要…

URL PDF HTML 收藏