arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12193 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12193 篇

2406.03614 2025-09-30 cs.LG cs.CL q-fin.RM 79%

Advancing Anomaly Detection: Non-Semantic Financial Data Encoding with LLMs

Alexander Bakumenko, Kateřina Hlaváčková-Schindler, Claudia Plant, Nina C. Hubig

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

Journal ref IEEE Access 13 (2025) 146757-146771

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14723 2025-09-16 cs.CL cs.AI 79%

Transplant Then Regenerate: A New Paradigm for Text Data Augmentation

Guangzhan Wang, Hongyu Zhang, Beijun Shen, Xiaodong Gu

机构 * Shanghai Jiao Tong University(上海交通大学) Chongqing University(重庆大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments Accepted by EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.09714 2025-09-15 cs.CL cs.AI 79%

How Small Transformation Expose the Weakness of Semantic Similarity Measures

Serge Lionel Nikiema, Albérick Euraste Djire, Abdoul Aziz Bonkoungou, Micheline Bénédicte Moumoula, Jordan Samhi, Abdoul Kader Kabore, Jacques Klein, Tegawendé F. Bissyande

机构 * University of Luxembourg(卢森堡大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.19860 2025-09-12 cs.CL cs.AI 79%

MIND: Towards Immersive Psychological Healing with Multi-agent Inner Dialogue

Yujia Chen, Changsong Li, Yiming Wang, Tianjie Ju, Qingqing Xiao, Nan Zhang, Zifan Kong, Peng Wang, Binyu Yan

机构 * Sichuan University(四川大学) Shanghai Jiao Tong University(上海交通大学) Mental Health Center, West China Hospital, Sichuan University(四川大学西昌医院心理健康中心) WestChina School of Nursing, Sichuan University(四川大学西昌护理学院) National University of Singapore(新加坡国立大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments Accepted by EMNLP 2025 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.04500 2025-09-08 cs.CL cs.AI 79%

Context Engineering for Trustworthiness: Rescorla Wagner Steering Under Mixed and Inappropriate Contexts

Rushi Wang, Jiateng Liu, Cheng Qian, Yifan Shen, Yanzhou Pan, Zhaozhuo Xu, Ahmed Abbasi, Heng Ji, Denghui Zhang

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments 36 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.19697 2025-08-28 cs.CR cs.AI cs.CL 79%

Safety Alignment Should Be Made More Than Just A Few Attention Heads

Chao Huang, Zefeng Zhang, Juewei Yue, Quangang Li, Chuang Zhang, Tingwen Liu

机构 * Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所) School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络安全学院)

专题命中 其他LLM :large language model(abstract);language model(abstract);prompting(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.09545 2025-08-26 cs.CL cs.AI cs.CY 79%

Does GPT-4 surpass human performance in linguistic pragmatics?

Ljubisa Bojic, Predrag Kovacevic, Milan Cabarkapa

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments 19 pages, 1 figure, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13948 2025-08-20 cs.HC cs.AI cs.CL cs.PL 79%

Prompt Orchestration Markup Language

Yuge Zhang, Nan Chen, Jiahang Xu, Yuqing Yang

机构 * Microsoft Research(微软研究院)

专题命中 其他LLM :large language model(abstract);language model(abstract);prompting(abstract);分类 cs.CL、cs.AI

Comments All findings in this paper are derived from a POML snapshot as of February 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.04820 2025-08-08 cs.SE cs.AI cs.LG 79%

Automated File-Level Logging Generation for Machine Learning Applications using LLMs: A Case Study using GPT-4o Mini

Mayra Sofia Ruiz Rodriguez, SayedHassan Khatoonabadi, Emad Shihab

机构 * Concordia University(康科德大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22160 2025-07-31 cs.CR cs.AI cs.CL 79%

Strategic Deflection: Defending LLMs from Logit Manipulation

Yassine Rachidy, Jihad Rbaiti, Youssef Hmamouche, Faissal Sehbaoui, Amal El Fallah Seghrouchni

机构 * International Artificial Intelligence Center of Morocco(摩洛哥国际人工智能中心) Mohammed VI Polytechnic University(摩洛哥穆莱·伊斯梅尔理工学院) AgriEdge Sorbonne University, LIP6 - UMR 7606 CNRS(索邦大学,LIP6 - UMR 7606 CNRS)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments 20 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.20788 2025-07-18 cs.CL cs.LG 79%

SCULPT: Systematic Tuning of Long Prompts

Shanu Kumar, Akhila Yesantarao Venkata, Shubhanshu Khandelwal, Bishal Santra, Parag Agrawal, Manish Gupta

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

Comments Accepted at ACL Main 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.10077 2025-07-16 cs.CL cs.AI cs.IR cs.IT math.IT 79%

A quantum semantic framework for natural language processing

Christopher J. Agostino, Quan Le Thien, Molly Apsel, Denizhan Pak, Elina Lesyk, Ashabari Majumdar

机构 * Department of Physics, Indiana University, Bloomington, Indiana 47405, USA(印第安纳大学物理系) Quantum Science and Engineering Center (QSEC), Indiana University, Bloomington, Indiana 47405, USA(印第安纳大学量子科学与工程中心) Cognitive Science Program, Indiana University, Bloomington, Indiana 47405, USA(印第安纳大学认知科学计划) Department of Psychological and Brain Sciences, Indiana University, Bloomington, Indiana 47405, USA(印第安纳大学心理与脑科学系) Luddy School of Informatics, Computing, and Engineering, Indiana University, Bloomington, Indiana 47405, USA(印第安纳大学信息科学、计算与工程学院) Department of Physics, University of Notre Dame, Notre Dame, IN 46556, USA(诺丁汉大学物理系)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments 12 pages, 2 figures, accepted submission to Quantum AI and NLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.03940 2025-07-15 cs.CL cs.AI 79%

Not all tokens are created equal: Perplexity Attention Weighted Networks for AI generated text detection

Pablo Miralles-González, Javier Huertas-Tato, Alejandro Martín, David Camacho

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.09639 2025-07-15 cs.MA cs.AI cs.CL cs.CY cs.HC 79%

Can A Society of Generative Agents Simulate Human Behavior and Inform Public Health Policy? A Case Study on Vaccine Hesitancy

Abe Bohan Hou, Hongru Du, Yichen Wang, Jingyu Zhang, Zixiao Wang, Paul Pu Liang, Daniel Khashabi, Lauren Gardner, Tianxing He

机构 * Department of Computer Science, Johns Hopkins University(约翰霍普金斯大学计算机科学系) Institute of Interdisciplinary Information Sciences, Tsinghua University(清华大学交叉信息研究院) Shanghai Qi Zhi Institute(上海启智研究院) Civil & Systems Engineering, Johns Hopkins University(约翰霍普金斯大学土木与系统工程系) Department of Computer Science, University of Chicago(芝加哥大学计算机科学系) Department of Epidemiology, Harvard University(哈佛大学流行病学系) MIT Media Lab and Department of EECS, Massachusetts Institute of Technology(麻省理工学院媒体实验室和电子工程与计算机科学系)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments Accepted to COLM 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.24030 2025-07-11 cs.LG cs.AI cs.CV 79%

From Images to Signals: Are Large Vision Models Useful for Time Series Analysis?

Ziming Zhao, ChengAo Shen, Hanghang Tong, Dongjin Song, Zhigang Deng, Qingsong Wen, Jingchao Ni

机构 * University of Houston(休斯顿大学) University of Illinois at Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) University of Connecticut(康涅狄格大学) Squirrel Ai Learning(squirrel Ai 学习)

专题命中 其他LLM :large language model(abstract);language model(abstract);foundation model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.00079 2025-07-02 cs.AI cs.LG 79%

VoyagerVision: Investigating the Role of Multi-modal Information for Open-ended Learning Systems

Ethan Smyth, Alessandro Suglia

机构 * School of Mathematical and Computer Sciences(数学与计算机科学学院) Heriot-Watt University(赫瑞斯韦大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

Comments website: https://esmyth-dev.github.io/VoyagerVision.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.23930 2025-07-01 cs.CL cs.AI 79%

Leveraging the Potential of Prompt Engineering for Hate Speech Detection in Low-Resource Languages

Ruhina Tabasshum Prome, Tarikul Islam Tamiti, Anomadarshi Barua

机构 * Bangladesh Institute of Governance and Management (BIGM)(孟加拉国治理与管理研究所) Department of Cyber Security Engineering, George Mason University(网络安全工程系,乔治·梅森大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);prompting(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.20426 2025-07-01 cs.CL cs.AI cs.CY cs.MA 79%

Among Them: A game-based framework for assessing persuasion capabilities of LLMs

Mateusz Idziejczak, Vasyl Korzavatykh, Mateusz Stawicki, Andrii Chmutov, Marcin Korcz, Iwo Błądek, Dariusz Brzezinski

机构 * Institute of Computing Science, Poznan University of Technology, Poland(计算机科学学院,波兹南技术大学,波兰)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.17367 2025-06-24 cs.CL cs.AI cs.MA 79%

Cash or Comfort? How LLMs Value Your Inconvenience

Mateusz Cedro, Timour Ichmoukhamedov, Sofie Goethals, Yifan He, James Hinns, David Martens

机构 * University of Antwerp(安特卫普大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments 12 pages, 4 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.03491 2025-06-19 cs.CL cs.AI 79%

Can LLMs Ask Good Questions?

Yueheng Zhang, Xiaoyuan Liu, Yiyou Sun, Atheer Alharbi, Hend Alzahrani, Tianneng Shi, Basel Alomair, Dawn Song

机构 * University of California Berkeley(加州大学伯克利分校) KACST(王国立科学与技术委员会) University of Washington(华盛顿大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.16526 2025-06-18 cs.SD cs.AI cs.CL eess.AS 79%

Text2midi: Generating Symbolic Music from Captions

Keshav Bhandari, Abhinaba Roy, Kyra Wang, Geeta Puri, Simon Colton, Dorien Herremans

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments 9 pages, 3 figures, Accepted at the 39th AAAI Conference on Artificial Intelligence (AAAI 2025)

Journal ref Proceedings of the 39th AAAI Conference on Artificial Intelligence (AAAI 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.02285 2025-06-11 cs.LG cs.AI 79%

Why Gradients Rapidly Increase Near the End of Training

Aaron Defazio

机构 * FAIR at Meta(Meta 的 FAIR)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.10997 2025-06-10 cs.CL cs.AI cs.CV 79%

RONA: Pragmatically Diverse Image Captioning with Coherence Relations

Aashish Anantha Ramakrishnan, Aadarsh Anantha Ramakrishnan, Dongwon Lee

机构 * The Pennsylvania State University(宾夕法尼亚州立大学) National Institute of Technology, Tiruchirappalli(特里奇里帕利理工学院)

专题命中 其他LLM :large language model(abstract);language model(abstract);prompting(abstract);分类 cs.CL、cs.AI

Comments Accepted in the NAACL Fourth Workshop on Intelligent and Interactive Writing Assistants (In2Writing), Albuquerque, New Mexico, May 2025, https://in2writing.glitch.me

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.13544 2025-06-10 cs.CL cs.AI 79%

From Sub-Ability Diagnosis to Human-Aligned Generation: Bridging the Gap for Text Length Control via MARKERGEN

Peiwen Yuan, Chuyi Tan, Shaoxiong Feng, Yiwei Li, Xinglin Wang, Yueqi Zhang, Jiayi Shi, Boyuan Pan, Yao Hu, Kan Li

机构 * School of Computer Science and Technology, Beijing Institute of Technology(北京理工大学计算机科学与技术学院) Xiaohongshu Inc(小红书公司)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.14676 2025-06-10 cs.CL cs.AI 79%

SudoLM: Learning Access Control of Parametric Knowledge with Authorization Alignment

Qin Liu, Fei Wang, Chaowei Xiao, Muhao Chen

机构 * UC Davis(加州大学戴维斯分校) USC(南加州大学) UW-Madison(威斯康星大学麦迪逊分校)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments ACL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.05390 2025-06-09 cs.CL cs.LG 79%

Understanding Gender Bias in AI-Generated Product Descriptions

Markelle Kelly, Mohammad Tahaei, Padhraic Smyth, Lauren Wilcox

机构 * University of California, Irvine(加州大学尔湾分校) eBay(eBay公司) Georgia Institute of Technology(佐治亚理工学院)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

Comments Accepted to FAccT 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.08979 2025-06-09 cs.CL cs.AI cs.MA cs.SE 79%

Multi-Agent Collaboration via Cross-Team Orchestration

Zhuoyun Du, Chen Qian, Wei Liu, Zihao Xie, YiFei Wang, Rennai Qiu, Yufan Dang, Weize Chen, Cheng Yang, Ye Tian, Xuantang Xiong, Lei Han

机构 * State Key Lab of CAD & CG(CAD与CG国家重点实验室) Zhejiang Polytechnic Institute(浙江工业大学) Polytechnic Institute, Zhejiang University(浙江大学 polytechnic 院) Shanghai Jiao Tong University(上海交通大学) King’s College London(伦敦国王学院) Tsinghua University(清华大学) Beijing University of Posts and Telecommunications(北京邮电大学) Tencent Robotics X(腾讯机器人科技)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments Accepted to Findings of ACL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.02389 2025-06-04 cs.LG cs.AI 79%

Univariate to Multivariate: LLMs as Zero-Shot Predictors for Time-Series Forecasting

Chamara Madarasingha, Nasrin Sohrabi, Zahir Tari

机构 * School of Computing Technologies RMIT University, Australia(计算机技术学院皇家墨尔本理工大学,澳大利亚) School of Information Technologies Deakin University, Australia(信息科技学院德肯大学,澳大利亚)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.00069 2025-06-03 cs.CL cs.AI 79%

Evaluating the Sensitivity of LLMs to Prior Context

Robert Hankache, Kingsley Nketia Acheampong, Liang Song, Marek Brynda, Raad Khraishi, Greig A. Cowan

机构 * NatWest AI Research(NatWest人工智能研究院) University College London(伦敦大学学院)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.00019 2025-06-03 cs.CL cs.AI 79%

Amadeus-Verbo Technical Report: The powerful Qwen2.5 family models trained in Portuguese

William Alberto Cruz-Castañeda, Marcellus Amadeus

机构 * Amadeus AI

专题命中 其他LLM :large language model(abstract);language model(abstract);foundation model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏