arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12266 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12266 篇

2510.12252 2025-10-20 cs.CR cs.AI 70%

PromptLocate: Localizing Prompt Injection Attacks

Yuqi Jia, Yupei Liu, Zedian Shao, Jinyuan Jia, Neil Gong

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

Comments To appear in IEEE Symposium on Security and Privacy, 2026. For slides, see https://people.duke.edu/~zg70/code/PromptInjection.pdf

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13857 2025-10-17 cs.SE cs.AI 70%

From Craft to Constitution: A Governance-First Paradigm for Principled Agent Engineering

Qiang Xu, Xiangyu Wen, Changran Xu, Zeju Li, Jianyuan Zhong

机构 * CURE Lab., Dept. of CSE, The Chinese University of Hong Kong, Hong Kong S.A.R.(CUHK计算机科学与工程系CURE实验室)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12364 2025-10-16 cs.SE cs.AI cs.HC 70%

(R)evolution of Programming: Vibe Coding as a Post-Coding Paradigm

Kevin Krings, Nino S. Bohn, Thomas Ludwig

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

Comments Workshop Contribution at the sixth decennial Aarhus conference in "The End of Programming (as we know it) - Envisioning Radical Re-Conceptualizations of Co-Coding with AI"

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.23144 2025-10-16 cs.AI cond-mat.stat-mech cs.MA nlin.AO physics.soc-ph 70%

Coordination Requires Simplification: Thermodynamic Bounds on Multi-Objective Compromise in Natural and Artificial Intelligence

Atma Anand

机构 * Department of Physics and Astronomy, University of Rochester(物理与天文学系,罗切斯特大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

Comments 15 pages, 1 figure, 9 pages supplementary material, submitted to Journal of Physics: Complexity

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.08891 2025-10-16 cs.IR cs.AI 70%

Reliable Decision Making via Calibration Oriented Retrieval Augmented Generation

Chaeyun Jang, Deukhwan Cho, Seanie Lee, Hyungi Lee, Juho Lee

机构 * KAIST(韩国科学技术院) Kookmin University(韩国釜山大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

Comments Accepted by NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11823 2025-10-15 cs.CR cs.AI 70%

BlackIce: A Containerized Red Teaming Toolkit for AI Security Testing

Caelin Kaplan, Alexander Warnecke, Neil Archibald

机构 * Databricks

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.22954 2025-10-15 cs.LG 70%

Retrieval-Augmented Generation with Estimation of Source Reliability

Jeongyeon Hwang, Junyoung Park, Hyejin Park, Dongwoo Kim, Sangdon Park, Jungseul Ok

机构 * Pohang University of Science and Technology (POSTECH)(釜山科学技术大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.08075 2025-10-14 cs.AI 70%

Multi-Condition Conformal Selection

Qingyang Hao, Wenbo Liao, Bingyi Jing, Hongxin Wei

机构 * Southern University of Science and Technology(南方科技大学) The Chinese University of Hong Kong(香港中文大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.05444 2025-10-14 cs.CL 70%

PhoniTale: Phonologically Grounded Mnemonic Generation for Typologically Distant Language Pairs

Sana Kang, Myeongseok Gwon, Su Young Kwon, Jaewook Lee, Andrew Lan, Bhiksha Raj, Rita Singh

机构 * KAIST(韩国科学技术院) University of Massachusetts Amherst(马萨诸塞大学阿默斯特分校) Carnegie Mellon University(卡内基梅隆大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

Comments Accepted to EMNLP 2025 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.08083 2025-10-13 cs.HC cs.AI q-bio.NC 70%

Confidence-weighted integration of human and machine judgments for superior decision-making

Felipe Yáñez, Xiaoliang Luo, Omar Valerio Minero, Bradley C. Love

机构 * Max Planck Institute for Neurobiology of Behavior – caesar, Bonn, Germany(马克斯·普朗克行为神经生物学研究所——caesar,波恩,德国) Department of Experimental Psychology, University College London, London, United Kingdom(实验心理学系,伦敦大学学院,伦敦,英国)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.05869 2025-10-08 cs.CL 70%

The fragility of "cultural tendencies" in LLMs

Kun Sun, Rong Wang

机构 * Department of Linguistics, Tongji University(同济大学语言学系) Department of Computational Linguistics, Tübingen University(图宾根大学计算语言学系) The Institute of Natural Language Processing, Stuttgart University(斯图加特大学自然语言处理研究所)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.05596 2025-10-08 cs.AI 70%

From Agentification to Self-Evolving Agentic AI for Wireless Networks: Concepts, Approaches, and Future Research Directions

Changyuan Zhao, Ruichen Zhang, Jiacheng Wang, Dusit Niyato, Geng Sun, Xianbin Wang, Shiwen Mao, Abbas Jamalipour

机构 * College of Computing and Data Science, Nanyang Technological University(计算与数据科学学院,南洋理工大学) College of Computer Science and Technology, Jilin University(计算机科学与技术学院,吉林大学) Department of Electrical and Computer Engineering, Western University(电气与计算机工程系,西方大学) Department of Electrical and Computer Engineering, Auburn University(电气与计算机工程系,阿伯茨罕大学) School of Electrical and Computer Engineering, University of Sydney(电气与计算机工程学院,悉尼大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

Comments 7 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.05441 2025-10-08 cs.SE cs.AI 70%

UnitTenX: Generating Tests for Legacy Packages with AI Agents Powered by Formal Verification

Yiannis Charalambous, Claudionor N. Coelho, Luis Lamb, Lucas C. Cordeiro

机构 * The University of Manchester, UK(曼彻斯特大学,英国) ECE Department, Santa Clara University, US(圣克拉拉大学电子与计算机工程系,美国)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.03886 2025-10-07 cs.AI 70%

Rare Text Semantics Were Always There in Your Diffusion Transformer

Seil Kang, Woojung Han, Dayun Ju, Seong Jae Hwang

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

Comments Accepted to NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.03678 2025-10-07 cs.LG stat.ML 70%

Towards Sampling Data Structures for Tensor Products in Turnstile Streams

Zhao Song, Shenghao Xie, Samson Zhou

机构 * University of California, Berkeley(加州大学伯克利分校) Texas A&M University(德克萨斯农工大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.08905 2025-10-07 cs.CL 70%

Forecasting Conversation Derailments Through Generation

Yunfan Zhang, Kathleen McKeown, Smaranda Muresan

机构 * Columbia University(哥伦比亚大学) Barnard College(巴纳德学院)

专题命中 其他LLM :LLM(abstract);language model(abstract);分类 cs.CL

Comments ACL INLG 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.03331 2025-10-07 cs.CY cs.AI 70%

Intelligent Healthcare Ecosystems: Optimizing the Iron Triangle of Healthcare (Access, Cost, Quality)

Vivek Acharya

机构 * Boston University(波士顿大学) MIT(麻省理工学院) Stanford(斯坦福大学) University of Texas at Austin(德克萨斯大学奥斯汀分校)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

Comments 8 pages, 4 figures, formatted per MDPI guidelines, APA-style numbered references

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.11840 2025-10-06 cs.LG math.OC 70%

On the $O(\frac{\sqrt{d}}{K^{1/4}})$ Convergence Rate of AdamW Measured by $\ell_1$ Norm

Huan Li, Yiming Dong, Zhouchen Lin

机构 * Institute of Robotics and Automatic Information Systems, College of Artificial Intelligence, Nankai University, Tianjin, China(机器人与自动信息系统研究所,人工智能学院,南开大学,天津,中国) National Key Lab of General AI, School of Intelligence Science and Technology, Peking University, Beijing, China(通用人工智能国家重点实验室,智能科学与技术学院,北京大学,北京,中国)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.LG

Comments V2: NeurIPS Camera-Ready. V3: expand upon the conference version by incorporating the analysis of NAdamW

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.03923 2025-10-06 cs.CL 70%

Did Translation Models Get More Robust Without Anyone Even Noticing?

Ben Peters, André F. T. Martins

机构 * Instituto de Telecomunicações(电信研究所) Instituto Superior Técnico(技术高等学院) Universidade de Lisboa(里斯本大学) ELLIS Unit Lisbon (LUMLIS)(里斯本ELLIS单位(LUMLIS)) Unbabel

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

Comments ACL 2025 (Main) camera ready

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01451 2025-10-03 q-fin.GN cs.LG 70%

Financial Stability Implications of Generative AI: Taming the Animal Spirits

Anne Lundgaard Hansen, Seung Jung Lee

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01185 2025-10-02 cs.LG 70%

Dirichlet-Prior Shaping: Guiding Expert Specialization in Upcycled MoEs

Leyla Mirvakhabova, Babak Ehteshami Bejnordi, Gaurav Kumar, Hanxue Liang, Wanru Zhao, Paul Whatmough

机构 * Qualcomm AI Research(高通人工智能研究)

专题命中 其他LLM :LLM(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01077 2025-10-02 cs.SE cs.AI 70%

CodeGenLink: A Tool to Find the Likely Origin and License of Automatically Generated Code

Daniele Bifolco, Guido Annicchiarico, Pierluigi Barbiero, Massimiliano Di Penta, Fiorella Zampetti

机构 * University of Sannio, Italy(萨尼亚大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

Comments Proceedings of the 40th IEEE/ACM International Conference on Automated Software Engineering (ASE 2025), November 16-20 2025, Seoul, South Korea

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.00552 2025-10-02 cs.AI cs.HC 70%

Data Quality Challenges in Retrieval-Augmented Generation

Leopold Müller, Joshua Holstein, Sarah Bause, Gerhard Satzger, Niklas Kühl

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

Comments Preprint version. Accepted for presentation at the International Conference on Information Systems (ICIS 2025). Please cite the published version when available

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.00283 2025-10-02 physics.optics cs.AI 70%

Data driven approaches in nanophotonics: A review of AI-enabled metadevices

Huanshu Zhang, Lei Kang, Sawyer D. Campbell, Jacob T. Young, Douglas H. Werner

机构 * Department of Electrical Engineering(电气工程系) The Pennsylvania State University(宾夕法尼亚州立大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.00046 2025-10-02 cs.CV cs.AI 70%

Reinforcement Learning-Based Prompt Template Stealing for Text-to-Image Models

Xiaotian Zou

机构 * Xiaotian Zou

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

Comments 10 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24216 2025-09-30 cs.CL cs.CY 70%

MoVa: Towards Generalizable Classification of Human Morals and Values

Ziyu Chen, Junfei Sun, Chenxi Li, Tuan Dung Nguyen, Jing Yao, Xiaoyuan Yi, Xing Xie, Chenhao Tan, Lexing Xie

机构 * The Australian National University(澳大利亚国立大学) University of Chicago(芝加哥大学) University of Pennsylvania(宾夕法尼亚大学) Microsoft Research Asia(微软亚洲研究院)

专题命中 其他LLM :LLM(abstract);prompting(abstract);分类 cs.CL

Comments 9 pages, 10 figures and tables, EMNLP 2025 main conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23512 2025-09-30 cs.LG math.OC 70%

Differentially Private Clipped-SGD: High-Probability Convergence with Arbitrary Clipping Level

Saleh Vatan Khah, Savelii Chezhegov, Shahrokh Farahmand, Samuel Horváth, Eduard Gorbunov

机构 * IUST(伊朗伊斯兰科技与应用大学) MBZUAI(马尔代夫人工智能研究所)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.LG

Comments 60 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.20817 2025-09-30 math.OC cs.LG 70%

Convergence of Clipped-SGD for Convex $(L_0,L_1)$-Smooth Optimization with Heavy-Tailed Noise

Savelii Chezhegov, Aleksandr Beznosikov, Samuel Horváth, Eduard Gorbunov

机构 * MCAS(麦吉尔-西奈大学(McGill-Sinai Academy)) MBZUAI(穆罕默德·本·拉希德智能研究院(MBZUAI))

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.LG

Comments 33 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.00689 2025-09-30 cs.CR cs.AI 70%

Ocassionally Secure: A Comparative Analysis of Code Generation Assistants

Ran Elgedawy, Porter Dosch, John Sadik, Senjuti Dutta, Anuj Gautam, Konstantinos Georgiou, Farzin Gholamrezae, Fujiao Ji, Kyungchan Lim, Qian Liu, Scott Ruoti

机构 * OpenAI Google(谷歌) DeepSeek

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

Comments 12 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.16838 2025-09-30 cs.CL 70%

If We May De-Presuppose: Robustly Verifying Claims through Presupposition-Free Question Decomposition

Shubhashis Roy Dipta, Francis Ferraro

机构 * Department of Computer Science and Electrical Engineering University of Maryland Baltimore County(计算机科学与电气工程系大学马里兰大学巴尔的摩县)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

Comments Published in *SEM 2025

详情

展开后加载摘要…

URL PDF HTML 收藏