arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12266 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12266 篇

2406.14737 2025-05-29 cs.CL 70%

Dissecting the Ullman Variations with a SCALPEL: Why do LLMs fail at Trivial Alterations to the False Belief Task?

Zhiqiang Pi, Annapurna Vadaparty, Benjamin K. Bergen, Cameron R. Jones

机构 * Department of Psychology, 1202 W. Johnson Street Madison, WI 53706 USA(心理学系) Department of Educational Psychology, 1025 W. Johnson Street Madison, WI 53706 USA(教育心理学系)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.21458 2025-05-28 cs.CL 70%

Do LLMs Need to Think in One Language? Correlation between Latent Language and Task Performance

Shintaro Ozaki, Tatsuya Hiraoka, Hiroto Otake, Hiroki Ouchi, Masaru Isonuma, Benjamin Heinzerling, Kentaro Inui, Taro Watanabe, Yusuke Miyao, Yohei Oseki, Yu Takagi

机构 * NAIST(日本国立信息与科技研究所) NII LLMC(日本信息处理学会大语言模型中心) MBZUAI(微软亚洲人工智能研究院) RIKEN(日本研究机构) Tohoku University(东北大学) The University of Tokyo(东京大学) Nagoya Institute of Technology(名古屋技术大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.07923 2025-05-28 math.OC cs.LG 70%

Sign Operator for Coping with Heavy-Tailed Noise in Non-Convex Optimization: High Probability Bounds Under $(L_0, L_1)$-Smoothness

Nikita Kornilov, Philip Zmushko, Andrei Semenov, Mark Ikonnikov, Alexander Gasnikov, Alexander Beznosikov

机构 * MIPT(莫斯科物理技术学院) Skoltech(斯克里普奇学院) Yandex EPFL(瑞士联邦理工学院) Innopolis University(因诺波利斯大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.00984 2025-05-28 cs.CL 70%

Predicting drug-gene relations via analogy tasks with word embeddings

Hiroaki Yamagiwa, Ryoma Hashimoto, Kiwamu Arakane, Ken Murakami, Shou Soeda, Momose Oyama, Yihua Zhu, Mariko Okada, Hidetoshi Shimodaira

机构 * Kyoto University(京都大学) Recruit Co., Ltd.(Recruit公司) Institute for Protein Research(蛋白质研究所) Osaka University(大阪大学) Research Institute of Molecular Pathology(分子病理研究所) RIKEN(理化学研究所)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

Journal ref Sci Rep 15, 17240 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.18864 2025-05-27 cs.CL 70%

Audio Jailbreak Attacks: Exposing Vulnerabilities in SpeechGPT in a White-Box Framework

Binhao Ma, Hanqing Guo, Zhengping Jay Luo, Rui Duan

机构 * Department of Computer Science University of Missouri-Kansas City(计算机科学系 密苏里大学-康科特分校) Department of Computer Science and Physics Rider University(计算机科学与物理系 Rider大学) Department of Electrical and Computer Engineering University of Hawai’i at Mānoa(电气与计算机工程系 夏威夷大学马诺亚分校)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.18657 2025-05-27 cs.AI 70%

MLLMs are Deeply Affected by Modality Bias

Xu Zheng, Chenfei Liao, Yuqian Fu, Kaiyu Lei, Yuanhuiyi Lyu, Lutao Jiang, Bin Ren, Jialei Chen, Jiawen Wang, Chengxin Li, Linfeng Zhang, Danda Pani Paudel, Xuanjing Huang, Yu-Gang Jiang, Nicu Sebe, Dacheng Tao, Luc Van Gool, Xuming Hu

机构 * HKUST(GZ)(香港科技大学(广州)) CSE, HKUST(香港科技大学计算机科学与工程系) Xi’an Jiaotong University(西安交通大学) University of Pisa, IT(比萨大学) University of Trento, IT(特伦特大学) Nagoya University(名古屋大学) China University of Mining & Technology, Beijing(中国矿业大学(北京)) Tongji University(同济大学) SPIC Energy Science and Technology Research Institute(SPIC能源科学与技术研究院) Shanghai Jiao Tong University(上海交通大学) Fudan University(复旦大学) College of Computing & Data Science, Nanyang Technological University(南洋理工大学计算机与数据科学学院)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.06549 2025-05-27 cs.CY cs.AI 70%

Societal Impacts Research Requires Benchmarks for Creative Composition Tasks

Judy Hanwen Shen, Carlos Guestrin

专题命中 其他LLM :language model(abstract);foundation model(abstract);分类 cs.AI

Comments v1: ICLR 2025 Workshop on Bidirectional Human-AI Alignment (BiAlign) v2: ICML 2025 Position Paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.01263 2025-05-27 cs.CV cs.CL 70%

Generalizable Prompt Learning of CLIP: A Brief Overview

Fangming Cui, Yonggang Zhang, Xuan Wang, Xule Wang, Liang Xiao

机构 * Meituan(美团)

专题命中 其他LLM :language model(abstract);prompting(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.17861 2025-05-26 cs.AI cs.CY cs.IR 70%

Superplatforms Have to Attack AI Agents

Jianghao Lin, Jiachen Zhu, Zheli Zhou, Yunjia Xi, Weiwen Liu, Yong Yu, Weinan Zhang

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

Comments Position paper under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.17147 2025-05-26 cs.CR cs.AI 70%

MTSA: Multi-turn Safety Alignment for LLMs through Multi-round Red-teaming

Weiyang Guo, Jing Li, Wenya Wang, YU LI, Daojing He, Jun Yu, Min Zhang

机构 * Harbin Institute of Technology, Shenzhen, China(哈尔滨工业大学(深圳)) Nanyang Technological University, Singapore(南洋理工大学) Zhejiang University, Zhejiang, China(浙江大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

Comments 19 pages,6 figures,ACL2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.17045 2025-05-26 cs.CL 70%

Assessing GPT's Bias Towards Stigmatized Social Groups: An Intersectional Case Study on Nationality Prejudice and Psychophobia

Afifah Kashif, Heer Patel

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.12746 2025-05-26 cs.AI 70%

Correspondence of high-dimensional emotion structures elicited by video clips between humans and Multimodal LLMs

Haruka Asanuma, Naoko Koide-Majima, Ken Nakamura, Takato Horii, Shinji Nishimoto, Masafumi Oizumi

机构 * The University of Tokyo, Graduate School of Arts and Sciences(东京大学艺术与科学研究生院) Center for Information and Neural Networks (CiNet), National Institute of Information and Communications Technology(信息与神经网络中心(CiNet),信息与通信技术国家研究所) The University of Osaka, Graduate School of Frontier Biosciences(大阪大学前沿生命科学研究生院) The University of Tokyo, Faculty of Engineering(东京大学工学部) The University of Osaka, Graduate School of Engineering Science(大阪大学工学研究院) The University of Osaka, Graduate School of Medicine(大阪大学医学研究院)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

Comments 25 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.16284 2025-05-23 cs.LG 70%

Only Large Weights (And Not Skip Connections) Can Prevent the Perils of Rank Collapse

Josh Alman, Zhao Song

机构 * Columbia University(哥伦比亚大学) University of California, Berkeley(加州大学伯克利分校)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.07367 2025-05-23 cs.CL 70%

My Words Imply Your Opinion: Reader Agent-based Propagation Enhancement for Personalized Implicit Emotion Analysis

Jian Liao, Yu Feng, Yujin Zheng, Jun Zhao, Suge Wang, Jianxing Zheng

机构 * School of Computer and Information Technology, Shanxi University, China(山西大学计算机与信息学院) Institute of Automation, Chinese Academy of Science, China(中国科学院自动化研究所) Key Laboratory of Computational Intelligence and Chinese Information Processing of Ministry of Education, Shanxi University, China(教育部计算智能与中文信息处理重点实验室) Joint Laboratory of Tourism Big Data in Shanxi Province, China(山西省旅游大数据联合实验室)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

Journal ref Proceedings of the 63th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) (ACL2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.10347 2025-05-23 cs.CL 70%

A Unified Approach to Routing and Cascading for LLMs

Jasper Dekoninck, Maximilian Baader, Martin Vechev

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.20049 2025-05-22 cs.CL 70%

It's the same but not the same: Do LLMs distinguish Spanish varieties?

Marina Mayor-Rocher, Cristina Pozo, Nina Melero, Gonzalo Martínez, María Grandury, Pedro Reviriego

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

Comments in Spanish language

Journal ref SEPLN, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.14286 2025-05-21 cs.CL cs.SD eess.AS 70%

Universal Acoustic Adversarial Attacks for Flexible Control of Speech-LLMs

Rao Ma, Mengjie Qian, Vyas Raina, Mark Gales, Kate Knill

机构 * ALTA Institute, Department of Engineering, University of Cambridge(ALTA研究院,工程系,剑桥大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.00932 2025-05-21 cs.SE cs.AI 70%

LLMs: A Game-Changer for Software Engineers?

Md Asraful Haque

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

Comments 20 pages, 7 figures, 3 tables

Journal ref BenchCouncil Transactions on Benchmarks, Standards and Evaluations (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.13581 2025-05-21 cs.IR cs.CL cs.CR 70%

RAR: Setting Knowledge Tripwires for Retrieval Augmented Rejection

Tommaso Mario Buonocore, Enea Parimbelli

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

Comments 7 pages, 4 figures, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.11380 2025-05-20 cs.CL 70%

From the New World of Word Embeddings: A Comparative Study of Small-World Lexico-Semantic Networks in LLMs

Zhu Liu, Ying Liu, KangYang Luo, Cunliang Kong, Maosong Sun

机构 * School of Humanities, Tsinghua University(清华大学人文学院) Department of Computer Science and Technology, Tsinghua University(清华大学计算机科学与技术系)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

Comments Paper under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.11677 2025-05-20 cs.SE cs.LG 70%

Enhancing Code Quality with Generative AI: Boosting Developer Warning Compliance

Hansen Chang, Christian DeLozier

机构 * Electrical and Computer Engineering Department(电气与计算机工程系) United States Naval Academy(美国海军学院)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.10198 2025-05-20 cs.CL 70%

DioR: Adaptive Cognitive Detection and Contextual Retrieval Optimization for Dynamic Retrieval-Augmented Generation

Hanghui Guo, Jia Zhu, Shimin Di, Weijie Shi, Zhangze Chen, Jiajie Xu

机构 * Zhejiang Key Laboratory of Intelligent Education Technology and Application, Zhejiang Normal University(浙江智能教育技术与应用重点实验室,浙江师范大学) School of Computer Science and Technology, Zhejiang Normal University(浙江师范大学计算机科学与技术学院) School of Computer Science and Engineering, Southeast University(东南大学计算机科学与工程学院) Department of Computer Science and Engineering, Hong Kong University of Science and Technology(香港科技大学计算机科学与工程系) School of Computer Science and Technology, Soochow University(苏州大学计算机科学与技术学院)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

Comments Accepted to ACL2025 Main

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.08234 2025-05-14 cs.CV cs.AI cs.CR 70%

Removing Watermarks with Partial Regeneration using Semantic Information

Krti Tallam, John Kevin Cava, Caleb Geniesse, N. Benjamin Erichson, Michael W. Mahoney

机构 * International Computer Science Institute(国际计算机科学研究所) School of Computing and Augmented Intelligence(计算与增强智能学院) Lawrence Berkeley National Laboratory(劳伦斯伯克利国家实验室) Department of Statistics(统计学系) University of California at Berkeley(加州大学伯克利分校)

专题命中 其他LLM :LLM(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.06817 2025-05-13 cs.AI 70%

Control Plane as a Tool: A Scalable Design Pattern for Agentic AI Systems

Sivasathivel Kandasamy

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

Comments 2 Figures and 2 Tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.01294 2025-05-13 cs.CL 70%

Endless Jailbreaks with Bijection Learning

Brian R. Y. Huang, Maximilian Li, Leonard Tang

机构 * Haize Labs(哈伊兹实验室)

专题命中 其他LLM :LLM(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.18501 2025-05-08 cs.CL 70%

Is In-Context Learning a Type of Error-Driven Learning? Evidence from the Inverse Frequency Effect in Structural Priming

Zhenghao Zhou, Robert Frank, R. Thomas McCoy

机构 * Yale University(耶鲁大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

Comments This version is accepted to NAACL 2025 (https://aclanthology.org/2025.naacl-long.586/)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.10316 2025-05-06 cs.CV cs.AI 70%

BrushEdit: All-In-One Image Inpainting and Editing

Yaowei Li, Yuxuan Bian, Xuan Ju, Zhaoyang Zhang, Junhao Zhuang, Ying Shan, Yuexian Zou, Qiang Xu

机构 * Peking University(北京大学) ARC Lab, Tencent PCG(腾讯PCG ARC实验室) The Chinese University of Hong Kong(香港中文大学) Tsinghua University(清华大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

Comments WebPage available at https://liyaowei-stu.github.io/project/BrushEdit/

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.20668 2025-04-30 cs.CL 70%

A Generative-AI-Driven Claim Retrieval System Capable of Detecting and Retrieving Claims from Social Media Platforms in Multiple Languages

Ivan Vykopal, Martin Hyben, Robert Moro, Michal Gregor, Jakub Simko

机构 * Faculty of Information Technology, Brno University of Technology(信息技术学院,布拉格技术大学) Kempelen Institute of Intelligent Technologies(智能技术研究所)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.18673 2025-04-29 cs.CL 70%

Can Third-parties Read Our Emotions?

Jiayi Li, Yingfan Zhou, Pranav Narayanan Venkit, Halima Binte Islam, Sneha Arya, Shomir Wilson, Sarah Rajtmajer

机构 * Pennsylvania State University(宾夕法尼亚州立大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.17309 2025-04-25 cs.CL 70%

CoheMark: A Novel Sentence-Level Watermark for Enhanced Text Quality

Junyan Zhang, Shuliang Liu, Aiwei Liu, Yubo Gao, Jungang Li, Xiaojie Gu, Xuming Hu

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) Tsinghua University(清华大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

Comments Published at the 1st workshop on GenAI Watermarking, collocated with ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏