arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-09-29 至 2025-09-29 共收录 243 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 19 篇

2506.22937 2025-09-29 cs.HC 50%

GamerAstra: Supporting 2D Non-Twitch Video Games for Blind and Low-Vision Players through a Multi-Agent Framework

Tianrun Qiu, Changxin Chen, Sizhe Cheng, Xuyang Liu, Xumeng Wang, Zhicong Lu, Yuxin Ma

专题命中 领域大模型 :language model(abstract)

Comments 17 pages, 11 figures, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 知识编辑与模型理解 12 篇

2509.22121 2025-09-29 cs.LG 89%

Mind the Missing: Variable-Aware Representation Learning for Irregular EHR Time Series using Large Language Models

Jeong Eul Kwon, Joo Heung Yoon, Hyo Kyung Lee

机构 * School of Industrial Management Engineering, Korea University(韩国大学工业管理工程学院) Division of Pulmonary, Allergy, Critical Care, and Sleep Medicine, Department of Medicine, University of Pittsburgh(匹兹堡大学医学部呼吸科、过敏科、重症医学科及睡眠医学科)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21534 2025-09-29 cs.LG 88%

A circuit for predicting hierarchical structure in-context in Large Language Models

Tankred Saanum, Can Demircan, Samuel J. Gershman, Eric Schulz

机构 * Harvard University(哈佛大学) Institute for Human-Centered AI(以人为中心的人工智能研究所) Helmholtz Computational Health Center(海德堡计算健康中心)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.16975 2025-09-29 cs.LG cs.AI cs.CL 85%

Latent Concept Disentanglement in Transformer-based Language Models

Guan Zhe Hong, Bhavya Vasudeva, Vatsal Sharan, Cyrus Rashtchian, Prabhakar Raghavan, Rina Panigrahy

机构 * Purdue University(普渡大学) University of Southern California(南加州大学) Google(谷歌)

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.12881 2025-09-29 cs.CL 79%

TEXT2AFFORD: Probing Object Affordance Prediction abilities of Language Models solely from Text

Sayantan Adak, Daivik Agrawal, Animesh Mukherjee, Somak Aditya

机构 * IIT, Kharagpur(印度Kharagpur理工学院)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

Comments Accepted at Conference on Computational Natural Language Learning 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.22251 2025-09-29 cs.CL cs.AI 79%

Beyond Textual Context: Structural Graph Encoding with Adaptive Space Alignment to alleviate the hallucination of LLMs

Yifang Zhang, Pengfei Duan, Yiwen Yang, Shengwu Xiong

机构 * Wuhan University of Technology(武汉理工大学)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments 11 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21999 2025-09-29 cs.CL cs.AI 79%

Black-Box Hallucination Detection via Consistency Under the Uncertain Expression

Seongho Joo, Kyungmin Min, Jahyun Koo, Kyomin Jung

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21357 2025-09-29 cs.CL cs.AI 73%

A Novel Differential Feature Learning for Effective Hallucination Detection and Classification

Wenkai Wang, Vincent Lee, Yizhen Zheng

机构 * Department of Data Science and AI(数据科学与人工智能系) Monash University(墨尔本大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments 10 pages, 7 figures, 13 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.17445 2025-09-29 cs.CL 70%

Semantic Reformulation Entropy for Robust Hallucination Detection in QA Tasks

Chaodong Tong, Qi Zhang, Lei Jiang, Yanbing Liu, Nannan Sun, Wei Li

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

Comments 5pages, 5 figures, submitted to ICASSP 2026,

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.11029 2025-09-29 cs.LG 70%

Exploiting the Asymmetric Uncertainty Structure of Pre-trained VLMs on the Unit Hypersphere

Li Ju, Max Andersson, Stina Fredriksson, Edward Glöckner, Andreas Hellander, Ekta Vats, Prashant Singh

机构 * Department of Information Technology, Uppsala University(信息科技系,乌普萨拉大学) Science for Life Laboratory, Uppsala University(生命科学实验室,乌普萨拉大学)

专题命中 知识编辑与模型理解 :language model(abstract);foundation model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21997 2025-09-29 cs.CV 67%

Exposing Hallucinations To Suppress Them: VLMs Representation Editing With Generative Anchors

Youxu Shi, Suorong Yang, Dong Liu

机构 * University of Science and Technology of China(中国科学技术大学) Nanjing University(南京大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.01986 2025-09-29 cs.CL cs.AI cs.LG 67%

Adaptively profiling models with task elicitation

Davis Brown, Prithvi Balehannina, Helen Jin, Shreya Havaldar, Hamed Hassani, Eric Wong

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI、cs.LG

Comments EMNLP 2025 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21695 2025-09-29 cs.LG 57%

Wav2Arrest 2.0: Long-Horizon Cardiac Arrest Prediction with Time-to-Event Modeling, Identity-Invariance, and Pseudo-Lab Alignment

Saurabh Kataria, Davood Fattahi, Minxiao Wang, Ran Xiao, Matthew Clark, Timothy Ruchti, Mark Mai, Xiao Hu

机构 * Nell Hodgson Woodruff School of Nursing, Emory University(埃默里大学护理学院) Department of Pediatrics, Emory School of Medicine(埃默里医学院儿科部)

专题命中 知识编辑与模型理解 :foundation model(abstract);分类 cs.LG

Comments Submitted to BPSC

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 其他LLM 20 篇

2509.22206 2025-09-29 cs.CL cs.AI 88%

The Outputs of Large Language Models are Meaningless

Anandi Hattiangadi, Anders J. Schoubye

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

Comments 24 pages, 2 figures, forthcoming in Herman Cappelen and Rachel Sterken, eds. Communicating with AI: Philosophical Perspectives. Oxford: Oxford University Press

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21424 2025-09-29 physics.chem-ph cs.AI 88%

PhenoMoler: Phenotype-Guided Molecular Optimization via Chemistry Large Language Model

Ran Song, Hui Liu

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.02091 2025-09-29 cs.CL cs.LG 86%

LLM-OptiRA: LLM-Driven Optimization of Resource Allocation for Non-Convex Problems in Wireless Communications

Xinyue Peng, Yanming Liu, Yihan Cang, Chaoqun Cao, Ming Chen

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

Comments 6 pages,4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.11022 2025-09-29 cs.SE cs.AI cs.CL cs.CR cs.LG 86%

Security Degradation in Iterative AI Code Generation -- A Systematic Analysis of the Paradox

Shivani Shukla, Himanshu Joshi, Romilla Syed

机构 * Department of Analytics and Information Systems(分析与信息系统系) University of San Francisco(旧金山大学) Department of Applied AI and Industry Innovation(应用人工智能与产业创新系) Vector Institute for Artificial Intelligence(人工智能研究院) Department of Management Science and Information Systems(管理科学与信息系统系) University of Massachusetts Boston(马萨诸塞大学波士顿分校)

专题命中 其他LLM :LLM(abstract,comments);large language model(abstract,comments);language model(abstract,comments);prompting(abstract,comments)

Comments Keywords - Large Language Models, Security Vulnerabilities, AI-Generated Code, Iterative Feedback, Software Security, Secure Coding Practices, Feedback Loops, LLM Prompting Strategies

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.22447 2025-09-29 cs.AI cs.NE 83%

Guiding Evolution of Artificial Life Using Vision-Language Models

Nikhil Baid, Hannah Erlebach, Paul Hellegouarch, Frederico Wieser

机构 * University College London(伦敦大学学院) Institut Pasteur(巴斯德研究院)

专题命中 其他LLM :language model(title,abstract);foundation model(abstract);分类 cs.AI

Comments 9 pages, 6 figures. Accepted for publication in the Proceedings of the Artificial Life Conference 2025 (MIT Press)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21602 2025-09-29 cs.RO 82%

Real-Time Indoor Object SLAM with LLM-Enhanced Priors

Yang Jiao, Yiding Qiu, Henrik I. Christensen

机构 * Students of Contextual Robotics Institute, University of California San Diego(情境机器人研究所研究生,加州大学圣地亚哥分校) Faculty of the Department of Computer Science and Engineering, University of California San Diego(计算机科学与工程系 faculty,加州大学圣地亚哥分校)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21673 2025-09-29 cs.LG cs.AI 81%

SlotFM: A Motion Foundation Model with Slot Attention for Diverse Downstream Tasks

Junyong Park, Oron Levy, Rebecca Adaimi, Asaf Liberman, Gierad Laput, Abdelkareem Bedri

机构 * Apple(苹果公司) KAIST(韩国科学技术院)

专题命中 其他LLM :foundation model(title,abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.16355 2025-09-29 cs.LG cs.AI 81%

How Strategic Agents Respond: Comparing Analytical Models with LLM-Generated Responses in Strategic Classification

Tian Xie, Pavan Rauch, Xueru Zhang

机构 * The Ohio State University(俄亥俄州立大学)

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI、cs.LG

Comments Add GPT 5 experiments

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.22030 2025-09-29 cs.CL 79%

From Outliers to Topics in Language Models: Anticipating Trends in News Corpora

Evangelia Zve, Benjamin Icard, Alice Breton, Lila Sainero, Gauvain Bourgne, Jean-Gabriel Ganascia

机构 * LIP6, Sorbonne University, CNRS, France(LIP6,索邦大学,国家科学研究中心,法国)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

Comments presented at ICNLSP 2025; to appear in the ACL Anthology; received the Best Full Paper Award

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.22638 2025-09-29 cs.CL cs.AI cs.LG 78%

Language Models Can Learn from Verbal Feedback Without Scalar Rewards

Renjie Luo, Zichen Liu, Xiangyan Liu, Chao Du, Min Lin, Wenhu Chen, Wei Lu, Tianyu Pang

机构 * Sea AI Lab(海智实验室) SUTD(新加坡科技设计大学) NUS(国立大学) NTU(南洋理工大学) University of Waterloo(滑铁卢大学)

专题命中 其他LLM :language model(title);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21360 2025-09-29 cs.CV cs.AI 77%

Multimodal Prompt Decoupling Attack on the Safety Filters in Text-to-Image Models

Xingkai Peng, Jun Jiang, Meng Tong, Shuai Li, Weiming Zhang, Nenghai Yu, Kejiang Chen

机构 * University of Science and Technology of China(中国科学技术大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21665 2025-09-29 cs.HC 75%

Alignment Without Understanding: A Message- and Conversation-Centered Approach to Understanding AI Sycophancy

Lihua Du, Xing Lyu, Lezi Xie, Bo Feng

专题命中 其他LLM :large language model(abstract);language model(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.17550 2025-09-29 cs.LG cs.AI stat.ML 73%

In-Context Algorithm Emulation in Fixed-Weight Transformers

Jerry Yao-Chieh Hu, Hude Liu, Jennifer Yuntong Zhang, Han Liu

专题命中 其他LLM :foundation model(abstract);prompting(abstract);分类 cs.AI、cs.LG

Comments Code is available at https://github.com/MAGICS-LAB/algo_emu

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.17993 2025-09-29 cs.CL 70%

DRS: Deep Question Reformulation With Structured Output

Zhecheng Li, Yiwei Wang, Bryan Hooi, Yujun Cai, Nanyun Peng, Kai-Wei Chang

机构 * University of California, San Diego(加州大学圣地亚哥分校) University of California, Los Angelas(加州大学洛杉矶分校) The University of Queensland(昆士兰大学) National University of Singapore(新加坡国立大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

Comments Accepted to ACL 2025 (Findings)

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.09265 2025-09-29 cs.CL 70%

Sharing Matters: Analysing Neurons Across Languages and Tasks in LLMs

Weixuan Wang, Barry Haddow, Minghao Wu, Wei Peng, Alexandra Birch

机构 * School of Informatics, University of Edinburgh(信息学院,爱丁堡大学) Monash University(墨尔本大学) Huawei Technologies Co., Ltd.(华为技术有限公司)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.22393 2025-09-29 cs.CV 67%

Text Adversarial Attacks with Dynamic Outputs

Wenqiang Wang, Siyuan Liang, Xiao Yan, Xiaochun Cao

专题命中 其他LLM :large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21747 2025-09-29 cs.CV 67%

Incorporating Scene Context and Semantic Labels for Enhanced Group-level Emotion Recognition

Qing Zhu, Wangdong Guo, Qirong Mao, Xiaohua Huang, Xiuyan Shao, Wenming Zheng

机构 * School of Computer Science and Communication Engineering, Jiangsu University(江苏大学计算机科学与通信工程学院) Oulu School, Nanjing Institute of Technology(南京理工大学奥卢学院) School of Management, Southeast University(东南大学管理学院) Key Laboratory of Child Development and Learning Science (Southeast University), Ministry of Education, Southeast University(教育部儿童发展与学习科学重点实验室(东南大学))

专题命中 其他LLM :large language model(abstract);language model(abstract)

Comments 10 pages, 5figures, submitted to IEEE Transactions on Human-Machine Systems

详情

展开后加载摘要…

URL PDF HTML 收藏