arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-10-13 至 2025-10-13 共收录 241 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 8 篇

2509.07979 2025-10-13 cs.CV 90%

Visual Representation Alignment for Multimodal Large Language Models

Heeji Yoon, Jaewoo Jung, Junwan Kim, Hyungyu Choi, Heeseong Shin, Sangbeom Lim, Honggyu An, Chaehyun Kim, Jisang Han, Donghyun Kim, Chanho Eom, Sunghwan Hong, Seungryong Kim

机构 * KAIST AI(韩国科学技术院人工智能研究所) New York University(纽约大学) Chung-Ang University(Chung-Ang 大学) Korea University(韩国大学) ETH Zürich(苏黎世联邦理工学院)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);foundation model(abstract);instruction tuning(abstract)

Comments Project Page: https://cvlab-kaist.github.io/VIRAL/

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.20684 2025-10-13 cs.RO cs.SE 89%

Identifying Uncertainty in Self-Adaptive Robotics with Large Language Models

Hassan Sartaj, Jalil Boudjadar, Mirgita Frasheri, Shaukat Ali, Peter Gorm Larsen

机构 * Simula Research Laboratory(Simula研究实验室) Aarhus University(奥胡斯大学)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);LLM(abstract)

Journal ref IEEE Software (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.05628 2025-10-13 cs.CL 85%

AnyEdit: Edit Any Knowledge Encoded in Language Models

Houcheng Jiang, Junfeng Fang, Ningyu Zhang, Guojun Ma, Mingyang Wan, Xiang Wang, Xiangnan He, Tat-seng Chua

机构 * MoE Key Lab of BIPC, University of Science(BIPC摩尔实验室,科学技术大学) National University of Singapore(新加坡国立大学) Zhejiang University(浙江大学) Douyin Co., Ltd(抖音公司)

专题命中 知识编辑与模型理解 :language model(title,abstract);LLM(abstract);large language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.09008 2025-10-13 cs.CV cs.AI cs.CL 84%

On Epistemic Uncertainty of Visual Tokens for Object Hallucinations in Large Vision-Language Models

Hoigi Seo, Dong Un Kang, Hyunjin Cho, Joohoon Lee, Se Young Chun

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.08931 2025-10-13 cs.AI cs.LG 82%

RADAR: Mechanistic Pathways for Detecting Data Contamination in LLM Evaluation

Ashish Kattamuri, Harshwardhan Fartale, Arpita Vats, Rahul Raja, Ishita Prasad

机构 * Proofpoint Indian Institute of Science(印度科学研究院) Linkedin Meta FAIR

专题命中 知识编辑与模型理解 :LLM(title,abstract);分类 cs.AI、cs.LG

Comments NeurIPS 2025 Workshop on Evaluating the Evolving LLM Lifecycle: Benchmarks, Emergent Abilities, and Scaling

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.08855 2025-10-13 cs.LG cs.AI cs.CL 75%

Time-Aware Feature Selection: Adaptive Temporal Masking for Stable Sparse Autoencoder Training

T. Ed Li, Junyu Ren

机构 * Yale University(耶鲁大学) University of Chicago(芝加哥大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

Comments First submitted on February 10th, 2025 to ICLR 2025 Workshop (XAI4Science: From Understanding Model Behavior to Discovering New Scientific Knowledge). The paper was accepted but the workshop does not generate proceedings. Now uploading to arXiv to make the paper publicly available

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.23684 2025-10-13 cs.CL 57%

Improbable Bigrams Expose Vulnerabilities of Incomplete Tokens in Byte-Level Tokenizers

Eugene Jang, Kimin Lee, Jin-Woo Chung, Keuntae Park, Seungwon Shin

机构 * Northeastern University(东北大学) KAIST(韩国科学技术院) S2W Inc.(S2W公司)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

Comments EMNLP 2025 Main

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.09135 2025-10-13 cs.CV cs.LG 57%

Training Feature Attribution for Vision Models

Aziz Bacha, Thomas George

机构 * Orange Research Châtillon, France(法国夏特隆橙色研究机构) École polytechnique, Palaiseau, France(法国帕莱索高等理工学院)

专题命中 知识编辑与模型理解 :prompting(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 其他LLM 23 篇

2510.09421 2025-10-13 cs.CL cs.AI 88%

On the Representations of Entities in Auto-regressive Large Language Models

Victor Morand, Josiane Mothe, Benjamin Piwowarski

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

Comments Accepted at BlackBoxNLP@EMNLP2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.08915 2025-10-13 cs.CL 86%

Artificial Impressions: Evaluating Large Language Model Behavior Through the Lens of Trait Impressions

Nicholas Deas, Kathleen McKeown

机构 * Columbia University(哥伦比亚大学)

专题命中 其他LLM :large language model(title);language model(title);LLM(abstract);分类 cs.CL

Comments EMNLP 2025 Camera Ready

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01644 2025-10-13 cs.CL cs.AI cs.CY 86%

Machine Learning for Detection and Analysis of Novel LLM Jailbreaks

John Hawkins, Aditya Pramar, Rodney Beard, Rohitash Chandra

机构 * Centre for Artificial Intelligence and Innovation(人工智能与创新中心) Pingla Institute(平拉研究所) Transitional Artificial Intelligence Research Group(过渡人工智能研究组) UNSW(新南威尔士大学)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.09007 2025-10-13 cs.LG 85%

LLM Unlearning on Noisy Forget Sets: A Study of Incomplete, Rewritten, and Watermarked Data

Changsheng Wang, Yihua Zhang, Dennis Wei, Jinghan Jia, Pin-Yu Chen, Sijia Liu

机构 * Michigan State University(密歇根州立大学) IBM Research(IBM研究院)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.LG

Comments Accepted by 18th ACM Workshop on Artificial Intelligence and Security (AISec'25)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.01694 2025-10-13 cs.CY cs.HC 85%

Student-AI Interaction in an LLM-Empowered Learning Environment: A Cluster Analysis of Engagement Profiles

Zhanxin Hao, Jianxiao Jiang, Jifan Yu, Zhiyuan Liu, Yu Zhang

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract)

Comments 15 pages, 10 figures; Reported on ECER 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.06594 2025-10-13 cs.CL 81%

Do Internal Layers of LLMs Reveal Patterns for Jailbreak Detection?

Sri Durga Sai Sowmya Kadali, Evangelos E. Papalexakis

机构 * Dept. of Computer Science and Engineering University of California, Riverside(计算机科学与工程系加州大学河滨分校)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.08632 2025-10-13 cs.CL cs.LG 81%

Next Semantic Scale Prediction via Hierarchical Diffusion Language Models

Cai Zhou, Chenyu Wang, Dinghuai Zhang, Shangyuan Tong, Yifei Wang, Stephen Bates, Tommi Jaakkola

机构 * Massachusetts Institute of Technology(麻省理工学院) Microsoft Research(微软研究院) Mila - Quebec AI Institute(魁北克AI研究所)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL、cs.LG

Comments Accepted to NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.09666 2025-10-13 cs.CL cs.AI cs.LG 80%

System Prompt Optimization with Meta-Learning

Yumin Choi, Jinheon Baek, Sung Ju Hwang

机构 * KAIST(韩国科学技术院)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

Comments NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.09037 2025-10-13 cs.AI cs.PL 77%

Repairing Regex Vulnerabilities via Localization-Guided Instructions

Sicheol Sung, Joonghyuk Hahn, Yo-Sub Han

机构 * Department of Computer Science, Yonsei University(延世大学计算机科学系)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

Comments 14 pages, 4 figures, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.08911 2025-10-13 cs.LG cs.NI 77%

Velocity and Density-Aware RRI Analysis and Optimization for AoI Minimization in IoV SPS

Maoxin Ji, Tong Wang, Qiong Wu, Pingyi Fan, Nan Cheng, Wen Chen

机构 * School of Internet of Things Engineering, Jiangnan University(江南大学物联网工程学院) Department of Electronic Engineering, State Key laboratory of Space Network and Communications, Beijing National Research Center for Information Science and Technology, Tsinghua University(清华大学电子工程系、空间网络与通信国家重点实验室、北京信息科学与技术国家研究中心) State Key Lab. of ISN and School of Telecommunications Engineering, Xidian University(西安电子科技大学信息与通信系统国家重点实验室、电信工程学院) Department of Electronic Engineering, Shanghai Jiao Tong University(上海交通大学电子工程系)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.LG

Comments This paper has been submitted to IEEE Communications Letters

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.16548 2025-10-13 cs.HC cs.AI cs.ET 77%

Exploring human-SAV interaction using LLMs: The impact of psychological factors on user experience

Lirui Guo, Michael G. Burke, Wynita M. Griggs

机构 * Monash University(墨尔本大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

Comments This work has been submitted to the IEEE for possible publication

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.08083 2025-10-13 cs.HC cs.AI q-bio.NC 70%

Confidence-weighted integration of human and machine judgments for superior decision-making

Felipe Yáñez, Xiaoliang Luo, Omar Valerio Minero, Bradley C. Love

机构 * Max Planck Institute for Neurobiology of Behavior – caesar, Bonn, Germany(马克斯·普朗克行为神经生物学研究所——caesar,波恩,德国) Department of Experimental Psychology, University College London, London, United Kingdom(实验心理学系,伦敦大学学院,伦敦,英国)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.08912 2025-10-13 cs.HC 67%

Beyond Words: Infusing Conversational Agents with Human-like Typing Behaviors

Jijie Zhou, Yuhan Hu

专题命中 其他LLM :large language model(abstract);language model(abstract)

Comments Author's version of a paper published at CUI '24 (ACM Conversational User Interfaces 2024)

Journal ref CUI '24: Proceedings of the ACM Conversational User Interfaces 2024, July 8-10, 2024, Luxembourg, Luxembourg. ACM, New York, NY, USA, 11 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06394 2025-10-13 cs.CL cs.LG 62%

How a Bilingual LM Becomes Bilingual: Tracing Internal Representations with Sparse Autoencoders

Tatsuro Inaba, Go Kamoda, Kentaro Inui, Masaru Isonuma, Yusuke Miyao, Yohei Oseki, Benjamin Heinzerling, Yu Takagi

机构 * MBZUAI(穆萨大学人工智能研究所) SOKENDAI(神户大学) NINJAL(日本国家研究所) Tohoku University(东北大学) RIKEN(日本研究机构) NII LLMC(日本信息处理学会大语言模型中心) University of Tokyo(东京大学) Nagoya Institute of Technology(名古屋技术大学)

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.LG

Comments 13 pages, 17 figures, accepted to EMNLP 2025 findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.08829 2025-10-13 cs.CR cs.AI cs.LG 62%

CommandSans: Securing AI Agents with Surgical Precision Prompt Sanitization

Debeshee Das, Luca Beurer-Kellner, Marc Fischer, Maximilian Baader

机构 * ETH Zurich(苏黎世联邦理工学院) Snyk(Snyk公司)

专题命中 其他LLM :LLM(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.08755 2025-10-13 cs.AI cs.CL cs.NI 62%

Robust Heuristic Algorithm Design with LLMs

Pantea Karimi, Dany Rouhana, Pooria Namyar, Siva Kesava Reddy Kakarla, Venkat Arun, Behnaz Arzani

机构 * MIT(麻省理工学院) Microsoft(微软公司) University of Southern California(南加州大学) Microsoft Research(微软研究院) The University of Texas at Austin(德克萨斯大学奥斯汀分校)

专题命中 其他LLM :LLM(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.19528 2025-10-13 cs.CV cs.AI cs.GR cs.LG 62%

RAGDiffusion: Faithful Cloth Generation via External Knowledge Assimilation

Xianfeng Tan, Yuhan Li, Wenxiang Shang, Yubo Wu, Jian Wang, Xuanhong Chen, Yi Zhang, Ran Lin, Bingbing Ni

机构 * Shanghai Jiao Tong University(上海交通大学) Alibaba Group(阿里巴巴集团)

专题命中 其他LLM :language model(abstract);分类 cs.AI、cs.LG

Comments Accept by ICCV 2025 (Highlight). Project website: https://colorful-liyu.github.io/RAGDiffusion-page/

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.09267 2025-10-13 cs.RO cs.LG 57%

Placeit! A Framework for Learning Robot Object Placement Skills

Amina Ferrad, Johann Huber, François Hélénon, Julien Gleyze, Mahdi Khoramshahi, Stéphane Doncieux

机构 * Sorbonne Université, CNRS, Institut des Systèmes Intelligents et de Robotique, ISIR(索邦大学、国家科学研究中心、智能系统与机器人研究所、ISIR)

专题命中 其他LLM :foundation model(abstract);分类 cs.LG

Comments 8 pages, 8 figures. Draft version

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01298 2025-10-13 q-bio.QM cs.CV cs.LG 57%

MorphGen: Controllable and Morphologically Plausible Generative Cell-Imaging

Berker Demirel, Marco Fumero, Theofanis Karaletsos, Francesco Locatello

专题命中 其他LLM :foundation model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.09078 2025-10-13 cs.GR cs.LG 57%

MCMC: Bridging Rendering, Optimization and Generative AI

Gurprit Singh, Wenzel Jakob

机构 * Max Planck Institute for Informatics(马克斯·普朗克研究所信息学研究所) EPFL(瑞士联邦理工学院)

专题命中 其他LLM :language model(abstract);分类 cs.LG

Comments SIGGRAPH Asia 2024 Courses. arXiv admin note: text overlap with arXiv:2208.11970 by other authors

Journal ref SIGGRAPH Asia 2024 Courses, Article No.: 8, Pages 1 - 27

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.08876 2025-10-13 cs.SE cs.AI 57%

Vector Graph-Based Repository Understanding for Issue-Driven File Retrieval

Kostiantyn Bevziuk, Andrii Fatula, Svetozar Lashin Yaroslav Opanasenko, Anna Tukhtarova, Ashok Jallepalli Pradeepkumar Sharma, Hritvik Shrivastava

专题命中 其他LLM :LLM(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.09522 2025-10-13 physics.chem-ph cond-mat.mtrl-sci 50%

Is Platinum a Proton Blocking Catalyst?

Aparna Saksena, Yujun Zhao, J. Manoj Prabhakar, Dierk Raabe, Baptiste Gault, Yug Joshi

专题命中 其他LLM :prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏