arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-10-22 至 2025-10-22 共收录 17 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 17 篇

2510.18476 2025-10-22 cs.AI cs.CL 86%

Probabilistic Modeling of Intentions in Socially Intelligent LLM Agents

Feifan Xia, Yuyang Fang, Defang Li, Yantong Xie, Weikang Li, Yang Li, Deguo Xia, Jizhou Huang

机构 * Baidu Inc(百度公司) Imperial College London(伦敦帝国学院) Zhejiang University(浙江大学) Carnegie Mellon University(卡内基梅隆大学) Peking University(北京大学)

专题命中 知识编辑与模型理解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09423 2025-10-22 cs.CV cs.RO 85%

Distilling LLM Prior to Flow Model for Generalizable Agent's Imagination in Object Goal Navigation

Badi Li, Ren-jie Lu, Yu Zhou, Jingke Meng, Wei-shi Zheng

机构 * Sun Yat-sen University(中山大学) The University of Hong Kong(香港大学) Key Laboratory of Machine Intelligence and Advanced Computing, Ministry of Education(教育部机器智能与高级计算重点实验室)

专题命中 知识编辑与模型理解 :LLM(title,abstract);large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.17148 2025-10-22 cs.SE cs.AI 83%

LLM Agents for Interactive Exploration of Historical Cadastre Data: Framework and Application to Venice

Tristan Karch, Jakhongir Saydaliev, Isabella Di Lenardo, Frédéric Kaplan

机构 * DH-Lab, EPFL, Lausanne, Switzerland(DH实验室,日内瓦联邦理工学院,洛桑,瑞士)

专题命中 知识编辑与模型理解 :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

Comments Accepted in Cambridge press - Computational Humanities Research 2025

Journal ref Comput. humanit. res. 1 (2025) e11

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17941 2025-10-22 cs.CL cs.AI 82%

Believe It or Not: How Deeply do LLMs Believe Implanted Facts?

Stewart Slocum, Julian Minder, Clément Dumas, Henry Sleight, Ryan Greenblatt, Samuel Marks, Rowan Wang

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24262 2025-10-22 q-bio.QM cs.AI cs.LG 81%

LAMP-PRo: Label-aware Attention for Multi-label Prediction of DNA- and RNA-binding Proteins using Protein Language Models

Nimisha Ghosh, Dheeran Sankaran, Rahul Balakrishnan Adhi, Sharath S, Amrut Anand

机构 * Department of Computer Science and Engineering, Shiv Nadar University Chennai, Tamil Nadu, India(计算机科学与工程系,Shiv Nadar大学 Chennai,印度 Tamil Nadu)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.09572 2025-10-22 cs.CL 79%

Rethinking LLM Uncertainty: A Multi-Agent Approach to Estimating Black-Box Model Uncertainty

Yu Feng, Phu Mon Htut, Zheng Qi, Wei Xiao, Manuel Mager, Nikolaos Pappas, Kishaloy Halder, Yang Li, Yassine Benajiba, Dan Roth

机构 * University of Pennsylvania(宾夕法尼亚大学) AWS AI Labs(AWS人工智能实验室) Johannes Gutenberg University of Mainz(美因茨约翰内斯·古滕贝格大学) Oracle AI(Oracle人工智能)

专题命中 知识编辑与模型理解 :LLM(title,abstract);分类 cs.CL

Comments EMNLP 2025 Findings

Journal ref EMNLP 2025 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17942 2025-10-22 cs.CY cs.AI 79%

Trust in foundation models and GenAI: A geographic perspective

Grant McKenzie, Krzysztof Janowicz, Carsten Kessler

机构 * McGill University, Canada(麦吉尔大学,加拿大) University of Vienna, Austria(维也纳大学,奥地利) Bochum University of Applied Sciences, Germany(波鸿应用科学大学,德国) Aalborg University Copenhagen, Denmark(奥胡斯大学哥本哈根分校,丹麦)

专题命中 知识编辑与模型理解 :foundation model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17833 2025-10-22 q-bio.NC cs.AI 79%

Brain-Language Model Alignment: Insights into the Platonic Hypothesis and Intermediate-Layer Advantage

Ángela López-Cardona, Sebastián Idesis, Mireia Masias-Bruns, Sergi Abadal, Ioannis Arapakis

机构 * Universitat Politècnica de Catalunya(加泰罗尼亚理工大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17918 2025-10-22 cs.CL cs.AI 79%

JT-Safe: Intrinsically Enhancing the Safety and Trustworthiness of LLMs

Junlan Feng, Fanyu Meng, Chong Long, Pengyu Cong, Duqing Wang, Yan Zheng, Yuyao Zhang, Xuanchang Gao, Ye Yuan, Yunfei Ma, Zhijie Ren, Fan Yang, Na Wu, Di Jin, Chao Deng

机构 * China Mobile Jiutian Research(中国移动九天研究所)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);post-training(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17910 2025-10-22 cs.CY cs.AI cs.CL 79%

Interpretability Framework for LLMs in Undergraduate Calculus

Sagnik Dakshit, Sushmita Sinha Roy

机构 * University of Texas at Tyler(德克萨斯理工大学) Florida Gulf Coast University(佛罗里达盖恩斯维尔大学)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17909 2025-10-22 cs.CL 74%

Atomic Literary Styling: Mechanistic Manipulation of Prose Generation in Neural Language Models

Tsogt-Ochir Enkhbayar

机构 * Mongol AI(蒙古AI)

专题命中 知识编辑与模型理解 :language model(title);分类 cs.CL

Comments 12 pages, 3 figures, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17842 2025-10-22 cs.SE cs.HC 67%

Vibe Coding: Toward an AI-Native Paradigm for Semantic and Intent-Driven Programming

Vinay Bamil

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

Comments 10 pages, 1 figure, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.11741 2025-10-22 cs.AI cs.CR 57%

MTRE: Multi-Token Reliability Estimation for Hallucination Detection in VLMs

Geigh Zollicoffer, Minh Vu, Manish Bhattarai

机构 * Los Alamos National Laboratory(洛斯阿拉莫斯国家实验室)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.04641 2025-10-22 cs.LG math.ST stat.ML stat.TH 57%

A Statistical Theory of Contrastive Pre-training and Multimodal Generative AI

Kazusato Oko, Licong Lin, Yuhang Cai, Song Mei

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.12539 2025-10-22 cs.AI cs.MA 57%

Counterfactual Effect Decomposition in Multi-Agent Sequential Decision Making

Stelios Triantafyllou, Aleksa Sukovic, Yasaman Zolfimoselo, Goran Radanovic

机构 * Max Planck Institute for Software Systems(马克斯·普朗克软件系统研究所)

专题命中 知识编辑与模型理解 :LLM(abstract);分类 cs.AI

Comments ICML 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.18703 2025-10-22 cs.CV 50%

Exploring a Unified Vision-Centric Contrastive Alternatives on Multi-Modal Web Documents

Yiqi Lin, Alex Jinpeng Wang, Linjie Li, Zhengyuan Yang, Mike Zheng Shou

机构 * Show Lab, National University of Singapore(新加坡国立大学展示实验室) Central South University(中南大学) Microsoft(微软公司)

专题命中 知识编辑与模型理解 :language model(abstract)

Comments Project page: this https://linyq17.github.io/VC2L/

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.18321 2025-10-22 cs.CV 50%

Beyond Single Models: Mitigating Multimodal Hallucinations via Adaptive Token Ensemble Decoding

Jinlin Li, Yuran Wang, Yifei Yuan, Xiao Zhou, Yingying Zhang, Xixian Yong, Yefeng Zheng, Xian Wu

机构 * Gaoling School of Artificial Intelligence, Renmin University of China(中国人民大学 Gallup 学院) Department of Electrical and Computer Engineering, McGill University(麦吉尔大学电气与计算机工程系) School of Statistics, Renmin University of China(中国人民大学统计学院) Tencent Jarvis Lab(腾讯 Jarvis 实验室) Medical Artificial Intelligence Lab, Westlake University(西湖大学医学人工智能实验室)

专题命中 知识编辑与模型理解 :language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏