arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 7596 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 7596 篇

2509.18997 2025-10-03 cs.LG 57%

Theoretical Foundations of Representation Learning using Unlabeled Data: Statistics and Optimization

Pascal Esser, Maximilian Fleissner, Debarghya Ghoshdastidar

机构 * Ludwig-Maximilians-Universität München(慕尼黑路德维希-马克西米利安大学) Technical University of Munich(慕尼黑技术大学) TUM School of Computation, Information and Technology(慕尼黑技术大学计算、信息与技术学院) Munich Data Science Institute (MDSI)(慕尼黑数据科学研究所) Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心)

专题命中 知识编辑与模型理解 :foundation model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.11627 2025-10-03 cs.LG 57%

Enhancing Electricity-System Resilience with Adaptive Robust Optimization and Conformal Uncertainty Characterization

Shuyi Chen, Shixiang Zhu, Ramteen Sioshansi

机构 * Heinz College of Information Systems and Public Policy, Carnegie Mellon University(信息系统与公共政策学院,卡内基梅隆大学) Carnegie Mellon Electricity Industry Center(卡内基梅隆电力产业中心) Wilton E. Scott Institute for Energy Innovation(威利特·E·斯科特能源创新研究所) Department of Engineering and Public Policy, Carnegie Mellon Electricity Industry Center(工程与公共政策系,卡内基梅隆电力产业中心) Department of Electrical and Computer Engineering, Heinz College of Information Systems and Public Policy(电气与计算机工程系,信息系统与公共政策学院) Department of Integrated Systems Engineering, The Ohio State University(整合系统工程系,俄亥俄州立大学)

专题命中 知识编辑与模型理解 :prompting(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24431 2025-09-30 cs.LG 57%

Semantic Compression via Multimodal Representation Learning

Eleonora Grassucci, Giordano Cicchetti, Aurelio Uncini, Danilo Comminiello

机构 * Dept. of Information Engineering, Electronics, and Telecomm.(信息工程、电子与电信系)

专题命中 知识编辑与模型理解 :post-training(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.00310 2025-09-30 cs.RO cs.AI 57%

TReF-6: Inferring Task-Relevant Frames from a Single Demonstration for One-Shot Skill Generalization

Yuxuan Ding, Shuangge Wang, Tesca Fitzgerald

机构 * Yale University(耶鲁大学)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.23717 2025-09-30 cs.AI 57%

Measuring Sparse Autoencoder Feature Sensitivity

Claire Tian, Katherine Tian, Nathan Hu

机构 * The Harker School(哈克尔学校) Stanford University(斯坦福大学)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI

Comments NeurIPS 2025 Workshop on Mechanistic Interpretability Camera Ready

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.12341 2025-09-30 cs.CV cs.AI 57%

Semantic Discrepancy-aware Detector for Image Forgery Identification

Ziye Wang, Minghang Yu, Chunyan Xu, Zhen Cui

机构 * Nanjing University of Science and Technology(南京理工大学) Beijing Normal University(北京师范大学)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI

Comments 10 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.15963 2025-09-30 cs.CV cs.CL 57%

OViP: Online Vision-Language Preference Learning for VLM Hallucination

Shujun Liu, Siyuan Wang, Zejun Li, Jianxiang Wang, Cheng Zeng, Zhongyu Wei

机构 * Fudan University(复旦大学) University of Southern California(南加州大学) ByteDance(字节跳动)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21695 2025-09-29 cs.LG 57%

Wav2Arrest 2.0: Long-Horizon Cardiac Arrest Prediction with Time-to-Event Modeling, Identity-Invariance, and Pseudo-Lab Alignment

Saurabh Kataria, Davood Fattahi, Minxiao Wang, Ran Xiao, Matthew Clark, Timothy Ruchti, Mark Mai, Xiao Hu

机构 * Nell Hodgson Woodruff School of Nursing, Emory University(埃默里大学护理学院) Department of Pediatrics, Emory School of Medicine(埃默里医学院儿科部)

专题命中 知识编辑与模型理解 :foundation model(abstract);分类 cs.LG

Comments Submitted to BPSC

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.02147 2025-09-26 cs.CL 57%

BabyLM's First Constructions: Causal probing provides a signal of learning

Joshua Rozner, Leonie Weissweiler, Cory Shain

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

Comments Accepted to EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.20065 2025-09-25 cs.CL 57%

From Input Perception to Predictive Insight: Modeling Model Blind Spots Before They Become Errors

Maggie Mi, Aline Villavicencio, Nafise Sadat Moosavi

机构 * University of Sheffield(谢菲尔德大学) University of Exeter(埃克塞特大学) The Alan Turing Institute(艾伦·图灵研究所) UFRN, Brazil(巴西UFRN)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

Comments EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.13390 2025-09-25 cs.CL 57%

Aligned Probing: Relating Toxic Behavior and Model Internals

Andreas Waldis, Vagrant Gautam, Anne Lauscher, Dietrich Klakow, Iryna Gurevych

机构 * Ubiquitous Knowledge Processing Lab (UKP Lab)(通用知识处理实验室) Technical University of Darmstadt(德累斯顿技术大学) Information Systems Research Lab(信息系统研究实验室) Lucerne University of Applied Sciences and Arts(卢塞恩应用科学与艺术大学) Spoken Language Systems(语音语言系统) Saarland University(萨尔兰大学) Data Science Group(数据科学组) University of Hamburg(汉堡大学)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.18691 2025-09-24 cs.SD cs.AI eess.AS 57%

An overview of neural architectures for self-supervised audio representation learning from masked spectrograms

Sarthak Yadav, Sergios Theodoridis, Zheng-Hua Tan

机构 * Department of Electronic Systems, Aalborg University(电子系统系,奥尔堡大学) Pioneer Centre for Artificial Intelligence(人工智能先锋中心) National and Kapodistrian University of Athens(雅典国家与卡波蒂斯坦大学)

专题命中 知识编辑与模型理解 :foundation model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.17482 2025-09-23 cs.CL 57%

Diagnosing Model Editing via Knowledge Spectrum

Tsung-Hsuan Pan, Chung-Chi Chen, Hen-Hsen Huang, Hsin-Hsi Chen

机构 * Department of Computer Science and Information Engineering(计算机科学与信息工程系) National Taiwan University(国立台湾大学) Artificial Intelligence Research Center(人工智能研究中心) AIST Japan(日本AIST) Institute of Information Science Academia Sinica(学术院信息科学研究所)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.14067 2025-09-22 cs.CV cs.AI 57%

VLA-Mark: A cross modal watermark for large vision-language alignment model

Shuliang Liu, Qi Zheng, Jesse Jiaxi Xu, Yibo Yan, Junyan Zhang, He Geng, Aiwei Liu, Peijie Jiang, Jia Liu, Yik-Cheung Tam, Xuming Hu

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) The Hong Kong University of Science and Technology(香港科技大学) University of Toronto(多伦多大学) Ant Group, Alibaba(蚂蚁集团,阿里巴巴) New York University Shanghai(纽约大学上海分校)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI

Comments Accepted by the main conference, EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.16846 2025-09-18 eess.AS cs.CL cs.SD 57%

KALL-E:Autoregressive Speech Synthesis with Next-Distribution Prediction

Kangxiang Xia, Xinfa Zhu, Jixun Yao, Wenjie Tian, Wenhao Li, Lei Xie

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

Comments 6 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.19505 2025-09-18 cs.AI 57%

Caught in the Act: a mechanistic approach to detecting deception

Gerard Boxo, Ryan Socha, Daniel Yoo, Shivam Raval

机构 * Barcelona Institute of Science and Technology(巴塞罗那科学与技术研究所) NorthWest Arkansas Community College(西北阿肯色社区学院) Carnegie Mellon University(卡内基梅隆大学) Harvard University(哈佛大学)

专题命中 知识编辑与模型理解 :LLM(abstract);分类 cs.AI

Comments 15 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.09315 2025-09-17 cs.RO cs.CV cs.LG 57%

TransDiffuser: Diverse Trajectory Generation with Decorrelated Multi-modal Representation for End-to-end Autonomous Driving

Xuefeng Jiang, Yuan Ma, Pengxiang Li, Leimeng Xu, Xin Wen, Kun Zhan, Zhongpu Xia, Peng Jia, Xianpeng Lang, Sheng Sun

机构 * LiAuto Inc(LiAuto公司) Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所) School of Vehicle and Mobility, Tsinghua University(清华大学车辆与移动研究所)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.LG

Comments Under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.07586 2025-09-15 cs.LG 57%

Unveiling Group-Specific Distributed Concept Drift: A Fairness Imperative in Federated Learning

Teresa Salazar, João Gama, Helder Araújo, Pedro Henriques Abreu

机构 * CISUC/LASI - Centre for Informatics and Systems of the University of Coimbra, Department of Informatics Engineering, University of Coimbra(科瓦姆布拉大学信息与系统中心/实验室,信息工程系,科瓦姆布拉大学) INESC TEC and the Faculty of Economy, University of Porto(INESC TEC和波尔图大学经济学院) Institute of Systems and Robotics, Department of Electrical and Computer Engineering, University of Coimbra(系统与机器人研究所,电子与计算机工程系,科瓦姆布拉大学)

专题命中 知识编辑与模型理解 :prompting(abstract);分类 cs.LG

Comments accepted for publication in IEEE Transactions on Neural Networks and Learning Systems (early access, Sep. 2025)

Journal ref IEEE Transactions on Neural Networks and Learning Systems (early access, Sep. 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.09541 2025-09-12 cs.AI 57%

Compositional Concept Generalization with Variational Quantum Circuits

Hala Hawashin, Mina Abbaszadeh, Nicholas Joseph, Beth Pearson, Martha Lewis, Mehrnoosh sadrzadeh

机构 * School of Computer Science Engineering University of New South Wales Sydney, Australia Stanford University California, USA Computer Science University College London London, UK School of Eng. Maths. \& Tech University of Bristol Bristol, UK Inst. Logic Language \& Computation University of Amsterdam Amsterdam, NL

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI

Comments Accepted to: 2025 IEEE International Conference on Quantum Artificial Intelligence (QAI), Naples, Italy, Nov 2-5, 2025. This is the authors' accepted manuscript (AAM). An IEEE copyright notice appears on page 1. The final published version will appear in IEEE Xplore; DOI to be added when available

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.11639 2025-09-12 eess.SP cs.LG stat.ML 57%

Recursive KalmanNet: Deep Learning-Augmented Kalman Filtering for State Estimation with Consistent Uncertainty Quantification

Hassan Mortada, Cyril Falcon, Yanis Kahil, Mathéo Clavaud, Jean-Philippe Michel

机构 * Exail

专题命中 知识编辑与模型理解 :prompting(abstract);分类 cs.LG

Comments 5 pages, 3 figures. Accepted for publication in EUSIPCO 2025 proceedings

Journal ref 33rd European Signal Processing Conference, (2025), pp. 885-889

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.08778 2025-09-11 cs.CL 57%

Do All Autoregressive Transformers Remember Facts the Same Way? A Cross-Architecture Analysis of Recall Mechanisms

Minyeong Choe, Haehyun Cho, Changho Seo, Hyunil Kim

机构 * Chosun University(全州大学) Soongsil University(顺天大学) Kongju National University(康州国立大学)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

Comments Accepted at EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.08640 2025-09-11 eess.IV cs.AI cs.CV 57%

RoentMod: A Synthetic Chest X-Ray Modification Model to Identify and Correct Image Interpretation Model Shortcuts

Lauren H. Cooke, Matthias Jung, Jan M. Brendel, Nora M. Kerkovits, Borek Foldyna, Michael T. Lu, Vineet K. Raghu

专题命中 知识编辑与模型理解 :foundation model(abstract);分类 cs.AI

Comments 25 + 8 pages, 4 + 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.03579 2025-09-09 cs.LG 57%

Hallucination Detection on a Budget: Efficient Bayesian Estimation of Semantic Entropy

Kamil Ciosek, Nicolò Felicioni, Sina Ghiassian

机构 * Spotify

专题命中 知识编辑与模型理解 :LLM(abstract);分类 cs.LG

Comments 24 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.17478 2025-09-09 cs.LG stat.ML 57%

Towards a General Time Series Forecasting Model with Unified Representation and Adaptive Transfer

Yihang Wang, Yuying Qiu, Peng Chen, Kai Zhao, Yang Shu, Zhongwen Rao, Lujia Pan, Bin Yang, Chenjuan Guo

机构 * East China Normal University, Shanghai, China(东华大学) Aalborg University, Aalborg, Denmark(奥尔堡大学) Huawei Noah's Ark Lab, Shenzhen, China(华为诺亚实验室)

专题命中 知识编辑与模型理解 :foundation model(abstract);分类 cs.LG

Comments Accepted by the Forty-second International Conference on Machine Learning (ICML2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.04998 2025-09-08 cs.LG q-bio.BM 57%

Directed Evolution of Proteins via Bayesian Optimization in Embedding Space

Matouš Soldát, Jiří Kléma

机构 * Department of Computer Science FEE, Czech Technical University in Prague Prague, Czech Republic(计算机科学系(FEE),捷克技术大学布拉格分校)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.LG

Comments 8 pages, 2 figures

Journal ref Proceedings of 2024 IEEE International Conference on Bioinformatics and Biomedicine (BIBM), Lisbon, Portugal, 2024, pp. 91-98

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.04735 2025-09-08 cs.CV cs.AI 57%

Enhancing Self-Driving Segmentation in Adverse Weather Conditions: A Dual Uncertainty-Aware Training Approach to SAM Optimization

Dharsan Ravindran, Kevin Wang, Zhuoyuan Cao, Saleh Abdelrahman, Jeffery Wu

机构 * Queen's University(女王大学)

专题命中 知识编辑与模型理解 :foundation model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.01626 2025-09-08 cs.CL cs.IR 57%

Modeling Sequential Sentence Relation to Improve Cross-lingual Dense Retrieval

Shunyu Zhang, Yaobo Liang, Ming Gong, Daxin Jiang, Nan Duan

机构 * Microsoft Research Asia(微软亚洲研究院)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

Comments Published at ICLR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.03961 2025-09-05 cs.CV cs.AI 57%

Multimodal Feature Fusion Network with Text Difference Enhancement for Remote Sensing Change Detection

Yijun Zhou, Yikui Zhai, Zilu Ying, Tingfeng Xian, Wenlve Zhou, Zhiheng Zhou, Xiaolin Tian, Xudong Jia, Hongsheng Zhang, C. L. Philip Chen

机构 * College of Electronics and Information Engineering, Wuyi University(威怡大学电子与信息工程学院) School of Electronic and Information Engineering and the Key Laboratory of Big Data and Intelligent Robot, Ministry of Education, South China University of Technology(电子与信息工程学院和大数据与智能机器人重点实验室,华南理工大学) State Key Laboratory of Lunar and Planetary Sciences, Macau University of Science and Technology(澳门大学地球和行星科学国家重点实验室) College of Engineering and Computer Science, California State University, Northridge(工程与计算机科学学院,加州大学北岭分校) Department of Geography, The University of Hong Kong(地理系,香港大学) Faculty of Computer Science and Engineering, S(计算机科学与工程学院,S)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.03863 2025-09-05 cs.AI 57%

Expedition & Expansion: Leveraging Semantic Representations for Goal-Directed Exploration in Continuous Cellular Automata

Sina Khajehabdollahi, Gautier Hamon, Marko Cvjetko, Pierre-Yves Oudeyer, Clément Moulin-Frier, Cédric Colas

机构 * Flowers AI & CogSci Lab, Inria, France(Flowers AI与认知科学实验室,Inria,法国) MIT, USA(麻省理工学院,美国)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.20757 2025-09-04 cs.CL 57%

GUARD: Glocal Uncertainty-Aware Robust Decoding for Effective and Efficient Open-Ended Text Generation

Yuanhao Ding, Esteban Garces Arias, Meimingwei Li, Julian Rodemann, Matthias Aßenmacher, Danlu Chen, Gaojuan Fan, Christian Heumann, Chongsheng Zhang

机构 * Henan University(河南大学) Department of Statistics, LMU Munich(慕尼黑大学统计系) Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心) CISPA Helmholtz Center for Information Security, Saarbrücken(萨尔布吕肯亥姆霍尔兹信息安全中心) University of California, San Diego(加州大学圣地亚哥分校)

专题命中 知识编辑与模型理解 :LLM(abstract);分类 cs.CL

Comments Accepted at Findings of the Association for Computational Linguistics: EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏