arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 7565 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 7565 篇

2510.21119 2025-10-27 stat.ME stat.ML 67%

Leveraging semantic similarity for experimentation with AI-generated treatments

Lei Shi, David Arbour, Raghavendra Addanki, Ritwik Sinha, Avi Feller

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

Comments 31 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13345 2025-10-27 cs.CY cs.AI cs.CL cs.HC cs.LG 67%

Beyond Accuracy: Rethinking Hallucination and Regulatory Response in Generative AI

Zihao Li, Weiwei Yi, Jiahong Chen

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17842 2025-10-22 cs.SE cs.HC 67%

Vibe Coding: Toward an AI-Native Paradigm for Semantic and Intent-Driven Programming

Vinay Bamil

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

Comments 10 pages, 1 figure, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12276 2025-10-20 cs.RO 67%

Spatial Forcing: Implicit Spatial Representation Alignment for Vision-language-action Model

Fuhao Li, Wenxuan Song, Han Zhao, Jingbo Wang, Pengxiang Ding, Donglin Wang, Long Zeng, Haoang Li

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) Tsinghua University(清华大学) Westlake University(西湖大学) Zhejiang University(浙江大学) South China University of Technology(华南理工大学)

专题命中 知识编辑与模型理解 :language model(abstract);foundation model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11646 2025-10-14 cs.SD 67%

BridgeCode: A Dual Speech Representation Paradigm for Autoregressive Zero-Shot Text-to-Speech Synthesis

Jingyuan Xing, Mingru Yang, Zhipeng Li, Xiaofen Xing, Xiangmin Xu

机构 * South China University of Technology, Guangzhou, China(华南理工大学) Foshan University, Foshan, China(佛山大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.10637 2025-10-14 cs.RO 67%

High-Fidelity Simulated Data Generation for Real-World Zero-Shot Robotic Manipulation Learning with Gaussian Splatting

Haoyu Zhao, Cheng Zeng, Linghao Zhuang, Yaxi Zhao, Shengke Xue, Hao Wang, Xingyue Zhao, Zhongyu Li, Kehan Li, Siteng Huang, Mingxiu Chen, Xin Li, Deli Zhao, Hua Zou

机构 * Wuhan University(武汉大学) DAMO Academy, Alibaba Group(达摩院,阿里巴巴集团) Hupan Lab(虎扑实验室) The Chinese University of Hong Kong(香港中文大学) Tsinghua University(清华大学) Huazhong University of Science and Technology(华中科技大学) Zhejiang University(浙江大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

Comments 13 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.11035 2025-10-14 cs.LG cs.AI cs.CL cs.CV 67%

Tversky Neural Networks: Psychologically Plausible Deep Learning with Differentiable Tversky Similarity

Moussa Koulako Bala Doumbouya, Dan Jurafsky, Christopher D. Manning

机构 * Department of Computer Science, 353 Jane Stanford Way(计算机科学系)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.18427 2025-10-08 q-fin.CP q-fin.RM 67%

Tracing Positional Bias in Financial Decision-Making: Mechanistic Insights from Qwen2.5

Fabrizio Dimino, Krati Saxena, Bhaskarjit Sarmah, Stefano Pasquali

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.16876 2025-10-07 cs.SE 67%

Revolutionizing Validation and Verification: Explainable Testing Methodologies for Intelligent Automotive Decision-Making Systems

Halit Eris, Stefan Wagner

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

Comments Preprint to be published at SE4ADS

Journal ref 2025 IEEE/ACM 1st International Workshop on Software Engineering for Autonomous Driving Systems (SE4ADS), Ottawa, ON, Canada, 2025, pp. 34-37

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.09747 2025-10-07 cs.NE 67%

BrainFLORA: Uncovering Brain Concept Representation via Multimodal Neural Embeddings

Dongyang Li, Haoyang Qin, Mingyang Wu, Chen Wei, Quanying Liu

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

Comments ACM MM 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.16357 2025-10-07 cs.CV 67%

Law of Vision Representation in MLLMs

Shijia Yang, Bohan Zhai, Quanzeng You, Jianbo Yuan, Hongxia Yang, Chenfeng Xu

机构 * Stanford University(斯坦福大学) UC Berkeley(加州大学伯克利分校) The Hong Kong Polytechnic University(香港理工大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

Comments The code is available at https://github.com/bronyayang/Law_of_Vision_Representation_in_MLLMs

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.03315 2025-10-07 cs.CL cs.AI cs.LG 67%

Decomposing Attention To Find Context-Sensitive Neurons

Alex Gibson

机构 * University of Cambridge(剑桥大学)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI、cs.LG

Comments 10 pages, 7 figures. Submitted to the Mechanistic Interpretability Workshop at NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25863 2025-10-01 cs.CV 67%

MAPLE: Multi-scale Attribute-enhanced Prompt Learning for Few-shot Whole Slide Image Classification

Junjie Zhou, Wei Shao, Yagao Yue, Wei Mu, Peng Wan, Qi Zhu, Daoqiang Zhang

机构 * The College of Artificial Intelligence, Nanjing University of Aeronautics and Astronautics(南京航空航天大学人工智能学院) The Key Laboratory of Brain-Machine Intelligence Technology, Ministry of Education(教育部脑机智能技术重点实验室)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25177 2025-09-30 cs.CV 67%

Mitigating Hallucination in Multimodal LLMs with Layer Contrastive Decoding

Bingkui Tong, Jiaer Xia, Kaiyang Zhou

机构 * Mohamed bin Zayed University of Artificial Intelligence(莫扎德·本·扎耶德人工智能大学) Hong Kong Baptist University(香港 Baptist大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.03271 2025-09-30 cs.HC 67%

Beyond Quantification: Navigating Uncertainty in Professional AI Systems

Sylvie Delacroix, Diana Robinson, Umang Bhatt, Jacopo Domenicucci, Jessica Montgomery, Gael Varoquaux, Carl Henrik Ek, Vincent Fortuin, Yulan He, Tom Diethe, Neill Campbell, Mennatallah El-Assady, Soren Hauberg, Ivana Dusparic, Neil Lawrence

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

Journal ref RSS Data Science (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.17550 2025-09-30 cs.CV 67%

T2VUnlearning: A Concept Erasing Method for Text-to-Video Diffusion Models

Xiaoyu Ye, Songjie Cheng, Yongtao Wang, Yajiao Xiong, Yishen Li

机构 * Peking University(北京大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.23457 2025-09-30 cs.CV 67%

No Concept Left Behind: Test-Time Optimization for Compositional Text-to-Image Generation

Mohammad Hossein Sameti, Amir M. Mansourian, Arash Marioriyad, Soheil Fadaee Oshyani, Mohammad Hossein Rohban, Mahdieh Soleymani Baghshah

机构 * Sharif University of Technology(沙里夫技术大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

Comments 8 pages, 8 figures, 1 table

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.17858 2025-09-30 cs.IR 67%

LexSemBridge: Fine-Grained Dense Representation Enhancement through Token-Aware Embedding Augmentation

Shaoxiong Zhan, Hai Lin, Hongming Tan, Xiaodong Cai, Hai-Tao Zheng, Xin Su, Zifei Shan, Ruitong Liu, Hong-Gee Kim

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

Comments 8 pages, 4 figures. Accepted to ECAI

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.14233 2025-09-30 cs.CL cs.AI cs.LG 67%

Mechanistic Fine-tuning for In-context Learning

Hakaze Cho, Peng Luo, Mariko Kato, Rin Kaenbyou, Naoya Inoue

机构 * Japan Advanced Institute of Science and Technology(日本科学技术先进研究院) Beijing Institute of Technology(北京理工大学) RIKEN(日本研究机构)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI、cs.LG

Comments 28 pages, 31 figures, 6 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.09129 2025-09-30 cs.LG cs.AI cs.CL 67%

NextLocLLM: Location Semantics Modeling and Coordinate-Based Next Location Prediction with LLMs

Shuai Liu, Ning Cao, Yile Chen, Yue Jiang, George Rosario Jagadeesh, Gao Cong

机构 * Nanyang Technological University(南洋理工大学)

专题命中 知识编辑与模型理解 :LLM(abstract);分类 cs.CL、cs.AI、cs.LG

Comments STIntelligence in CIKM 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21997 2025-09-29 cs.CV 67%

Exposing Hallucinations To Suppress Them: VLMs Representation Editing With Generative Anchors

Youxu Shi, Suorong Yang, Dong Liu

机构 * University of Science and Technology of China(中国科学技术大学) Nanjing University(南京大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.01986 2025-09-29 cs.CL cs.AI cs.LG 67%

Adaptively profiling models with task elicitation

Davis Brown, Prithvi Balehannina, Helen Jin, Shreya Havaldar, Hamed Hassani, Eric Wong

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI、cs.LG

Comments EMNLP 2025 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.16146 2025-09-16 cs.CV cs.AI cs.CL cs.LG 67%

Steering LVLMs via Sparse Autoencoder for Hallucination Mitigation

Zhenglin Hua, Jinghan He, Zijun Yao, Tianxu Han, Haiyun Guo, Yuheng Jia, Junfeng Fang

机构 * School of Computer Science and Engineering, Southeast University(东南大学计算机科学与工程学院) Key Laboratory of New Generation Artificial Intelligence Technology and Its Interdisciplinary Applications (Southeast University)(东南大学新一代人工智能技术及其交叉应用关键实验室) Foundation Model Research Center, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所基础模型研究中心) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) Department of Computer Science and Technology, Tsinghua University(清华大学计算机科学与技术系) Wuhan University of Technology(武汉理工大学) National University of Singapore(新加坡国立大学)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI、cs.LG

Comments Accepted to Findings of EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.06700 2025-09-09 cs.NI 67%

Sovereign AI for 6G: Towards the Future of AI-Native Networks

Swarna Bindu Chetty, David Grace, Simon Saunders, Paul Harris, Eirini Eleni Tsiropoulou, Tony Quek, Hamed Ahmadi

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.04801 2025-09-08 eess.SP 67%

KGRAG-SC: Knowledge Graph RAG-Assisted Semantic Communication

Dayu Fan, Rui Meng, Song Gao, Xiaodong Xu

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

Comments 7 pages,4 figures,conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.02879 2025-09-04 econ.TH 67%

Artificial or Human Intelligence?

Eric Gao

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

Comments 20 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.02388 2025-09-03 q-fin.ST 67%

Bridging Human Cognition and AI: A Framework for Explainable Decision-Making Systems

N. Jean, G. Le Pera

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

Comments We introduce a practical framework for explainable AI that combines modern interpretability methods with Malle's five-category model of human behavior explanation, illustrated through real-world case studies in credit risk and regulatory analysis

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.00371 2025-09-03 cs.CV 67%

Two Causes, Not One: Rethinking Omission and Fabrication Hallucinations in MLLMs

Guangzong Si, Hao Yin, Xianfei Li, Qing Ding, Wenlong Liao, Tao He, Pai Peng

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

Comments Preprint,Underreview

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.19209 2025-08-27 cs.CV 67%

OmniHuman-1.5: Instilling an Active Mind in Avatars via Cognitive Simulation

Jianwen Jiang, Weihong Zeng, Zerong Zheng, Jiaqi Yang, Chao Liang, Wang Liao, Han Liang, Yuan Zhang, Mingyuan Gao

机构 * Intelligent Creation Lab, ByteDance(字节跳动智能创作实验室)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

Comments Homepage: https://omnihuman-lab.github.io/v1_5/

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.19036 2025-08-27 cs.CY 67%

Of the People, By the Algorithm: How AI Transforms Democratic Representation

Yuval Rymon

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

Comments 13 pages

详情

展开后加载摘要…

URL PDF HTML 收藏