arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-09-16 至 2025-09-16 共收录 266 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 30 篇

2509.11944 2025-09-16 cs.AI 74%

Agentic Temporal Graph of Reasoning with Multimodal Language Models: A Potential AI Aid to Healthcare

Susanta Mitra

专题命中 领域大模型 :language model(title);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11431 2025-09-16 cs.AI cs.CL 73%

Securing AI Agents: Implementing Role-Based Access Control for Industrial Applications

Aadil Gani Ganie

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.10467 2025-09-16 cs.IR cs.AI cs.CL cs.CV cs.MM 73%

DSRAG: A Domain-Specific Retrieval Framework Based on Document-derived Multimodal Knowledge Graph

Mengzheng Yang, Yanfei Ren, David Osei Opoku, Ruochang Li, Peng Ren, Chunxiao Xing

机构 * School of Software, Henan University, Kaifeng 475004, China(河南大学软件学院) BNRist, DCST, RIIT, Tsinghua University, Beijing 100084, China(清华大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments 12 pages, 5 figures. Accepted to the 22nd International Conference on Web Information Systems and Applications (WISA 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02505 2025-09-16 cs.RO 71%

Would you let a humanoid play storytelling with your child? A usability study on LLM-powered narrative Human-Robot Interaction

Maria Lombardi, Carmela Calabrese, Davide Ghiglino, Caterina Foglino, Davide De Tommaso, Giulia Da Lisca, Lorenzo Natale, Agnieszka Wykowska

机构 * Humanoid Sensing and Perception, Italian Institute of Technology (IIT)(人形感知与感知,意大利技术研究院) Social Cognition in Human-Robot Interaction, IIT(人机交互中的社会认知,IIT)

专题命中 领域大模型 :LLM(title)

Journal ref 2025 IEEE/RSJ International Conference on Intelligent Robots and Systems, Hangzhou, China, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11803 2025-09-16 cs.CL 70%

From Fuzzy Speech to Medical Insight: Benchmarking LLMs on Noisy Patient Narratives

Eden Mama, Liel Sheri, Yehudit Aperstein, Alexander Apartsin

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL

Comments 6 pages, 1 figure

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11714 2025-09-16 eess.IV cs.LG 70%

EMeRALDS: Electronic Medical Record Driven Automated Lung Nodule Detection and Classification in Thoracic CT Images

Hafza Eman, Furqan Shaukat, Muhammad Hamza Zafar, Syed Muhammad Anwar

机构 * Faculty of Electrical and Electronics Engineering, University of Engineering(电气电子工程学院,工程大学) Department of Engineering Sciences, University of Agder(工程科学系,阿格德大学) Sheikh Zayed Institute for Pediatric Surgical Innovation, Children’s National Hospital(谢赫扎耶德小儿外科创新研究所,儿童医院) School of Medicine and Health Sciences, George Washington University(医学与健康科学学院,乔治华盛顿大学)

专题命中 领域大模型 :language model(abstract);pretraining(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11444 2025-09-16 cs.CL cs.SI 70%

CognitiveSky: Scalable Sentiment and Narrative Analysis for Decentralized Social Media

Gaurab Chhetri, Anandi Dutta, Subasish Das

机构 * Department of Computer Science(计算机科学系) Ingram School of Engineering(Ingram工程学院) Texas State University(德克萨斯州立大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL

Comments This is the author's preprint version of a paper accepted for presentation at HICSS 59 (Hawaii International Conference on System Sciences), 2026, Hawaii, USA. The final published version will appear in the official conference proceedings. Conference site: https://hicss.hawaii.edu/

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11330 2025-09-16 cs.AI 70%

Decoding Plastic Toxicity: An Intelligent Framework for Conflict-Aware Relational Metapath Extraction from Scientific Abstracts

Sudeshna Jana, Manjira Sinha, Tirthankar Dasgupta

机构 * TCS Research India(塔塔咨询印度研究)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

Comments 11 pages, 6 figures, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.09686 2025-09-16 cs.IR cs.AI 70%

GeoGPT-RAG Technical Report

Fei Huang, Fan Wu, Zeqing Zhang, Qihao Wang, Long Zhang, Grant Michael Boquet, Hongyang Chen

机构 * GeoGPT Team(GeoGPT团队) Zhejiang Lab(浙江实验室)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

Comments 19 pages, 10 figures, 10 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11878 2025-09-16 cs.CV 67%

Do It Yourself (DIY): Modifying Images for Poems in a Zero-Shot Setting Using Weighted Prompt Manipulation

Sofia Jamil, Kotla Sai Charan, Sriparna Saha, Koustava Goswami, K J Joseph

机构 * Department of Computer Science & Engineering, Indian Institute of Technology Patna(计算机科学与工程系,印度理工学院帕纳布分校) Adobe Research(Adobe研究)

专题命中 领域大模型 :large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.03562 2025-09-16 cs.LG cs.AI 62%

Kolb-Based Experiential Learning for Generalist Agents with Human-Level Kaggle Data Science Performance

Antoine Grosnit, Alexandre Maraval, Refinath S N, Zichao Zhao, James Doran, Giuseppe Paolo, Albert Thomas, Jonas Gonzalez, Abhineet Kumar, Khyati Khandelwal, Abdelhakim Benechehab, Hamza Cherkaoui, Youssef Attia El-Hili, Kun Shao, Jianye Hao, Jun Yao, Balázs Kégl, Haitham Bou-Ammar, Jun Wang

专题命中 领域大模型 :LLM(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.06336 2025-09-16 cs.CV cs.AI cs.CR 57%

Multi-View Slot Attention Using Paraphrased Texts for Face Anti-Spoofing

Jeongmin Yu, Susang Kim, Kisu Lee, Taekyoung Kwon, Won-Yong Shin, Ha Young Kim

机构 * Yonsei University(延世大学)

专题命中 领域大模型 :language model(abstract);分类 cs.AI

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.16207 2025-09-16 cs.HC cs.AI cs.CY 57%

A Human-Centered Approach to Identifying Promises, Risks, & Challenges of Text-to-Image Generative AI in Radiology

Katelyn Morrison, Arpit Mathur, Aidan Bradshaw, Tom Wartmann, Steven Lundi, Afrooz Zandifar, Weichang Dai, Kayhan Batmanghelich, Motahhare Eslami, Adam Perer

专题命中 领域大模型 :prompting(abstract);分类 cs.AI

Comments 10 pages of main content, Appendix attached after references, accepted to AAAI/ACM AIES 2025

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 知识编辑与模型理解 17 篇

2505.23657 2025-09-16 cs.CL cs.AI cs.LG 89%

Active Layer-Contrastive Decoding Reduces Hallucination in Large Language Model Generation

Hongxiang Zhang, Hao Chen, Muhao Chen, Tianyi Zhang

机构 * Purdue University(普渡大学) University of California, Davis(加州大学戴维斯分校)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI、cs.LG

Comments 19 pages, 3 figures, EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11952 2025-09-16 cs.CV 88%

CLAIRE: A Dual Encoder Network with RIFT Loss and Phi-3 Small Language Model Based Interpretability for Cross-Modality Synthetic Aperture Radar and Optical Land Cover Segmentation

Debopom Sutradhar, Arefin Ittesafun Abian, Mohaimenul Azam Khan Raiaan, Reem E. Mohamed, Sheikh Izzal Azid, Sami Azam

机构 * Department of Computer Science and Engineering, United International University(计算机科学与工程系,国际大学) Faculty of Science and Information Technology, Charles Darwin University(科学与信息技术学院,查尔斯·达尔文大学) School of Engineering and Energy , Murdoch University(工程与能源学院,默多尼大学) Faculty of Science and Technology, Charles Darwin University(科学与技术学院,查尔斯·达尔文大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);small language model(title,abstract)

Comments 23 pages, 6 figures, 10 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.09931 2025-09-16 cs.LG cs.AI 86%

Mechanistic Interpretability of LoRA-Adapted Language Models for Nuclear Reactor Safety Applications

Yoon Pyo Lee

专题命中 知识编辑与模型理解 :language model(title,abstract);LLM(abstract);large language model(abstract);分类 cs.AI、cs.LG

Comments Accepted for publication in Nuclear Technology. 24 pages, 2 tables, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.12065 2025-09-16 cs.CL 83%

Steering Language Models in Multi-Token Generation: A Case Study on Tense and Aspect

Alina Klerings, Jannik Brinkmann, Daniel Ruffinelli, Simone Ponzetto

机构 * University of Mannheim(曼海姆大学) Technical University Clausthal(克劳斯泰尔大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract);分类 cs.CL

Comments to be published in The 2025 Conference on Empirical Methods in Natural Language Processing

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11986 2025-09-16 cs.CV cs.CL 79%

Lost in Embeddings: Information Loss in Vision-Language Models

Wenyan Li, Raphael Tang, Chengzu Li, Caiqi Zhang, Ivan Vulić, Anders Søgaard

机构 * University of Copenhagen(哥本哈根大学) Microsoft(微软) University of Cambridge(剑桥大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11287 2025-09-16 cs.CV cs.CL 79%

Mitigating Hallucinations in Large Vision-Language Models by Self-Injecting Hallucinations

Yifan Lu, Ziqi Zhang, Chunfeng Yuan, Jun Gao, Congxuan Zhang, Xiaojuan Qi, Bing Li, Weiming Hu

机构 * Beijing Key Laboratory of Super Intelligent Security of Multi-Modal Information, CASIA(北京多模态信息超级智能安全重点实验室,中国科学院自动化所) State Key Laboratory of Multimodal Artificial Intelligence Systems, CASIA(多模态人工智能系统国家重点实验室,中国科学院自动化所) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) Hello Group(Hello集团) Nanchang Hangkong University(南昌航空大学) The University of Hong Kong(香港大学) School of Information Science and Technology, ShanghaiTech University(上海科技大学信息科学与技术学院)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

Comments emnlp 2025 accepted

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.07233 2025-09-16 eess.AS cs.CL 79%

Reducing Object Hallucination in Large Audio-Language Models via Audio-Aware Decoding

Tzu-wen Hsu, Ke-Han Lu, Cheng-Han Chiang, Hung-yi Lee

机构 * Purdue University(普渡大学) National Taiwan University(国立台湾大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11369 2025-09-16 cs.LG 77%

Decoding Musical Origins: Distinguishing Human and AI Composers

Cheng-Yang Tsai, Tzu-Wei Huang, Shao-Yu Wei, Guan-Wei Chen, Hung-Ying Chu, Yu-Cheng Lin

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.10875 2025-09-16 cs.AI cond-mat.soft 77%

Is the `Agent' Paradigm a Limiting Framework for Next-Generation Intelligent Systems?

Jesse Gardner, Vladimir A. Baulin

机构 * Active Inference Institute(主动推断研究所)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.04335 2025-09-16 cs.CL cs.AI cs.LG 75%

Hallucinated Span Detection with Multi-View Attention Features

Yuya Ogasa, Yuki Arase

机构 * Grad. Sch. of Information Science and Tech.(信息科学与技术研究生院) The University of Osaka(大阪大学) School of Computing(计算学部) Institute of Science(科学研究所) LY Corporation(LY公司)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.03106 2025-09-16 cs.CL cs.LG 73%

Monitoring Decoding: Mitigating Hallucination via Evaluating the Factuality of Partial Response during Generation

Yurui Chang, Bochuan Cao, Lu Lin

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

Comments Accepted to ACL 2025 (Findings)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11915 2025-09-16 cs.CL 70%

Uncertainty in Authorship: Why Perfect AI Detection Is Mathematically Impossible

Aadil Gani Ganie

机构 * UNIVERSITAT POLITECNICA DE VALENCIA(瓦伦西亚理工大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11569 2025-09-16 cs.CL 70%

D$^2$HScore: Reasoning-Aware Hallucination Detection via Semantic Breadth and Depth Analysis in LLMs

Yue Ding, Xiaofang Zhu, Tianze Xia, Junfei Wu, Xinlong Chen, Qiang Liu, Liang Wang

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

Comments under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.16146 2025-09-16 cs.CV cs.AI cs.CL cs.LG 67%

Steering LVLMs via Sparse Autoencoder for Hallucination Mitigation

Zhenglin Hua, Jinghan He, Zijun Yao, Tianxu Han, Haiyun Guo, Yuheng Jia, Junfeng Fang

机构 * School of Computer Science and Engineering, Southeast University(东南大学计算机科学与工程学院) Key Laboratory of New Generation Artificial Intelligence Technology and Its Interdisciplinary Applications (Southeast University)(东南大学新一代人工智能技术及其交叉应用关键实验室) Foundation Model Research Center, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所基础模型研究中心) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) Department of Computer Science and Technology, Tsinghua University(清华大学计算机科学与技术系) Wuhan University of Technology(武汉理工大学) National University of Singapore(新加坡国立大学)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI、cs.LG

Comments Accepted to Findings of EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.07082 2025-09-16 cs.CV cs.AI cs.LG 62%

On the Generalization of Representation Uncertainty in Earth Observation

Spyros Kondylatos, Nikolaos Ioannis Bountos, Dimitrios Michail, Xiao Xiang Zhu, Gustau Camps-Valls, Ioannis Papoutsis

机构 * National Observatory of Athens(雅典国家天文台) National Technical University of Athens(雅典技术大学) University of Valencia(瓦伦西亚大学) Harokopio University of Athens(雅典惠克罗波利斯大学) Technical University of Munich(慕尼黑技术大学) Munich Center for Machine Learning(慕尼黑机器学习中心) Archimedes/Athena RC(阿基米德/雅典RC)

专题命中 知识编辑与模型理解 :pretraining(abstract);分类 cs.AI、cs.LG

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.12039 2025-09-16 cs.CV 50%

RAM++: Robust Representation Learning via Adaptive Mask for All-in-One Image Restoration

Zilong Zhang, Chujie Qin, Chunle Guo, Yong Zhang, Chao Xue, Ming-Ming Cheng, Chongyi Li

机构 * VCIP, CS, Nankai University(VCIP、计算机科学系、南开大学) Chongqing Chang’an Wangjiang Industrial Group Co., Ltd(重庆长安王江工业集团有限公司) Tiandy Technologies(天鼎科技)

专题命中 知识编辑与模型理解 :pretraining(abstract)

Comments 18 pages, 22 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.22104 2025-09-16 eess.AS 50%

M2D-CLAP: Exploring General-purpose Audio-Language Representations Beyond CLAP

Daisuke Niizumi, Daiki Takeuchi, Masahiro Yasuda, Binh Thien Nguyen, Yasunori Ohishi, Noboru Harada

专题命中 知识编辑与模型理解 :LLM(abstract)

Comments Formerly M2D2, reverted to M2D-CLAP. 15 pages, 7 figures, 13 tables. Accepted by IEEE Access

详情

展开后加载摘要…

URL PDF HTML 收藏