arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-10-01 至 2025-10-01 共收录 274 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 24 篇

2509.26110 2025-10-01 cs.SE astro-ph.IM 67%

Agent-based code generation for the Gammapy framework

Dmitriy Kostunin, Vladimir Sotnikov, Sergo Golovachev, Abhay Mehta, Tim Lukas Holch, Elisa Jones

专题命中 领域大模型 :large language model(abstract);language model(abstract)

Comments ICRC2025 proceedings PoS(ICRC2025)753

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.26001 2025-10-01 cs.DL 67%

First Workshop on Building Innovative Research Systems for Digital Libraries (BIRDS 2025)

Christin Katharina Kreutz, Hermann Kroll

专题命中 领域大模型 :large language model(abstract);language model(abstract)

Comments Workshop accepted at and held @ TPDL'25 in Tampere, Finland. Webpage: https://ws-birds.github.io/birds2025.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25449 2025-10-01 cs.LG cs.AI 62%

Joint Embeddings Go Temporal

Sofiane Ennadir, Siavash Golkar, Leopoldo Sarra

机构 * KTH Stockholm(瑞典斯德哥尔摩皇家理工学院) New York University(纽约大学) Flatiron Institute(Flatiron研究所)

专题命中 领域大模型 :foundation model(abstract);分类 cs.AI、cs.LG

Comments Accepted at the Workshop on Time Series in the Age of Large Models - NeurIPS 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.16514 2025-10-01 cs.CL cs.AI 62%

The Ever-Evolving Science Exam

Junying Wang, Zicheng Zhang, Yijin Guo, Farong Wen, Ye Shen, Yingji Liang, Yalun Wu, Wenzhe Li, Chunyi Li, Zijian Chen, Qi Jia, Guangtao Zhai

机构 * Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)

专题命中 领域大模型 :foundation model(abstract);分类 cs.CL、cs.AI

Comments 33 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.20411 2025-10-01 cs.CR cs.AI 57%

Adversarial Defense in Cybersecurity: A Systematic Review of GANs for Threat Detection and Mitigation

Tharcisse Ndayipfukamiye, Jianguo Ding, Doreen Sebastian Sarwatt, Adamu Gaston Philipo, Huansheng Ning

专题命中 领域大模型 :LLM(abstract);分类 cs.AI

Comments 36 pages, 10 tables, 4figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25711 2025-10-01 cs.CV 50%

ProbMed: A Probabilistic Framework for Medical Multimodal Binding

Yuan Gao, Sangwook Kim, Jianzhong You, Chris McIntosh

机构 * Peter Munk Cardiac Centre(彼得·默克心脏中心) Ted Rogers Centre for Heart Research(泰德·罗杰斯心脏病研究中心) University Health Network(大学健康网络) Joint Department of Medical Imaging(联合医学影像部门) University of Toronto(多伦多大学) Vector Institute(向量研究所)

专题命中 领域大模型 :pretraining(abstract)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.26463 2025-10-01 cs.SE 50%

ErrorPrism: Reconstructing Error Propagation Paths in Cloud Service Systems

Junsong Pu, Yichen Li, Zhuangbin Chen, Jinyang Liu, Zhihan Jiang, Jianjun Chen, Rui Shi, Zibin Zheng, Tieying Zhang

专题命中 领域大模型 :LLM(abstract)

Comments 12 pages, 6 figures, 1 table, this paper has been accepted by the 40th IEEE/ACM International Conference on Automated Software Engineering, ASE 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25963 2025-10-01 cs.CV 50%

Self-Supervised Anatomical Consistency Learning for Vision-Grounded Medical Report Generation

Longzhen Yang, Zhangkai Ni, Ying Wen, Yihang Liu, Lianghua He, Heng Tao Shen

机构 * Tongji University(同济大学) East China Normal University(华东师范大学) Shanghai Eye Disease Prevention and Treatment Center(上海眼病防治中心)

专题命中 领域大模型 :foundation model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25499 2025-10-01 cs.HC 50%

Atlas of Human-AI Interaction (v1): An Interactive Meta-Science Platform for Large-Scale Research Literature Sensemaking

Chayapatr Archiwaranguprok, Awu Chen, Sheer Karny, Hiroshi Ishii, Pattie Maes, Pat Pataranutaporn

专题命中 领域大模型 :LLM(abstract)

Comments 37 pages, 6 figures, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 知识编辑与模型理解 15 篇

2509.21761 2025-10-01 cs.CR cs.AI 85%

Backdoor Attribution: Elucidating and Controlling Backdoor in Language Models

Miao Yu, Zhenhong Zhou, Moayad Aloqaily, Kun Wang, Biwei Huang, Stephen Wang, Yueming Jin, Qingsong Wen

机构 * University of Science and Technology of China(中国科学技术大学) Nanyang Technological University(南洋理工大学) United Arab Emirates University(阿拉伯联合酋长国大学) University of California San Diego(加州大学圣地亚哥分校) Abel AI National University of Singapore(新加坡国立大学) Squirrel Ai Learning

专题命中 知识编辑与模型理解 :language model(title,abstract);LLM(abstract);large language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25525 2025-10-01 cs.CR cs.LG 83%

Defeating Cerberus: Concept-Guided Privacy-Leakage Mitigation in Multimodal Language Models

Boyang Zhang, Istemi Ekin Akkus, Ruichuan Chen, Alice Dethise, Klaus Satzke, Ivica Rimac, Yang Zhang

机构 * CISPA Helmholtz Center for Information Security(CISPA赫尔姆霍茨信息安全中心) Nokia Bell Labs(诺基亚贝尔实验室)

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25568 2025-10-01 cs.CL cs.AI 81%

Probing the Limits of Stylistic Alignment in Vision-Language Models

Asma Farajidizaji, Akash Gupta, Vatsal Raina

机构 * Imperial College London(伦敦帝国学院) Apta AI

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI

Comments 5 pages, 1 figure, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25552 2025-10-01 cs.AI 79%

Evaluating Foundation Models with Pathological Concept Learning for Kidney Cancer

Shangqi Gao, Sihan Wang, Yibo Gao, Boming Wang, Xiahai Zhuang, Anne Warren, Grant Stewart, James Jones, Mireia Crispin-Ortuzar

机构 * University of Cambridge, Cambridge, UK(剑桥大学) Fudan University, Shanghai, China(复旦大学) Cambridge University Hospitals NHS Foundation Trust, Cambridge, UK(剑桥大学医院 NHS 基础信托)

专题命中 知识编辑与模型理解 :foundation model(title,abstract);分类 cs.AI

Comments Best Paper Award at MICCAI AMAI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25974 2025-10-01 cs.NI cs.MA 78%

OpenID Connect for Agents (OIDC-A) 1.0: A Standard Extension for LLM-Based Agent Identity and Authorization

Subramanya Nagabhushanaradhya

专题命中 知识编辑与模型理解 :LLM(title,abstract)

Comments 10 pages, 5 tables, 2 code listings. Specification proposal available at https://github.com/subramanya1997/oidc-a/

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.26116 2025-10-01 cs.LG cs.CE 77%

UncertainGen: Uncertainty-Aware Representations of DNA Sequences for Metagenomic Binning

Abdulkadir Celikkanat, Andres R. Masegosa, Mads Albertsen, Thomas D. Nielsen

机构 * Aalborg University(奥尔堡大学)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25669 2025-10-01 cs.AI 74%

GroundSight: Augmenting Vision-Language Models with Grounding Information and De-hallucination

Xinxi Chen, Tianyang Chen, Lijia Hong

机构 * Independent Researcher(独立研究者)

专题命中 知识编辑与模型理解 :language model(title);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25220 2025-10-01 cs.CL cs.LG 73%

Cyclic Ablation: Testing Concept Localization against Functional Regeneration in AI

Eduard Kapelko

机构 * Eduard Kapelko(独立研究者)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

Comments Code is available at: https://www.kaggle.com/code/kapedalex/cycleablationpublic/

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.26461 2025-10-01 cs.CL 70%

CreAgentive: An Agent Workflow Driven Multi-Category Creative Generation Engine

Yuyang Cheng, Linyue Cai, Changwei Peng, Yumiao Xu, Rongfang Bie, Yong Zhao

机构 * Sichuan University(四川大学) Beijing Normal University(北京师范大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04943 2025-10-01 cs.CV cs.CL 70%

ReLoop: "Seeing Twice and Thinking Backwards" via Closed-loop Training to Mitigate Hallucinations in Multimodal understanding

Jianjiang Yang, Yanshu li, Ziyan Huang

机构 * University of Bristol(布里斯托大学) Brown University(布朗大学) South China University of Technology(华南理工大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

Comments Accepted by conference EMNLP2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.07309 2025-10-01 cs.CL 70%

ConfRAG: Confidence-Guided Retrieval-Augmenting Generation

Yin Huang, Yifan Ethan Xu, Kai Sun, Vera Yan, Alicia Sun, Haidar Khan, Jimmy Nguyen, Jingxiang Chen, Mohammad Kachuee, Zhaojiang Lin, Yue Liu, Aaron Colak, Anuj Kumar, Wen-tau Yih, Xin Luna Dong

机构 * Meta Reality Labs(Meta现实实验室) FAIR at Meta(Meta的FAIR)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

Comments 10 pages main content, 7 pages appendix, 6 figures, 10 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.01872 2025-10-01 cs.CL 70%

Turning Logic Against Itself : Probing Model Defenses Through Contrastive Questions

Rachneet Sachdeva, Rima Hazra, Iryna Gurevych

机构 * Ubiquitous Knowledge Processing Lab (UKP Lab), Department of Computer Science and Hessian Center for AI (hessian.AI), Technical University of Darmstadt(德累斯顿技术大学计算机科学系、海斯堡人工智能中心(hessian.AI)、通用知识处理实验室(UKP Lab))

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

Comments Accepted at EMNLP 2025 (Main)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25863 2025-10-01 cs.CV 67%

MAPLE: Multi-scale Attribute-enhanced Prompt Learning for Few-shot Whole Slide Image Classification

Junjie Zhou, Wei Shao, Yagao Yue, Wei Mu, Peng Wan, Qi Zhu, Daoqiang Zhang

机构 * The College of Artificial Intelligence, Nanjing University of Aeronautics and Astronautics(南京航空航天大学人工智能学院) The Key Laboratory of Brain-Machine Intelligence Technology, Ministry of Education(教育部脑机智能技术重点实验室)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.16002 2025-10-01 cs.CL cs.AI 62%

Causal Interventions Reveal Shared Structure Across English Filler-Gap Constructions

Sasha Boguraev, Christopher Potts, Kyle Mahowald

机构 * The University of Texas at Austin(德克萨斯大学奥斯汀分校) Stanford University(斯坦福大学)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI

Comments 22 pages, 21 figures, 10 tables; EMNLP (Main) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11625 2025-10-01 cs.LG cs.AI cs.CR 62%

Inducing Uncertainty on Open-Weight Models for Test-Time Privacy in Image Recognition

Muhammad H. Ashiq, Peter Triantafillou, Hung Yun Tseng, Grigoris G. Chrysos

机构 * University of Wisconsin-Madison(威斯康星大学麦迪逊分校) University of Warwick(沃里克大学)

专题命中 知识编辑与模型理解 :pretraining(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 其他LLM 10 篇

2509.25767 2025-10-01 cs.AI 88%

Galton's Law of Mediocrity: Why Large Language Models Regress to the Mean and Fail at Creativity in Advertising

Matt Keon, Aabid Karim, Bhoomika Lohana, Abdul Karim, Thai Nguyen, Tara Hamilton, Ali Abbas

机构 * mv Research Lab(55mv研究实验室) Monash University(莫纳什大学) School of Engineering, Western Sydney University(西澳大学工程学院)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25593 2025-10-01 cs.AI cs.CL cs.HC cs.IR 86%

Causal Autoencoder-like Generation of Feedback Fuzzy Cognitive Maps with an LLM Agent

Akash Kumar Panda, Olaoluwa Adigun, Bart Kosko

机构 * Department of Electrical and Computer Engineering(电气与计算机工程系) University of Southern California(南加州大学) School of Computing and Information Sciences(计算与信息科学学院) Florida International University(佛罗里达国际大学)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments 8 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25973 2025-10-01 cs.AI 83%

Scalable and Robust LLM Unlearning by Correcting Responses with Retrieved Exclusions

Junbeom Kim, Kyuyoung Kim, Jihoon Tack, Dongha Lim, Jinwoo Shin

机构 * KAIST AI(韩国科学技术院人工智能研究所) Yonsei University(延世大学)

专题命中 其他LLM :LLM(title);language model(abstract);prompting(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.26643 2025-10-01 cs.CL cs.LG 81%

Convergence and Divergence of Language Models under Different Random Seeds

Finlay Fehlauer, Kyle Mahowald, Tiago Pimentel

机构 * ETH Zürich(苏黎世联邦理工学院) University of Texas at Austin(德克萨斯大学奥斯汀分校)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL、cs.LG

Comments Published at EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.26224 2025-10-01 cs.CL cs.AI 81%

Type-Less yet Type-Aware Inductive Link Prediction with Pretrained Language Models

Alessandro De Bellis, Salvatore Bufi, Giovanni Servedio, Vito Walter Anelli, Tommaso Di Noia, Eugenio Di Sciascio

专题命中 其他LLM :language model(title,abstract);分类 cs.CL、cs.AI

Comments Accepted and to appear in Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing (EMNLP 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25268 2025-10-01 cs.LG cs.AI physics.ao-ph 81%

A Weather Foundation Model for the Power Grid

Cristian Bodnar, Raphaël Rousseau-Rizzi, Nikhil Shankar, James Merleau, Stylianos Flampouris, Guillem Candille, Slavica Antic, François Miralles, Jayesh K. Gupta

机构 * Silurian AI Hydro-Québec(水电公司)

专题命中 其他LLM :foundation model(title,abstract);分类 cs.AI、cs.LG

Comments 31 pages, 22 figures

详情

展开后加载摘要…

URL PDF HTML 收藏