arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-09-10 至 2025-09-10 共收录 152 信号源:cs.CL, cs.AI, cs.LG

1. 效率与部署 28 篇

2506.18407 2025-09-10 cs.GR cs.CV 67%

IntuiTF: MLLM-Guided Transfer Function Optimization for Direct Volume Rendering

Yiyao Wang, Bo Pan, Ke Wang, Han Liu, Jinyuan Mao, Yuxin Liu, Minfeng Zhu, Xiuqi Huang, Weifeng Chen, Bo Zhang, Wei Chen

机构 * State Key Lab of CAD&CG, Zhejiang University(浙江大学CAD与CG国家重点实验室) Laboratory of Art and Archaeology Image (Zhejiang University), Ministry of Education, China(浙江大学艺术与考古图像实验室(教育部,中国)) Zhejiang University of Finance&Economics(浙江财经大学) Zhejiang University(浙江大学)

专题命中 效率与部署 :large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.07324 2025-09-10 cs.CL cs.AI 62%

Mitigating Attention Localization in Small Scale: Self-Attention Refinement via One-step Belief Propagation

Nakyung Lee, Yeongoon Kim, Minhae Oh, Suhwan Kim, Jin Woo Koo, Hyewon Jo, Jungwoo Lee

机构 * Seoul National University(首尔国立大学)

专题命中 效率与部署 :language model(abstract);分类 cs.CL、cs.AI

Comments Accepted at EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.07727 2025-09-10 cs.LG cs.DC 57%

MoE-Compression: How the Compression Error of Experts Affects the Inference Accuracy of MoE Model?

Songkai Ma, Zhaorui Zhang, Sheng Di, Benben Liu, Xiaodong Yu, Xiaoyi Lu, Dan Wang

机构 * Department of Computing(计算系) Hong Kong Polytechnic University(香港理工大学) Mathematics and Computer Science Division(数学与计算机科学 division) Argonne National Laboratory(阿贡国家实验室) The University of Hong Kong(香港大学) Stevens Institute of Technology(史蒂文斯理工学院) Department of Computer Science and Engineering(计算机科学与工程系) University of California, Merced(加州大学默塞德分校)

专题命中 效率与部署 :LLM(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.07391 2025-09-10 cs.CV cs.LG 57%

A Data-Free Analytical Quantization Scheme for Deep Learning Models

Ahmed Luqman, Khuzemah Qazi, Murray Patterson, Malik Jahan Khan, Imdadullah Khan

机构 * Georgia State University(佐治亚州立大学)

专题命中 效率与部署 :post-training(abstract);分类 cs.LG

Comments Accepted for publication in IEEE International Conference on Data Mining (ICDM 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.00969 2025-09-10 cs.CV 50%

Seeing More, Saying More: Lightweight Language Experts are Dynamic Video Token Compressors

Xiangchen Wang, Jinrui Zhang, Teng Wang, Haigang Zhang, Feng Zheng

机构 * Southern University of Science and Technology(南方科技大学) The University of Hong Kong(香港大学) Shenzhen Polytechnic University(深圳职业技术大学)

专题命中 效率与部署 :language model(abstract)

Comments 17 pages, 8 figures, EMNLP2025

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 领域大模型 12 篇

2509.07135 2025-09-10 cs.CL 88%

MedBench-IT: A Comprehensive Benchmark for Evaluating Large Language Models on Italian Medical Entrance Examinations

Ruggero Marino Lazzaroni, Alessandro Angioi, Michelangelo Puliga, Davide Sanna, Roberto Marras

机构 * University of Graz(格拉茨大学) OnePix Academy(OnePix学院)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL

Comments Accepted as an oral presentation at CLiC-it 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04431 2025-09-10 cs.AI cs.CL 88%

MedGellan: LLM-Generated Medical Guidance to Support Physicians

Debodeep Banerjee, Burcu Sayin, Stefano Teso, Andrea Passerini

机构 * University of Pisa(帕尔米斯大学) University of Trento(特伦托大学) DISI, University of Trento(特伦托大学DISI学院) CIMEC, university of Trento(特伦托大学CIMEC中心)

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.07512 2025-09-10 cs.CL cs.AI cs.IR 86%

ALLabel: Three-stage Active Learning for LLM-based Entity Recognition using Demonstration Retrieval

Zihan Chen, Lei Shi, Weize Wu, Qiji Zhou, Yue Zhang

机构 * Beihang University(北航大学) Westlake University(西湖大学)

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.07933 2025-09-10 cs.SE cs.AI 85%

Breaking Android with AI: A Deep Dive into LLM-Powered Exploitation

Wanni Vidulige Ishan Perera, Xing Liu, Fan liang, Junyi Zhang

机构 * Sam Houston State University(萨姆霍茨州立大学)

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.07622 2025-09-10 cs.CL 85%

MaLei at MultiClinSUM: Summarisation of Clinical Documents using Perspective-Aware Iterative Self-Prompting with LLMs

Libo Ren, Yee Man Ng, Lifeng Han

机构 * University of Manchester, UK(曼彻斯特大学) Modul University Vienna, Austria(维也纳Modul大学) Leiden Institute of Advanced Computer Science (LIACS), Leiden University, The Netherlands(莱顿先进计算机科学研究所(LIACS)、莱顿大学) Leiden University Medical Center, The Netherlands(莱顿大学医学中心)

专题命中 领域大模型 :prompting(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL

Comments system paper at CLEF 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.07860 2025-09-10 cs.IR 85%

KLIPA: A Knowledge Graph and LLM-Driven QA Framework for IP Analysis

Guanzhi Deng, Yi Xie, Yu-Keung Ng, Mingyang Liu, Peijun Zheng, Jie Liu, Dapeng Wu, Yinqiao Li, Linqi Song

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.06966 2025-09-10 eess.SP cs.AI cs.LG 82%

Cross-device Zero-shot Label Transfer via Alignment of Time Series Foundation Model Embeddings

Neal G. Ravindra, Arijit Sehanobish

机构 * Independent Researcher(独立研究者)

专题命中 领域大模型 :foundation model(title,abstract);分类 cs.AI、cs.LG

Comments 5 pages, 3 figures, 1 table. tl;dr: Adversarial alignment of Time-Series Foundation Model (TSFM) embeddings enables transfer of high-quality clinical labels from medical-grade to consumer-grade wearables, enabling zero-shot prediction of gestational age without requiring paired data

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.07588 2025-09-10 cs.CL cs.AI 81%

BALI: Enhancing Biomedical Language Representations through Knowledge Graph and Language Model Alignment

Andrey Sakhovskiy, Elena Tutubalina

机构 * AIRI Sber AI ISP RAS Research Center for Trusted AI(俄罗斯科学院信息与系统研究所可信人工智能研究中心)

专题命中 领域大模型 :language model(title,abstract);分类 cs.CL、cs.AI

Comments 9 pages, 1 figure, published in "The 48th International ACM SIGIR Conference on Research and Development in Information Retrieval (SIGIR 2025)"

Journal ref Proceedings of the 48th International ACM SIGIR Conference on Research and Development in Information Retrieval (2025). Association for Computing Machinery, 1152-1164

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07353 2025-09-10 cs.AI cs.CL cs.LG 80%

Benchmarking for Domain-Specific LLMs: A Case Study on Academia and Beyond

Rubing Chen, Jiaxin Wu, Jian Wang, Xulu Zhang, Wenqi Fan, Chenghua Lin, Xiao-Yong Wei, Qing Li

机构 * The Hong Kong Polytechnic University(香港理工大学) The University of Manchester(曼彻斯特大学)

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

Comments Accepted by EMNLP2025 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.19667 2025-09-10 cs.LG cs.AI 79%

Tripartite-GraphRAG via Plugin Ontologies

Michael Banf, Johannes Kuhn

机构 * Perelyn GmbH(佩雷林公司)

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.11892 2025-09-10 cs.HC 67%

RPKT: Learning What You Don't -- Know Recursive Prerequisite Knowledge Tracing in Conversational AI Tutors for Personalized Learning

Jinwen Tang, Qiming Guo, Zhicheng Tang, Yi Shang

专题命中 领域大模型 :large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.16304 2025-09-10 cs.CV 50%

SAMba-UNet: SAM2-Mamba UNet for Cardiac MRI in Medical Robotic Perception

Guohao Huo, Ruiting Dai, Ling Shao, Hao Tang

机构 * University of Electronic Science and Technology of China(电子科技大学) University of Chinese Academy of Sciences(中国科学院大学) School of Computer Science, Peking University(北京大学计算机学院)

专题命中 领域大模型 :foundation model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 知识编辑与模型理解 9 篇

2509.07925 2025-09-10 cs.CL cs.AI cs.LG 89%

GENUINE: Graph Enhanced Multi-level Uncertainty Estimation for Large Language Models

Tuo Wang, Adithya Kulkarni, Tyler Cody, Peter A. Beling, Yujun Yan, Dawei Zhou

机构 * Virginia Polytechnic Institute and State University(弗吉尼亚理工学院和州立大学) Ball State University(巴尔的摩州立大学) Dartmouth College(达特茅斯学院)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI、cs.LG

Comments Accepted by EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.07334 2025-09-10 cs.HC 80%

SpecifyUI: Supporting Iterative UI Design Intent Expression through Structured Specifications and Generative AI

Yunnong Chen, Chengwei Shi, Liuqing Chen

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);prompting(abstract)

Comments 27 pages, 12 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.00700 2025-09-10 cs.CV 80%

Prompt the Unseen: Evaluating Visual-Language Alignment Beyond Supervision

Raehyuk Jung, Seungjun Yu, Hyunjung Shim

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);prompting(abstract)

Comments Link to publicly available codes is added

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.07190 2025-09-10 cs.CL cs.HC 77%

Rule-Based Moral Principles for Explaining Uncertainty in Natural Language Generation

Zahra Atf, Peter R Lewis

机构 * Faculty of Business and Information Technology(商业与信息技术学院) Ontario Tech University(安大略技术大学)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

Comments This paper was accepted for presentation at the 35th IEEE International Conference on Collaborative Advances in Software and Computing. Conference website:https://conf.researchr.org/home/cascon-2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.02884 2025-09-10 cs.LG cs.AI 73%

Unlearning vs. Obfuscation: Are We Truly Removing Knowledge?

Guangzhi Sun, Potsawee Manakul, Xiao Zhan, Mark Gales

机构 * Department of Engineering, University of Cambridge(剑桥大学工程系) Department of Informatics, King’s College London(伦敦大学国王学院信息学院) SCB 10X, SCBX Group(SCB 10X,SCBX集团)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

Comments To Appear in EMNLP 2025 main conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.07322 2025-09-10 cs.CL cs.LG 73%

MEMIT-Merge: Addressing MEMIT's Key-Value Conflicts in Same-Subject Batch Editing for LLMs

Zilu Dong, Xiangqing Shen, Rui Xia

机构 * School of Computer Science and Engineering, Nanjing University of Science and Technology, China(计算机科学与工程学院,南京理工大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

Comments Accepted by ACL2025 findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.07617 2025-09-10 cs.AI 70%

Transferable Direct Prompt Injection via Activation-Guided MCMC Sampling

Minghui Li, Hao Zhang, Yechao Zhang, Wei Wan, Shengshan Hu, pei Xiaobing, Jing Wang

机构 * School of Software Engineering, Huazhong University of Science and Technology(华中科技大学软件学院) School of Cyber Science and Engineering, Huazhong University of Science and Technology(华中科技大学网络安全学院) College of Computing and Data Science, Nanyang Technological University(南洋理工大学计算与数据科学学院) Faculty of Data Science, City University of Macau(澳门城市大学数据科学学院)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI

Comments Accepted to EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.09297 2025-09-10 cs.LG 70%

When Do Neural Networks Learn World Models?

Tianren Zhang, Guanyu Chen, Feng Chen

机构 * Department of Automation, Tsinghua University, Beijing, China(自动化系,清华大学,北京,中国)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.LG

Comments ICML 2025; ICLR 2025 World Models Workshop (oral, outstanding paper award)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.07475 2025-09-10 cs.CL cs.AI 62%

HALT-RAG: A Task-Adaptable Framework for Hallucination Detection with Calibrated NLI Ensembles and Abstention

Saumya Goswami, Siddharth Kurra

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 其他LLM 6 篇

2509.07642 2025-09-10 cs.AI 89%

Getting In Contract with Large Language Models -- An Agency Theory Perspective On Large Language Model Alignment

Sascha Kaltenpoth, Oliver Müller

机构 * Paderborn University, Department of Business Administration and Economics(帕德博恩大学商业管理与经济学系)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

Comments Presented at the 19th International Conference on Wirtschaftsinformatik 2024, Würzburg, Germany https://aisel.aisnet.org/wi2024/91/

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.17674 2025-09-10 cs.CR cs.AI cs.LG 87%

Attacking LLMs and AI Agents: Advertisement Embedding Attacks Against Large Language Models

Qiming Guo, Jinwen Tang, Xingran Huang

机构 * Department of Computer Science Texas A\&M University–Corpus Christi Corpus Christi, TX, USA EECS Department University of Missouri Columbia, MO, USA Department of Computer Engineering University of California–Riverside Riverside, CA, USA

专题命中 其他LLM :large language model(title);language model(title);LLM(abstract);分类 cs.AI、cs.LG

Comments 6 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21929 2025-09-10 cs.CL cs.LG math.DS 87%

Local Normalization Distortion and the Thermodynamic Formalism of Decoding Strategies for Large Language Models

Tom Kempton, Stuart Burrell

机构 * Department of Mathematics, University of Manchester(曼彻斯特大学数学系) Innovation Lab, Featurespace(Featurespace创新实验室)

专题命中 其他LLM :language model(title,abstract);large language model(title);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.17065 2025-09-10 cs.MA cs.AI 77%

Creative Agents: Simulating the Systems Model of Creativity with Generative Agents

Naomi Imasato, Kazuki Miyazawa, Takayuki Nagai, Takato Horii

机构 * Graduate School of Engineering Science, Osaka University, Osaka, Japan(大阪大学工学研究科) Artificial Intelligence Exploration Research Center, The University of Electro-Communications, Tokyo, Japan(电通大学人工智能探索研究中心) International Research Center for Neurointelligence (WPI-IRCN), The University of Tokyo, Tokyo, Japan(东京大学神经智能国际研究中心)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏