arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-09-18 至 2025-09-18 共收录 154 信号源:cs.CL, cs.AI, cs.LG

1. 效率与部署 25 篇

2506.02824 2025-09-18 cs.RO 50%

Efficient Tactile Perception with Soft Electrical Impedance Tomography and Pre-trained Transformer

Huazhi Dong, Ronald B. Liu, Sihao Teng, Delin Hu, Peisan, E, Francesco Giorgio-Serchi, Yunjie Yang

机构 * SMART Group, Institute for Imaging, Data and Communications, School of Engineering, The University of Edinburgh(SMART集团,成像与数据通信研究所,工程学院,爱丁堡大学)

专题命中 效率与部署 :pretraining(abstract)

Comments IEEE Transactions on Industrial Electronics

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.21233 2025-09-18 cs.CV 50%

CROP: Contextual Region-Oriented Visual Token Pruning

Jiawei Guo, Feifei Zhai, Pu Jian, Qianrun Wei, Yu Zhou

机构 * State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, CAS, Beijing, China(多模态人工智能系统国家重点实验室,自动化研究所,中国科学院,北京,中国) School of Artificial Intelligence, University of Chinese Academy of Sciences, Beijing, China(人工智能学院,中国科学院大学,北京,中国) Fanyu AI Laboratory, Zhongke Fanyu Technology Co., Ltd, Beijing, China(凡语AI实验室,中科创始人技术有限公司,北京,中国)

专题命中 效率与部署 :LLM(abstract)

Comments EMNLP2025 Main

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 领域大模型 12 篇

2505.13259 2025-09-18 cs.CL 89%

From Automation to Autonomy: A Survey on Large Language Models in Scientific Discovery

Tianshi Zheng, Zheye Deng, Hong Ting Tsang, Weiqi Wang, Jiaxin Bai, Zihao Wang, Yangqiu Song

机构 * Department of Computer Science and Engineering, HKUST, Hong Kong SAR, China(计算机科学与工程系,香港科技大学,香港特别行政区,中国)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL

Comments EMNLP 2025 Main

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13691 2025-09-18 cs.RO 85%

SPAR: Scalable LLM-based PDDL Domain Generation for Aerial Robotics

Songhao Huang, Yuwei Wu, Guangyao Shi, Gaurav S. Sukhatme, Vijay Kumar

机构 * GRASP Lab, University of Pennsylvania(宾夕法尼亚大学GRASP实验室) Department of Computer Science, University of Southern California(南加州大学计算机科学系)

专题命中 领域大模型 :LLM(title);large language model(abstract);language model(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13696 2025-09-18 cs.CL 83%

Integrating Text and Time-Series into (Large) Language Models to Predict Medical Outcomes

Iyadh Ben Cheikh Larbi, Ajay Madhavan Ravichandran, Aljoscha Burchardt, Roland Roller

机构 * German Research Center for Artificial Intelligence (DFKI)(德国人工智能研究中心) Technical University Berlin(柏林技术大学)

专题命中 领域大模型 :language model(title,abstract);large language model(abstract);分类 cs.CL

Comments Presented and published at BioCreative IX

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13785 2025-09-18 eess.AS cs.SD 82%

Summary on The Multilingual Conversational Speech Language Model Challenge: Datasets, Tasks, Baselines, and Methods

Bingshen Mu, Pengcheng Guo, Zhaokai Sun, Shuai Wang, Hexin Liu, Mingchen Shao, Lei Xie, Eng Siong Chng, Longshuai Xiao, Qiangze Feng, Daliang Wang

机构 * School of Intelligence Science and Technology, Nanjing University(智能科学与技术学院,南京大学) College of Computing and Data Science, Nanyang Technological University(计算与数据科学学院,南洋理工大学) Huawei Technologies, China(华为技术有限公司) Nexdata Technology Inc., USA(Nexdata技术公司)

专题命中 领域大模型 :language model(title,abstract);SLM(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13899 2025-09-18 cs.HC 75%

AI as a teaching tool and learning partner

Steven Watterson, Sarah Atkinson, Elaine Murray, Andrew McDowell

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract)

Comments 6 Pages, 1 Figure

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13731 2025-09-18 cs.RO 75%

Reinforcement Learning for Robotic Insertion of Flexible Cables in Industrial Settings

Jeongwoo Park, Seabin Lee, Changmin Park, Wonjong Lee, Changjoo Nam

机构 * Dept. of Electronic Engineering, Sogang University(电子工程系,成均馆大学) Dept. of Artificial Intelligence, Sogang University(人工智能系,成均馆大学)

专题命中 领域大模型 :language model(abstract);foundation model(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13846 2025-09-18 cs.CV cs.LG 74%

Consistent View Alignment Improves Foundation Models for 3D Medical Image Segmentation

Puru Vaish, Felix Meister, Tobias Heimann, Christoph Brune, Jelmer M. Wolterink

机构 * Department of Applied Mathematics, Technical Medical Centre, University of Twente(代尔夫特理工大学应用数学系) Digital Technology and Innovation, Siemens Healthineers, Erlangen, Germany(西门子医疗创新部,埃尔朗根,德国)

专题命中 领域大模型 :foundation model(title);分类 cs.LG

Comments MICCAI 2025: 1st Place in Transformer track and 2nd Place in Convolution track of SSL3D-OpenMind challenge

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13888 2025-09-18 cs.CL cs.AI cs.IR 73%

Combating Biomedical Misinformation through Multi-modal Claim Detection and Evidence-based Verification

Mariano Barone, Antonio Romano, Giuseppe Riccio, Marco Postiglione, Vincenzo Moscato

机构 * University of Naples Federico II(那不勒斯费迪里奇二世大学) Northwestern University(西北大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Journal ref SIGIR '25: Proceedings of the 48th International ACM SIGIR Conference on Research and Development in Information Retrieval, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13773 2025-09-18 cs.AI cs.IR 70%

MIRA: Empowering One-Touch AI Services on Smartphones with MLLM-based Instruction Recommendation

Zhipeng Bian, Jieming Zhu, Xuyang Xie, Quanyu Dai, Zhou Zhao, Zhenhua Dong

机构 * Shenzhen University(深圳大学) Huawei Noah’s Ark Lab(华为诺亚实验室) Zhejiang University(浙江大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

Comments Published in Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 6: Industry Track), ACL 2025. Official version: https://doi.org/10.18653/v1/2025.acl-industry.103

Journal ref Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 6: Industry Track) ACL 2025 1457-1465

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13730 2025-09-18 cs.CY 67%

Perspectives and potential issues in using artificial intelligence for computer science education

Juho Vepsäläinen, Petri Juntunen

专题命中 领域大模型 :large language model(abstract);language model(abstract)

Comments 12 pages, 1 figure, 2 tables, preprint (not approved for publication yet)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.17602 2025-09-18 eess.IV 67%

Attention-ResUNet and EfficientSASM-UNet: UNet based frameworks for Lung and Nodule segmentation

Muhammad Abdullah, Furqan Shaukat

专题命中 领域大模型 :language model(abstract);foundation model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13393 2025-09-18 cs.OH 50%

Vehicle-to-Grid Integration: Ensuring Grid Stability, Strengthening Cybersecurity, and Advancing Energy Market Dynamics

Bilal Ahmad, Jianguo Ding, Tayyab Ali, Doreen Sebastain Sarwatt, Ramsha Arshad, Adamu Gaston Philipo, Huansheng Ning

专题命中 领域大模型 :prompting(abstract)

Comments 65 pages, 10 figures, 6 tables

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 知识编辑与模型理解 5 篇

2509.13702 2025-09-18 cs.CL cs.AI 90%

DSCC-HS: A Dynamic Self-Reinforcing Framework for Hallucination Suppression in Large Language Models

Xiao Zheng

机构 * School of Computing and Technology(计算机学院) China University of Petroleum(中国石油大学) Qingdao(青岛)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13664 2025-09-18 cs.CL cs.AI 82%

Sparse Neurons Carry Strong Signals of Question Ambiguity in LLMs

Zhuoxuan Zhang, Jinhao Duan, Edward Kim, Kaidi Xu

机构 * Brown University(布朗大学) Drexel University(德雷塞尔大学)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);prompting(abstract)

Comments To be appeared in EMNLP 2025 (main)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.08380 2025-09-18 cs.AI cs.LG 73%

Co-Investigator AI: The Rise of Agentic AI for Smarter, Trustworthy AML Compliance Narratives

Prathamesh Vasudeo Naik, Naresh Kumar Dintakurthi, Zhanghao Hu, Yue Wang, Robby Qiu

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.16846 2025-09-18 eess.AS cs.CL cs.SD 57%

KALL-E:Autoregressive Speech Synthesis with Next-Distribution Prediction

Kangxiang Xia, Xinfa Zhu, Jixun Yao, Wenjie Tian, Wenhao Li, Lei Xie

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

Comments 6 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.19505 2025-09-18 cs.AI 57%

Caught in the Act: a mechanistic approach to detecting deception

Gerard Boxo, Ryan Socha, Daniel Yoo, Shivam Raval

机构 * Barcelona Institute of Science and Technology(巴塞罗那科学与技术研究所) NorthWest Arkansas Community College(西北阿肯色社区学院) Carnegie Mellon University(卡内基梅隆大学) Harvard University(哈佛大学)

专题命中 知识编辑与模型理解 :LLM(abstract);分类 cs.AI

Comments 15 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 其他LLM 15 篇

2411.06207 2025-09-18 cs.CL 89%

KBM: Delineating Knowledge Boundary for Adaptive Retrieval in Large Language Models

Zhen Zhang, Xinyu Wang, Yong Jiang, Zile Qiao, Zhuo Chen, Guangyu Li, Feiteng Mu, Mengting Hu, Pengjun Xie, Fei Huang

机构 * College of Software, Nankai University(南开大学软件学院) Tongyi Lab, Alibaba Group(阿里巴巴集团 Tongyi 实验室) School of Information Science and Technology, ShanghaiTech University(上海科技大学信息科学与技术学院)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13326 2025-09-18 cs.HC cs.LG 85%

LLM Chatbot-Creation Approaches

Hemil Mehta, Tanvi Raut, Kohav Yadav, Edward F. Gehringer

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.LG

Comments Forthcoming in Frontiers in Education (FIE 2025), Nashville, Tennessee, USA, Nov 2-5, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.04537 2025-09-18 cs.MA cs.AI cs.CY 85%

Emergent Social Dynamics of LLM Agents in the El Farol Bar Problem

Ryosuke Takata, Atsushi Masumori, Takashi Ikegami

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.12497 2025-09-18 cs.LG 79%

Prediction and Causality of functional MRI and synthetic signal using a Zero-Shot Time-Series Foundation Model

Alessandro Crimi, Andrea Brovelli

机构 * AGH University of Krakow, Poland(克拉科夫AGH大学) Institut de Neurosciences de la Timone UMR 7289, Aix Marseille Université, CNRS, 13005, Marseille, France(里莫内神经科学研究所)

专题命中 其他LLM :foundation model(title,abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13982 2025-09-18 cs.PL 78%

CLMTracing: Black-box User-level Watermarking for Code Language Model Tracing

Boyu Zhang, Ping He, Tianyu Du, Xuhong Zhang, Lei Yun, Kingsum Chow, Jianwei Yin

专题命中 其他LLM :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.14030 2025-09-18 cs.AI 77%

CrowdAgent: Multi-Agent Managed Multi-Source Annotation System

Maosheng Qin, Renyu Zhu, Mingxuan Xia, Chenkai Chen, Zhen Zhu, Minmin Lin, Junbo Zhao, Lu Xu, Changjie Fan, Runze Wu, Haobo Wang

机构 * Zhejiang University(浙江大学) NetEase Fuxi AI Lab(网易凤凰AI实验室) Zhejiang Sci-tech University(浙江科技学院)

专题命中 其他LLM :large language model(abstract);language model(abstract);small language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.23512 2025-09-18 cs.CL 77%

SCORE: Story Coherence and Retrieval Enhancement for AI Narratives

Qiang Yi, Yangfan He, Jianhui Wang, Xinyuan Song, ShiYao Qian, Xinhang Yuan, Yi Xin, Yijin Wang, Jingqun Tang, Yuchen Li, Junjiang Lin, Hongyang He, Zhen Tian, Tianxiang Xu, Keqin Li, Kuan Lu, Menghao Huo, Jiaqi Chen, Miao Zhang, Tianyu Shi, Jianyuan Ni

机构 * UCB(加州大学伯克利分校) UMN(明尼苏达大学) UESTC(电子科技大学) Emory(埃默里大学) UofT(多伦多大学) WUSTL(华盛顿大学) NJU(南京大学) XDU(西安电子科技大学) ByteDance(字节跳动) Baidu(百度) UofWarwick(沃里克大学) UofGlasgow(格拉斯哥大学) AMA(美国医学协会) Google(谷歌) SCU(四川大学) THU-SZ(清华大学深圳研究院) Amazon(亚马逊) PKU(北京大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.12824 2025-09-18 cs.IR 75%

DiffHash: Text-Guided Targeted Attack via Diffusion Models against Deep Hashing Image Retrieval

Zechao Liu, Zheng Zhou, Xiangkun Chen, Tao Liang, Dapeng Lang

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.00132 2025-09-18 cs.CL cs.LG 73%

Contextualize-then-Aggregate: Circuits for In-Context Learning in Gemma-2 2B

Aleksandra Bakalova, Yana Veitsman, Xinting Huang, Michael Hahn

机构 * Saarland Informatics Campus(萨尔兰大学信息技术校区)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13869 2025-09-18 cs.CL 70%

Do LLMs Align Human Values Regarding Social Biases? Judging and Explaining Social Biases with LLMs

Yang Liu, Chenhui Chu

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

Comments 38 pages, 31 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.10663 2025-09-18 cs.CL 70%

Context Copying Modulation: The Role of Entropy Neurons in Managing Parametric and Contextual Knowledge Conflicts

Zineddine Tighidet, Andrea Mogini, Hedi Ben-younes, Jiali Mei, Patrick Gallinari, Benjamin Piwowarski

机构 * BNP Paribas(BNP巴黎银行) Sorbonne Université(索邦大学) Criteo AI Lab(Criteo人工智能实验室)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

Comments Accepted at EMNLP 2025

Journal ref EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏