arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12193 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12193 篇

2508.03654 2025-08-06 cs.CL cs.CV 79%

Can Large Vision-Language Models Understand Multimodal Sarcasm?

Xinyu Wang, Yue Zhang, Liqiang Jing

机构 * The University of Texas at Dallas(德克萨斯大学达拉斯分校)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

Comments Accepted by CIKM 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.03524 2025-08-06 cs.CV cs.LG 79%

Semantic Mosaicing of Histo-Pathology Image Fragments using Visual Foundation Models

Stefan Brandstätter, Maximilian Köller, Philipp Seeböck, Alissa Blessing, Felicitas Oberndorfer, Svitlana Pochepnia, Helmut Prosch, Georg Langs

机构 * Computational Imaging Research Lab(计算成像研究实验室) Division of General and Pediatric Radiology(普通放射科与小儿放射科 division) Christian Doppler Laboratory for Machine Learning Driven Precision Imaging(Christian Doppler 机器学习驱动的精准成像实验室) Comprehensive Center for Artificial Intelligence in Medicine(医学人工智能综合中心) Department of Pathology(病理学系)

专题命中 其他LLM :foundation model(title,abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01380 2025-08-05 cs.CV cs.AI 79%

Effective Damage Data Generation by Fusing Imagery with Human Knowledge Using Vision-Language Models

Jie Wei, Erika Ardiles-Cruz, Aleksey Panasyuk, Erik Blasch

专题命中 其他LLM :language model(title,abstract);分类 cs.AI

Comments 6 pages, IEEE NAECON'25

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.06509 2025-08-04 cs.SE cs.AI 79%

Private GPTs for LLM-driven testing in software development and machine learning

Jakub Jagielski, Consuelo Rojas, Markus Abel

机构 * Ambrosys Gmbh(Ambrosys公司)

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI

Comments 5 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.02701 2025-07-22 cs.CL 79%

On Entity Identification in Language Models

Masaki Sakata, Benjamin Heinzerling, Sho Yokoi, Takumi Ito, Kentaro Inui

机构 * Tohoku University(东大大学) RIKEN(日本研究机构) NINJAL(日本国家情报机构) Langsmith Inc.(Langsmith公司) MBZUAI

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

Comments ACL 2025 Findings; 26 pages, 13 figures, 9 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.14799 2025-07-22 cs.CR cs.AI 79%

Manipulating LLM Web Agents with Indirect Prompt Injection Attack via HTML Accessibility Tree

Sam Johnson, Viet Pham, Thai Le

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI

Comments EMNLP 2025 System Demonstrations Submission

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.13761 2025-07-21 cs.CL 79%

Innocence in the Crossfire: Roles of Skip Connections in Jailbreaking Visual Language Models

Palash Nandi, Maithili Joshi, Tanmoy Chakraborty

机构 * Department of Electrical Engineering(电气工程系) Indian Institute of Technology Delhi(印度理工学院德里)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.12443 2025-07-17 cs.NI cs.AI cs.HC cs.PL 79%

LLM-Based Config Synthesis requires Disambiguation

Rajdeep Mondal, Nikolaj Bjorner, Todd Millstein, Alan Tang, George Varghese

机构 * University of California Los Angeles(加州大学洛杉矶分校) Microsoft Research(微软研究院) Microsoft(微软)

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.00043 2025-07-15 q-bio.BM cs.LG 79%

RiNALMo: General-Purpose RNA Language Models Can Generalize Well on Structure Prediction Tasks

Rafael Josip Penić, Tin Vlašić, Roland G. Huber, Yue Wan, Mile Šikić

专题命中 其他LLM :language model(title,abstract);分类 cs.LG

Comments 31 pages, 9 figures

Journal ref Nat. Commun. 16, 5671 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04509 2025-07-08 cs.CV cs.AI 79%

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization

Zhendong Xiao, Wu Wei, Shujie Ji, Shan Yang, Changhao Chen

机构 * School of Automation Science and Engineering, South China University of Technology(自动化科学与工程学院,华南理工大学) Thrust of Artificial Intelligence, The Hong Kong University of Science and Technology (Guangzhou)(人工智能方向,香港科学与技术大学(广州))

专题命中 其他LLM :language model(title,abstract);分类 cs.AI

Comments PRCV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.16789 2025-06-26 cs.CL 79%

Conversational User-AI Intervention: A Study on Prompt Rewriting for Improved LLM Response Generation

Rupak Sarkar, Bahareh Sarrafzadeh, Nirupama Chandrasekaran, Nagu Rangan, Philip Resnik, Longqi Yang, Sujay Kumar Jauhar

专题命中 其他LLM :LLM(title,abstract);分类 cs.CL

Comments 8 pages, ACL style

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.17073 2025-06-23 cs.CY cs.AI 79%

LLM-Based Bot Broadens the Range of Arguments in Online Discussions, Even When Transparently Disclosed as AI

Valeria Vuk, Cristina Sarasua, Fabrizio Gilardi

机构 * Department of Political Science University of Zurich(政治学系苏黎世大学) Department of Informatics University of Zurich(信息学系苏黎世大学)

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.15975 2025-06-23 cs.CR cs.CL 79%

Multi-use LLM Watermarking and the False Detection Problem

Zihao Fu, Chris Russell

机构 * Oxford Internet Institute(牛津互联网研究所) University of Oxford(牛津大学)

专题命中 其他LLM :LLM(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.04619 2025-06-18 cs.CL 79%

SynGraph: A Dynamic Graph-LLM Synthesis Framework for Sparse Streaming User Sentiment Modeling

Xin Zhang, Qiyu Wei, Yingjie Zhu, Linhai Zhang, Deyu Zhou, Sophia Ananiadou

机构 * The University of Manchester(曼彻斯特大学) Harbin Institute of Technology(哈尔滨工业大学) King’s College London(伦敦大学国王学院) Southeast University(东南大学)

专题命中 其他LLM :LLM(title,abstract);分类 cs.CL

Comments Accepted at ACL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.13028 2025-06-17 cs.OS cs.AI 79%

NaSh: Guardrails for an LLM-Powered Natural Language Shell

Bimal Raj Gyawali, Saikrishna Achalla, Konstantinos Kallas, Sam Kumar

机构 * University of California, Los Angeles(加州大学洛杉矶分校)

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI

Comments 7 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.12543 2025-06-17 cs.LG math.OC 79%

Is your batch size the problem? Revisiting the Adam-SGD gap in language modeling

Teodora Srećković, Jonas Geiping, Antonio Orvieto

机构 * Max Planck Institute for Intelligent Systems(马克斯·普朗克智能系统研究所) ELLIS Institute(ELLIS研究所) Tübingen AI Center(图宾根人工智能中心)

专题命中 其他LLM :language model(title,abstract);分类 cs.LG

Comments Short version accepted at the 2025 HiLD Workshop at ICML

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.05739 2025-06-09 cs.CR cs.AI 79%

To Protect the LLM Agent Against the Prompt Injection Attack with Polymorphic Prompt

Zhilong Wang, Neha Nagaraja, Lan Zhang, Hayretdin Bahsi, Pawan Patil, Peng Liu

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI

Comments To appear in the Industry Track of the 55th Annual IEEE/IFIP International Conference on Dependable Systems and Networks (DSN 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.00725 2025-06-03 cond-mat.mtrl-sci cs.LG 79%

A Foundation Model for Non-Destructive Defect Identification from Vibrational Spectra

Mouyang Cheng, Chu-Liang Fu, Bowen Yu, Eunbi Rha, Abhijatmedhi Chotrattanapituk, Douglas L Abernathy, Yongqiang Cheng, Mingda Li

机构 * Quantum Measurement Group, MIT(麻省理工学院量子测量组) Center for Computational Science and Engineering, MIT(麻省理工学院计算科学与工程中心) Department of Materials Science and Engineering, MIT(麻省理工学院材料科学与工程系) Department of Nuclear Science and Engineering, MIT(麻省理工学院核科学与工程系) Department of Physics, MIT(麻省理工学院物理系) Department of Electrical Engineering and Computer Science, MIT(麻省理工学院电气工程与计算机科学系) Neutron Scattering Division, Oak Ridge National Laboratory(橡树岭国家实验室中子散射部)

专题命中 其他LLM :foundation model(title,abstract);分类 cs.LG

Comments 14 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.16173 2025-06-03 cs.CL 79%

Mapping 1,000+ Language Models via the Log-Likelihood Vector

Momose Oyama, Hiroaki Yamagiwa, Yusuke Takase, Hidetoshi Shimodaira

机构 * Kyoto University(京都大学) RIKEN(理化学研究所)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

Comments ACL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.11940 2025-06-03 cs.CL 79%

The Impact of Token Granularity on the Predictive Power of Language Model Surprisal

Byung-Doh Oh, William Schuler

机构 * Center for Data Science(数据科学中心) New York University(纽约大学) Department of Linguistics(语言学系) The Ohio State University(俄亥俄州立大学)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

Comments ACL 2025; results with Natural Stories alignment issue corrected (commit 4700daa)

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.17820 2025-06-03 cs.CV cs.AI 79%

Don't Miss the Forest for the Trees: Attentional Vision Calibration for Large Vision Language Models

Sangmin Woo, Donguk Kim, Jaehyuk Jang, Yubin Choi, Changick Kim

机构 * KAIST(韩国科学技术院)

专题命中 其他LLM :language model(title,abstract);分类 cs.AI

Comments ACL 2025 Findings; Project: https://sangminwoo.github.io/AvisC/

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.10304 2025-06-03 cs.SE cs.LG 79%

LLM-Powered Test Case Generation for Detecting Bugs in Plausible Programs

Kaibo Liu, Zhenpeng Chen, Yiyang Liu, Jie M. Zhang, Mark Harman, Yudong Han, Yun Ma, Yihong Dong, Ge Li, Gang Huang

机构 * Peking University(北京大学) Nanyang Technological University(南洋理工大学) King’s College London(伦敦国王学院) University College London(伦敦大学学院) National Key Laboratory of Data Space Technology and System(数据空间技术与系统国家重点实验室)

专题命中 其他LLM :LLM(title,abstract);分类 cs.LG

Comments Accepted by the 63rd Annual Meeting of the Association for Computational Linguistics (ACL 2025) Main Track

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.01135 2025-05-28 cs.SD cs.IR cs.LG eess.AS 79%

Music Foundation Model as Generic Booster for Music Downstream Tasks

WeiHsiang Liao, Yuhta Takida, Yukara Ikemiya, Zhi Zhong, Chieh-Hsin Lai, Giorgio Fabbro, Kazuki Shimada, Keisuke Toyama, Kinwai Cheuk, Marco A. Martínez-Ramírez, Shusuke Takahashi, Stefan Uhlich, Taketo Akama, Woosung Choi, Yuichiro Koyama, Yuki Mitsufuji

机构 * SonyAI(索尼人工智能) Sony Group Corporation(索尼集团) Sony Europe B.V.(索尼欧洲B.V.) Sony CSL(索尼计算机科学实验室)

专题命中 其他LLM :foundation model(title,abstract);分类 cs.LG

Comments 41 pages with 14 figures

Journal ref Published in Transactions on Machine Learning Research (TMLR), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.10164 2025-05-27 cs.CL 79%

Analyzing FOMC Minutes: Accuracy and Constraints of Language Models

Wonseong Kim, Jan Frederic Spörer, Siegfried Handschuh

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

Comments 15pages, 4 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.10681 2025-05-19 cs.CY cs.AI cs.HC 79%

Towards an LLM-powered Social Digital Twinning Platform

Önder Gürcan, Vanja Falck, Markus G. Rousseau, Larissa L. Lima

机构 * Center for Modeling Social Systems(社会科学建模中心) NORCE Norwegian Research Center AS(挪威NORCE研究机构) Kristiansand, Norway(挪威克里斯蒂安桑)

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI

Comments 13 pages, 3 figures, 23rd International Conference on Practical applications of Agents and Multi-Agent Systems (PAAMS 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.10453 2025-05-16 cs.CV cs.AI 79%

Vision language models have difficulty recognizing virtual objects

Tyler Tran, Sangeet Khemlani, J. G. Trafton

机构 * US Naval Research Laboratory(美国海军研究实验室)

专题命中 其他LLM :language model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.08215 2025-05-14 cs.AI cs.SD eess.AS 79%

Unveiling the Best Practices for Applying Speech Foundation Models to Speech Intelligibility Prediction for Hearing-Impaired People

Haoshuai Zhou, Boxuan Cao, Changgeng Mo, Linkai Li, Shan Xiang Wang

专题命中 其他LLM :foundation model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.06004 2025-05-12 cs.CL 79%

Exploring the Feasibility of Multilingual Grammatical Error Correction with a Single LLM up to 9B parameters: A Comparative Study of 17 Models

Dawid Wisniewski, Antoni Solarski, Artur Nowakowski

机构 * Poznan University of Technology(波兹南技术大学) Adam Mickiewicz University(亚当·密茨凯维奇大学)

专题命中 其他LLM :LLM(title);language model(abstract);分类 cs.CL

Comments Accepted at MTSummit 2025 (The 20th Machine Translation Summit)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.05970 2025-05-12 cs.CL 79%

Towards Developmentally Plausible Rewards: Communicative Success as a Learning Signal for Interactive Language Models

Lennart Stöpler, Rufat Asadli, Mitja Nikolaus, Ryan Cotterell, Alex Warstadt

机构 * CerCo, CNRS(CerCo与法国国家科学研究中心) University of California San Diego(加州大学圣地亚哥分校)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.01900 2025-05-06 cs.CL 79%

CAMOUFLAGE: Exploiting Misinformation Detection Systems Through LLM-driven Adversarial Claim Transformation

Mazal Bethany, Nishant Vishwamitra, Cho-Yu Jason Chiang, Peyman Najafirad

机构 * University of Texas at San Antonio(德克萨斯大学圣安东尼奥分校)

专题命中 其他LLM :LLM(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏