arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-11-13 至 2025-11-13 共收录 183 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 21 篇

2504.03789 2025-11-13 cs.CY cs.MA 82%

Steve: LLM Powered ChatBot for Career Progression

Naveen Mathews Renji, Balaji Rao, Carlo Lipizzi

专题命中 领域大模型 :LLM(title);large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.09411 2025-11-13 cs.CL 81%

GSAP-ERE: Fine-Grained Scholarly Entity and Relation Extraction Focused on Machine Learning

Wolfgang Otto, Lu Gan, Sharmila Upadhyaya, Saurav Karmakar, Stefan Dietze

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);prompting(abstract)

Comments Accepted at AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.09325 2025-11-13 cs.AI cs.CL 79%

Not Everything That Counts Can Be Counted: A Case for Safe Qualitative AI

Stine Beltoft, Lukas Galke

机构 * University of Southern Denmark(丹麦南丹麦大学)

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments Accepted at 3rd International Conference on Frontiers of Artificial Intelligence, Ethics, and Multidisciplinary Applications (FAIEMA 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.10776 2025-11-13 cs.LG 77%

Nearly Optimal Algorithms for Contextual Dueling Bandits from Adversarial Feedback

Qiwei Di, Jiafan He, Quanquan Gu

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.LG

Comments 33pages, 2 figures, 1 table, ICML2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.12503 2025-11-13 cs.CY 75%

Qualitative Research in an Era of AI: A Pragmatic Approach to Data Analysis, Workflow, and Computation

Corey M. Abramson, Tara Prendergast, Zhuofan Li, Daniel Dohan

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract)

Comments FORTHCOMING: Abramson, Corey M., Tara Prendergast, Zhuofan Li, Daniel Dohan. 2026 (forthcoming). "Qualitative Research in an Era of AI: A Pragmatic Approach to Data Analysis, Workflow, and Computation". Annual Review of Sociology. pre-print, methodology, workflow article

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.06227 2025-11-13 cs.CL 70%

LExT: Towards Evaluating Trustworthiness of Natural Language Explanations

Krithi Shailya, Shreya Rajpal, Gokul S Krishnan, Balaraman Ravindran

机构 * Centre for Responsible AI, IIT Madras(责任人工智能中心,IIT马德拉斯)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08593 2025-11-13 cs.CL 70%

Knowledge Graph Analysis of Legal Understanding and Violations in LLMs

Abha Jha, Abel Salinas, Fred Morstatter

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.00477 2025-11-13 cs.IR 67%

Read the Docs Before Rewriting: Equip Rewriter with Domain Knowledge via Continual Pre-training

Qi Wang, Yixuan Cao, Yifan Liu, Jiangtao Zhao, Ping Luo

专题命中 领域大模型 :large language model(abstract);language model(abstract)

Comments The paper is about to undergo significant revisions

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.09247 2025-11-13 cs.AI 57%

MedFuse: Multiplicative Embedding Fusion For Irregular Clinical Time Series

Yi-Hsien Hsieh, Ta-Jung Chien, Chun-Kai Huang, Shao-Hua Sun, Che Lin

机构 * Graduate Institute of Communication Engineering, National Taiwan University (NTU)(国立台湾大学通信工程研究所) Department of Electrical Engineering, NTU(国立台湾大学电子工程系) Center for Advanced Computing and Imaging in Biomedicine, NTU(国立台湾大学生物医学先进计算与影像中心) Smart Medicine and Health Informatics Program, NTU(国立台湾大学智慧医疗与健康资讯学程)

专题命中 领域大模型 :pretraining(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.09088 2025-11-13 cs.CR cs.AI 57%

Improving Sustainability of Adversarial Examples in Class-Incremental Learning

Taifeng Liu, Xinjing Liu, Liangqiu Dong, Yang Liu, Yilong Yang, Zhuo Ma

机构 * ∗ Corresponding author(通讯作者)

专题命中 领域大模型 :language model(abstract);分类 cs.AI

Comments This paper is accepted to AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.00452 2025-11-13 cs.IR cs.AI 57%

M^2VAE: Multi-Modal Multi-View Variational Autoencoder for Cold-start Item Recommendation

Chuan He, Yongchao Liu, Qiang Li, Wenliang Zhong, Chuntao Hong, Xinwei Yao

机构 * Ant Group(蚂蚁集团)

专题命中 领域大模型 :pretraining(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.15282 2025-11-13 cs.LG 57%

AutoG: Towards automatic graph construction from tabular data

Zhikai Chen, Han Xie, Jian Zhang, Xiang song, Jiliang Tang, Huzefa Rangwala, George Karypis

机构 * Michigan State University(密歇根州立大学) Amazon(亚马逊)

专题命中 领域大模型 :LLM(abstract);分类 cs.LG

Comments camera ready version, update meta info,accepted by ICLR 2025 https://openreview.net/forum?id=hovDbX4Gh6

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07983 2025-11-13 cs.CV 50%

ChexFract: From General to Specialized -- Enhancing Fracture Description Generation

Nikolay Nechaev, Evgeniia Przhezdzetskaia, Dmitry Umerenkov, Dmitry V. Dylov

机构 * Artificial Intelligence Research Institute (AIRI)(人工智能研究 institute (AIRI))

专题命中 领域大模型 :language model(abstract)

Comments 13 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23693 2025-11-13 physics.flu-dyn 50%

CFDagent: A Language-Guided, Zero-Shot Multi-Agent System for Complex Flow Simulation

Zhaoyue Xu, Long Wang, Chunyu Wang, Yixin Chen, Qingyong Luo, Hua-Dong Yao, Shizhao Wang, Guowei He

专题命中 领域大模型 :LLM(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.15503 2025-11-13 cs.CV 50%

Domain Adaptation from Generated Multi-Weather Images for Unsupervised Maritime Object Classification

Dan Song, Shumeng Huo, Wenhui Li, Lanjun Wang, Chao Xue, An-An Liu

机构 * The School of Electrical and Information Engineering, Tianjin University, China(天津大学电气与信息工程学院) Tiandy Technologies Co., Ltd, Tianjin, China(天津天盾科技有限公司)

专题命中 领域大模型 :language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 知识编辑与模型理解 11 篇

2508.02087 2025-11-13 cs.CL 88%

When Truth Is Overridden: Uncovering the Internal Origins of Sycophancy in Large Language Models

Keyu Wang, Jin Li, Shu Yang, Zhuoran Zhang, Di Wang

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.09453 2025-11-13 eess.SP 85%

LLM Enabled Beam Training for Pinching Antenna Systems (PASS)

Deqiao Gan, Xiaoxia Xu, Xiaohu Ge, Yuanwei Liu

专题命中 知识编辑与模型理解 :LLM(title,abstract);large language model(abstract);language model(abstract)

Comments submitted to IEEE journal

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.17047 2025-11-13 cs.CL 79%

How Linguistics Learned to Stop Worrying and Love the Language Models

Richard Futrell, Kyle Mahowald

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08978 2025-11-13 cs.MM cs.CV 78%

Spatio-Temporal Data Enhanced Vision-Language Model for Traffic Scene Understanding

Jingtian Ma, Jingyuan Wang, Wayne Xin Zhao, Guoping Liu, Xiang Wen

机构 * School of Computer Science and Engineering, and the MOE Engineering Research Center of Advanced Computer Application Technology, Beihang University(计算机科学与工程学院,以及教育部先进计算机应用技术工程研究中心,北京航空航天大学) School of Computer Science and Engineering, the School of Economics and Management, and the MIIT Key Laboratory of Data Intelligence and Management, Beihang University(计算机科学与工程学院,经济管理学院,以及工信部数据智能与管理重点实验室,北京航空航天大学) Gaoling School of Artificial Intelligence, Renmin University of China(中关村人工智能学院,中国人民大学) DiDi Global Inc.(滴滴出行公司)

专题命中 知识编辑与模型理解 :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.09458 2025-11-13 cs.HC 75%

Exploring The Interaction-Outcome Paradox: Seemingly Richer and More Self-Aware Interactions with LLMs May Not Yet Lead to Better Learning

Rahul R. Divekar, Sophia Guerra, Lisette Gonzalez, Natasha Boos

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.15538 2025-11-13 cs.LG cs.AI cs.CL 75%

Capturing Polysemanticity with PRISM: A Multi-Concept Feature Description Framework

Laura Kopf, Nils Feldhus, Kirill Bykov, Philine Lou Bommer, Anna Hedström, Marina M. -C. Höhne, Oliver Eberle

机构 * Technische Universität Berlin(柏林技术大学) BIFOLD UMI Lab(UMI实验室) Fraunhofer Heinrich-Hertz-Institute(弗劳恩霍夫海因里希-赫兹研究所) ETH AI Center(苏黎世联邦理工学院人工智能中心) Universität Potsdam(波茨坦大学) Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.21239 2025-11-13 cs.CL 70%

Semantic Volume: Quantifying and Detecting both External and Internal Uncertainty in LLMs

Xiaomin Li, Zhou Yu, Ziji Zhang, Yingying Zhuang, Swair Shah, Narayanan Sadagopan, Anurag Beniwal

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08641 2025-11-13 cs.CR cs.CY cs.MA 67%

QOC DAO -- Stepwise Development Towards an AI Driven Decentralized Autonomous Organization

Marc Jansen, Christophe Verdot

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.09018 2025-11-13 cs.CV cs.AI 57%

Causally-Grounded Dual-Path Attention Intervention for Object Hallucination Mitigation in LVLMs

Liu Yu, Zhonghao Chen, Ping Kuang, Zhikun Feng, Fan Zhou, Lan Wang, Gillian Dobbie

机构 * University of Auckland(奥克兰大学) China Scholarship Council(中国留学基金委)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI

Comments 9 pages, published to AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.08802 2025-11-13 cs.LG 57%

The Non-Linear Representation Dilemma: Is Causal Abstraction Enough for Mechanistic Interpretability?

Denis Sutter, Julian Minder, Thomas Hofmann, Tiago Pimentel

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.LG

Comments NeurIPS 2025 (Spotlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.10180 2025-11-13 cs.CV cs.LG 57%

CART: Compositional Auto-Regressive Transformer for Image Generation

Siddharth Roheda, Rohit Chowdhury, Aniruddha Bala, Rohan Jaiswal

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.LG

Comments figures compressed to meet arxiv size limit

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 其他LLM 7 篇

2411.16602 2025-11-13 cs.CV cs.GR 89%

Chat2SVG: Vector Graphics Generation with Large Language Models and Image Diffusion Models

Ronghuan Wu, Wanchao Su, Jing Liao

机构 * City University of Hong Kong(香港城市大学) Monash University(墨尔本大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract)

Comments Project Page: https://chat2svg.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.24197 2025-11-13 cs.AI 77%

Learning API Functionality from In-Context Demonstrations for Tool-based Agents

Bhrij Patel, Ashish Jagmohan, Aditya Vempaty

机构 * University of Maryland, College Park(马里兰大学学院公园分校) Emergence AI NYC(纽约Emergence AI)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

Comments 19 Pages, 14 Figures, 7 Tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08785 2025-11-13 econ.GN q-fin.EC 75%

Making Talk Cheap: Generative AI and Labor Market Signaling

Anais Galdin, Jesse Silbert

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08604 2025-11-13 econ.GN q-fin.EC 75%

Generative Agents and Expectations: Do LLMs Align with Heterogeneous Agent Models?

Filippo Gusella, Eugenio Vicario

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏