arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-11-12 至 2025-11-12 共收录 211 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 14 篇

2508.10669 2025-11-12 cs.AI cs.IR 57%

STEP: Stepwise Curriculum Learning for Context-Knowledge Fusion in Conversational Recommendation

Zhenye Yang, Jinpeng Chen, Huan Li, Xiongnan Jin, Xuanyang Li, Junwei Zhang, Hongbo Gao, Kaimin Wei, Senzhang Wang

机构 * School of Artificial Intelligence Shenzhen University(人工智能学院深圳大学) Beijing University of Posts and Telecommunications(北京邮电大学) Jinan University(暨南大学) Central South University(中南大学)

专题命中 领域大模型 :language model(abstract);分类 cs.AI

Comments 10 pages; 4 figures; 6 tables; code available at https://github.com/Alex-bupt/STEP

Journal ref CIKM '2025: Proceedings of the 34th ACM International Conference on Information and Knowledge Management

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07966 2025-11-12 cs.CV 50%

Multi-Modal Assistance for Unsupervised Domain Adaptation on Point Cloud 3D Object Detection

Shenao Zhao, Pengpeng Liang, Zhoufan Yang

专题命中 领域大模型 :language model(abstract)

Comments Accepted to AAAI-26

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.00598 2025-11-12 cs.CV 50%

DGL-RSIS: Decoupling Global Spatial Context and Local Class Semantics for Training-Free Remote Sensing Image Segmentation

Boyi Li, Ce Zhang, Richard M. Timmerman, Wenxuan Bao

专题命中 领域大模型 :language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 知识编辑与模型理解 13 篇

2511.07694 2025-11-12 cs.LG 89%

Probabilities Are All You Need: A Probability-Only Approach to Uncertainty Estimation in Large Language Models

Manh Nguyen, Sunil Gupta, Hung Le

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07295 2025-11-12 cs.IR cs.AI 89%

Hard vs. Noise: Resolving Hard-Noisy Sample Confusion in Recommender Systems via Large Language Models

Tianrui Song, Wen-Shuo Chao, Hao Liu

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

Comments Accepted by AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.02431 2025-11-12 cs.CL 89%

From Anger to Joy: How Nationality Personas Shape Emotion Attribution in Large Language Models

Mahammed Kamruzzaman, Abdullah Al Monsur, Gene Louis Kim, Anshuman Chhabra

机构 * University of South Florida(佛罗里达州立大学) North South University(北南大学) Bellini College of AI, Cybersecurity and Computing(人工智能、网络安全与计算学院)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL

Comments Accepted at AACL-2025 (main)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.16565 2025-11-12 cs.CL cs.AI cs.LG 89%

Language over Content: Tracing Cultural Understanding in Multilingual Large Language Models

Seungho Cho, Changgeon Ko, Eui Jun Hwang, Junmyeong Lee, Huije Lee, Jong C. Park

机构 * KAIST(韩国釜山科学技术院)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI、cs.LG

Comments Accepted to CIKM 2025 Workshop on Human Centric AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.11770 2025-11-12 cs.LG cs.AI cs.CL stat.ML 82%

Internal Causal Mechanisms Robustly Predict Language Model Out-of-Distribution Behaviors

Jing Huang, Junyi Tao, Thomas Icard, Diyi Yang, Christopher Potts

机构 * stanford(斯坦福大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI、cs.LG

Comments ICML 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08319 2025-11-12 cs.CL cs.AI cs.MA 79%

Adaptive Multi-Agent Response Refinement in Conversational Systems

Soyeong Jeong, Aparna Elangovan, Emine Yilmaz, Oleg Rokhlenko

机构 * KAIST(韩国科学技术院) Amazon(亚马逊) Collate University College London(伦敦大学学院)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments LaCATODA Workshop @ AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07941 2025-11-12 cs.CV cs.AI 77%

Libra-MIL: Multimodal Prototypes Stereoscopic Infused with Task-specific Language Priors for Few-shot Whole Slide Image Classification

Zhenfeng Zhuang, Fangyu Zhou, Liansheng Wang

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07480 2025-11-12 cs.CR cs.AI 77%

KG-DF: A Black-box Defense Framework against Jailbreak Attacks Based on Knowledge Graphs

Shuyuan Liu, Jiawei Chen, Xiao Yang, Hang Su, Zhaoxia Yin

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.15267 2025-11-12 cs.CL cs.AI 73%

TraceCoder: Towards Traceable ICD Coding via Multi-Source Knowledge Integration

Mucheng Ren, He Chen, Yuchen Yan, Danqing Hu, Jun Xu, Xian Zeng

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments Accpeted as BIBM 2025 Regular. 6 pages. Camera-Ready version

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08440 2025-11-12 cs.LG 70%

Coherence Mechanisms for Provable Self-Improvement

Mehryar Mohri, Jon Schneider, Yifan Wu

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.17120 2025-11-12 cs.CL 70%

Self-Interpretability: LLMs Can Describe Complex Internal Processes that Drive Their Decisions

Dillon Plunkett, Adam Morris, Keerthi Reddy, Jorge Morales

机构 * Northeastern University(东北大学) Princeton University(普林斯顿大学) Independent Researcher(独立研究者)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08075 2025-11-12 cs.CV cs.AI 57%

CLIP is All You Need for Human-like Semantic Representations in Stable Diffusion

Cameron Braunstein, Mariya Toneva, Eddy Ilg

机构 * Saarland University, Saarbrücken Germany(萨尔兰州大学) MPI for Software Systems, Saarbrücken Germany(软件系统研究所) University of Technology Nuremberg, Nuremberg Germany(纽伦堡技术大学)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI

Comments 28 pages, 8 figures, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07682 2025-11-12 cs.HC cs.AI 57%

Designing and Evaluating Malinowski's Lens: An AI-Native Educational Game for Ethnographic Learning

Michael Hoffmann, Jophin John, Jan Fillies, Adrian Paschke

机构 * Leibniz Supercomputing Centre(莱比锡超算中心) Stanford University(斯坦福大学)

专题命中 知识编辑与模型理解 :prompting(abstract);分类 cs.AI

Comments 21 pages, 8 figures. Full preprint version; shorter version in preparation

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 其他LLM 15 篇

2505.17510 2025-11-12 cs.CL 89%

Large Language Models Do Multi-Label Classification Differently

Marcus Ma, Georgios Chochlakis, Niyantha Maruthu Pandiyan, Jesse Thomason, Shrikanth Narayanan

机构 * University of Southern California(南加州大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL

Comments To be published in the Main Conference Proceedings of EMNLP 2025, 24 pages, 16 figures, 7 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.20797 2025-11-12 cs.CL cs.CY cs.SI 89%

"Whose Side Are You On?" Estimating Ideology of Political and News Content Using Large Language Models and Few-shot Demonstration Selection

Muhammad Haroon, Magdalena Wojcieszak, Anshuman Chhabra

机构 * University of California, Davis(加州大学戴维斯分校) University of South Florida(佛罗里达州立大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08143 2025-11-12 cs.CL cs.AI 86%

Relation as a Prior: A Novel Paradigm for LLM-based Document-level Relation Extraction

Qiankun Pi, Yepeng Sun, Jicang Lu, Qinlong Fan, Ningbo Huang, Shiyu Wang

机构 * Information Engineering University(信息工程大学) Academy of Military Science(军事科学院)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments 17 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.17222 2025-11-12 cs.CL 85%

Humans Hallucinate Too: Language Models Identify and Correct Subjective Annotation Errors With Label-in-a-Haystack Prompts

Georgios Chochlakis, Peter Wu, Arjun Bedi, Marcus Ma, Kristina Lerman, Shrikanth Narayanan

机构 * University of Southern California(南加州大学)

专题命中 其他LLM :language model(title,abstract);LLM(abstract);large language model(abstract);分类 cs.CL

Comments Accepted to the Main Proceedings of EMNLP, 2025. 20 pages, 16 figures, 10 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.11773 2025-11-12 cs.CV cs.HC 85%

AgentSense: Virtual Sensor Data Generation Using LLM Agents in Simulated Home Environments

Zikang Leng, Megha Thukral, Yaqi Liu, Hrudhai Rajasekhar, Shruthi K. Hiremath, Jiaman He, Thomas Plötz

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract)

Comments Accepted by AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08535 2025-11-12 cs.CV cs.AI 83%

Large Sign Language Models: Toward 3D American Sign Language Translation

Sen Zhang, Xiaoxiao He, Di Liu, Zhaoyang Xia, Mingyu Zhao, Chaowei Tan, Vivian Li, Bo Liu, Dimitris N. Metaxas, Mubbasir Kapadia

机构 * Rutgers University(罗格斯大学) Meta Reality Labs(Meta现实实验室) Qualcomm(高通公司) Walmart Global Tech(沃尔玛全球技术) Roblox PRISMS

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.18108 2025-11-12 cs.CV 82%

Unveiling Visual Perception in Language Models: An Attention Head Analysis Approach

Jing Bi, Junjia Guo, Yunlong Tang, Lianggong Bruce Wen, Zhang Liu, Chenliang Xu

机构 * University of Rochester(罗切斯特大学) Corning Inc(康宁公司)

专题命中 其他LLM :language model(title,abstract);large language model(abstract)

Journal ref CVPR 2025 (IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08389 2025-11-12 eess.AS cs.AI cs.CL 81%

Unifying Model and Layer Fusion for Speech Foundation Models

Yi-Jen Shih, David Harwath

专题命中 其他LLM :foundation model(title,abstract);分类 cs.CL、cs.AI

Comments Accepted by IEEE ASRU 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08282 2025-11-12 cs.NI cs.CR cs.ET 78%

SRE-Llama -- Fine-Tuned Meta's Llama LLM, Federated Learning, Blockchain and NFT Enabled Site Reliability Engineering(SRE) Platform for Communication and Networking Software Services

Eranga Bandara, Safdar H. Bouk, Sachin Shetty, Ravi Mukkamala, Abdul Rahman, Peter Foytik, Ross Gore, Xueping Liang, Ng Wee Keong, Kasun De Zoysa

专题命中 其他LLM :LLM(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.11154 2025-11-12 cs.CR cs.CL 77%

MPMA: Preference Manipulation Attack Against Model Context Protocol

Zihan Wang, Rui Zhang, Yu Liu, Wenshu Fan, Wenbo Jiang, Qingchuan Zhao, Hongwei Li, Guowen Xu

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

Comments This is an extended version of the copyrighted publication at AAAI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08177 2025-11-12 cs.HC cs.SE 71%

GazeCopilot: Evaluating Novel Gaze-Informed Prompting for AI-Supported Code Comprehension and Readability

Yasmine Elfares, Gül Çalikli, Mohamed Khamis

专题命中 其他LLM :prompting(title)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.07810 2025-11-12 cs.CL 70%

Understanding and Controlling Repetition Neurons and Induction Heads in In-Context Learning

Nhi Hoai Doan, Tatsuya Hiraoka, Kentaro Inui

机构 * Mohamed bin Zayed University of Artificial Intelligence(Mohamed bin Zayed人工智能大学) RIKEN(日本理化学研究所)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08232 2025-11-12 cs.SE 67%

OWLAPY: A Pythonic Framework for OWL Ontology Engineering

Alkid Baci, Luke Friedrichs, Caglar Demir, Axel-Cyrille Ngonga Ngomo

专题命中 其他LLM :large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08139 2025-11-12 cs.CL 57%

On the Interplay between Positional Encodings, Morphological Complexity, and Word Order Flexibility

Kushal Tatariya, Wessel Poelman, Miryam de Lhoneux

机构 * NLP, Department of Computer Science, KU Leuven(自然语言处理,计算机科学系,鲁汶大学)

专题命中 其他LLM :language model(abstract);分类 cs.CL

Comments IJCNLP-AACL: Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏