arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-10-22 至 2025-10-22 共收录 233 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 17 篇

2412.09572 2025-10-22 cs.CL 79%

Rethinking LLM Uncertainty: A Multi-Agent Approach to Estimating Black-Box Model Uncertainty

Yu Feng, Phu Mon Htut, Zheng Qi, Wei Xiao, Manuel Mager, Nikolaos Pappas, Kishaloy Halder, Yang Li, Yassine Benajiba, Dan Roth

机构 * University of Pennsylvania(宾夕法尼亚大学) AWS AI Labs(AWS人工智能实验室) Johannes Gutenberg University of Mainz(美因茨约翰内斯·古滕贝格大学) Oracle AI(Oracle人工智能)

专题命中 知识编辑与模型理解 :LLM(title,abstract);分类 cs.CL

Comments EMNLP 2025 Findings

Journal ref EMNLP 2025 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17942 2025-10-22 cs.CY cs.AI 79%

Trust in foundation models and GenAI: A geographic perspective

Grant McKenzie, Krzysztof Janowicz, Carsten Kessler

机构 * McGill University, Canada(麦吉尔大学,加拿大) University of Vienna, Austria(维也纳大学,奥地利) Bochum University of Applied Sciences, Germany(波鸿应用科学大学,德国) Aalborg University Copenhagen, Denmark(奥胡斯大学哥本哈根分校,丹麦)

专题命中 知识编辑与模型理解 :foundation model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17833 2025-10-22 q-bio.NC cs.AI 79%

Brain-Language Model Alignment: Insights into the Platonic Hypothesis and Intermediate-Layer Advantage

Ángela López-Cardona, Sebastián Idesis, Mireia Masias-Bruns, Sergi Abadal, Ioannis Arapakis

机构 * Universitat Politècnica de Catalunya(加泰罗尼亚理工大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17918 2025-10-22 cs.CL cs.AI 79%

JT-Safe: Intrinsically Enhancing the Safety and Trustworthiness of LLMs

Junlan Feng, Fanyu Meng, Chong Long, Pengyu Cong, Duqing Wang, Yan Zheng, Yuyao Zhang, Xuanchang Gao, Ye Yuan, Yunfei Ma, Zhijie Ren, Fan Yang, Na Wu, Di Jin, Chao Deng

机构 * China Mobile Jiutian Research(中国移动九天研究所)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);post-training(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17910 2025-10-22 cs.CY cs.AI cs.CL 79%

Interpretability Framework for LLMs in Undergraduate Calculus

Sagnik Dakshit, Sushmita Sinha Roy

机构 * University of Texas at Tyler(德克萨斯理工大学) Florida Gulf Coast University(佛罗里达盖恩斯维尔大学)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17909 2025-10-22 cs.CL 74%

Atomic Literary Styling: Mechanistic Manipulation of Prose Generation in Neural Language Models

Tsogt-Ochir Enkhbayar

机构 * Mongol AI(蒙古AI)

专题命中 知识编辑与模型理解 :language model(title);分类 cs.CL

Comments 12 pages, 3 figures, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17842 2025-10-22 cs.SE cs.HC 67%

Vibe Coding: Toward an AI-Native Paradigm for Semantic and Intent-Driven Programming

Vinay Bamil

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

Comments 10 pages, 1 figure, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.11741 2025-10-22 cs.AI cs.CR 57%

MTRE: Multi-Token Reliability Estimation for Hallucination Detection in VLMs

Geigh Zollicoffer, Minh Vu, Manish Bhattarai

机构 * Los Alamos National Laboratory(洛斯阿拉莫斯国家实验室)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.04641 2025-10-22 cs.LG math.ST stat.ML stat.TH 57%

A Statistical Theory of Contrastive Pre-training and Multimodal Generative AI

Kazusato Oko, Licong Lin, Yuhang Cai, Song Mei

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.12539 2025-10-22 cs.AI cs.MA 57%

Counterfactual Effect Decomposition in Multi-Agent Sequential Decision Making

Stelios Triantafyllou, Aleksa Sukovic, Yasaman Zolfimoselo, Goran Radanovic

机构 * Max Planck Institute for Software Systems(马克斯·普朗克软件系统研究所)

专题命中 知识编辑与模型理解 :LLM(abstract);分类 cs.AI

Comments ICML 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.18703 2025-10-22 cs.CV 50%

Exploring a Unified Vision-Centric Contrastive Alternatives on Multi-Modal Web Documents

Yiqi Lin, Alex Jinpeng Wang, Linjie Li, Zhengyuan Yang, Mike Zheng Shou

机构 * Show Lab, National University of Singapore(新加坡国立大学展示实验室) Central South University(中南大学) Microsoft(微软公司)

专题命中 知识编辑与模型理解 :language model(abstract)

Comments Project page: this https://linyq17.github.io/VC2L/

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.18321 2025-10-22 cs.CV 50%

Beyond Single Models: Mitigating Multimodal Hallucinations via Adaptive Token Ensemble Decoding

Jinlin Li, Yuran Wang, Yifei Yuan, Xiao Zhou, Yingying Zhang, Xixian Yong, Yefeng Zheng, Xian Wu

机构 * Gaoling School of Artificial Intelligence, Renmin University of China(中国人民大学 Gallup 学院) Department of Electrical and Computer Engineering, McGill University(麦吉尔大学电气与计算机工程系) School of Statistics, Renmin University of China(中国人民大学统计学院) Tencent Jarvis Lab(腾讯 Jarvis 实验室) Medical Artificial Intelligence Lab, Westlake University(西湖大学医学人工智能实验室)

专题命中 知识编辑与模型理解 :language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 其他LLM 11 篇

2412.10582 2025-10-22 cs.CL 92%

WHAT-IF: Exploring Branching Narratives by Meta-Prompting Large Language Models

Runsheng "Anson" Huang, Lara J. Martin, Chris Callison-Burch

机构 * University of Pennsylvania(宾夕法尼亚大学) University of Maryland, Baltimore County(马里兰大学巴尔的摩县分校)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);prompting(title,abstract);LLM(abstract)

Comments Published in Wordplay: When Language Meets Games Workshop (EMNLP 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.00481 2025-10-22 cs.SE cs.AI 89%

Challenges in Testing Large Language Model Based Software: A Faceted Taxonomy

Felix Dobslaw, Robert Feldt, Juyeon Yoon, Shin Yoo

机构 * Mid Sweden University(Mid Sweden大学) Chalmers University of Technology(查尔姆斯理工大学) KAIST(韩国科学技术院)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.18561 2025-10-22 cs.CL cs.AI 88%

Large language models for folktale type automation based on motifs: Cinderella case study

Tjaša Arčon, Marko Robnik-Šikonja, Polona Tratnik

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.18728 2025-10-22 cs.CR cs.AI 88%

HarmNet: A Framework for Adaptive Multi-Turn Jailbreak Attacks on Large Language Models

Sidhant Narula, Javad Rafiei Asl, Mohammad Ghasemigol, Eduardo Blanco, Daniel Takabi

机构 * University of Arizona(亚利桑那大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.AI

Comments This paper has been accepted for presentation at the Conference on Applied Machine Learning in Information Security (CAMLIS 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.22724 2025-10-22 cs.CL 88%

The Translation Barrier Hypothesis: Multilingual Generation with Large Language Models Suffers from Implicit Translation Failure

Niyati Bafna, Tianjian Li, Kenton Murray, David R. Mortensen, David Yarowsky, Hale Sirin, Daniel Khashabi

机构 * Johns Hopkins University, Center for Language and Speech Processing(约翰霍普金斯大学,语言与语音处理中心) Language Technologies Institute, Carnegie Mellon University(语言技术研究所,卡内基梅隆大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

Comments 28 pages, incl. appendix

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.18113 2025-10-22 cs.CR 85%

Investigating the Impact of Dark Patterns on LLM-Based Web Agents

Devin Ersoy, Brandon Lee, Ananth Shreekumar, Arjun Arunasalam, Muhammad Ibrahim, Antonio Bianchi, Z. Berkay Celik

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract)

Comments At IEEE S&P 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17575 2025-10-22 cs.HC 85%

DeTAILS: Deep Thematic Analysis with Iterative LLM Support

Ansh Sharma, Karen Cochrane, James R. Wallace

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.14456 2025-10-22 cs.CL cs.AI 79%

Correct-Detect: Balancing Performance and Ambiguity Through the Lens of Coreference Resolution in LLMs

Amber Shore, Russell Scheinberg, Ameeta Agrawal, So Young Lee

机构 * Portland State University(波特兰州立大学) Miami University(迈阿密大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);prompting(abstract);分类 cs.CL、cs.AI

Comments Accepted at EMNLP 2025 (main)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.19398 2025-10-22 cs.AI cs.GT cs.LG 79%

Do LLMs Strategically Reveal, Conceal, and Infer Information? A Theoretical and Empirical Analysis in The Chameleon Game

Mustafa O. Karabag, Jan Sobotka, Ufuk Topcu

机构 * The University of Texas at Austin(德克萨斯大学奥斯汀分校)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.18493 2025-10-22 cs.CR cs.AI cs.HC 77%

One Size Fits All? A Modular Adaptive Sanitization Kit (MASK) for Customizable Privacy-Preserving Phone Scam Detection

Kangzhong Wang, Zitong Shen, Youqian Zhang, Michael MK Cheung, Xiapu Luo, Grace Ngai, Eugene Yujun Fu

机构 * The Hong Kong Polytechnic University(香港理工大学) The Education University of Hong Kong(香港教育大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

Comments 9 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17875 2025-10-22 cs.CV cs.AI 57%

3D Weakly Supervised Semantic Segmentation via Class-Aware and Geometry-Guided Pseudo-Label Refinement

Xiaoxu Xu, Xuexun Liu, Jinlong Li, Yitian Yuan, Qiudan Zhang, Lin Ma, Nicu Sebe, Xu Wang

机构 * College of Computer Science, Beihang University(北京航空航天大学计算机科学学院) College of Computer Science and Software Engineering, Shenzhen University(深圳大学计算机科学与软件工程学院) Department of Information Engineering and Computer Science, University of Trento(特伦托大学信息工程与计算机科学系) Meituan(美团)

专题命中 其他LLM :language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏