arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-08-26 至 2025-08-26 共收录 16 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 16 篇

2508.17953 2025-08-26 cs.CL cs.AI cs.LG 90%

Understanding Subword Compositionality of Large Language Models

Qiwei Peng, Yekun Chai, Anders Søgaard

机构 * University of Copenhagen(哥本哈根大学) ETH Zurich(苏黎世联邦理工学院)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI、cs.LG

Comments EMNLP 2025 Main

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.11308 2025-08-26 cs.AI cs.CL cs.CR 88%

Defending against Jailbreak through Early Exit Generation of Large Language Models

Chongwen Zhao, Zhihao Dou, Kaizhu Huang

机构 * Duke Kunshan University(杜克昆山大学)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

Comments ICONIP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.21520 2025-08-26 cs.LG cs.CL 86%

LLM-Forest: Ensemble Learning of LLMs with Graph-Augmented Prompts for Data Imputation

Xinrui He, Yikun Ban, Jiaru Zou, Tianxin Wei, Curtiss B. Cook, Jingrui He

机构 * University of Illinois at Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Mayo Clinic Arizona(梅奥诊所阿兹根)

专题命中 知识编辑与模型理解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

Journal ref Findings of the Association for Computational Linguistics: ACL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.16969 2025-08-26 cs.CL cs.AI cs.DB 81%

Explaining Black-box Language Models with Knowledge Probing Systems: A Post-hoc Explanation Perspective

Yunxiao Zhao, Hao Xu, Zhiqiang Wang, Xiaoli Li, Jiye Liang, Ru Li

机构 * School of Computer and Information Technology, Shanxi University, China(山西大学计算机与信息学院) Key Laboratory of Computational Intelligence and Chinese Information Processing of Ministry of Education, Shanxi University, China(教育部计算智能与中文信息处理重点实验室) Institute for Infocomm Research, A*Star, Singapore(A*Star信息与通信研究所)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI

Comments 16 pages, 8 figures. This paper has been accepted by DASFAA 2025: The 30th International Conference on Database Systems for Advanced Applications

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.16697 2025-08-26 cs.CL cs.AI cs.LG 80%

QueryBandits for Hallucination Mitigation: Exploiting Semantic Features for No-Regret Rewriting

Nicole Cho, William Watson, Alec Koppel, Sumitra Ganesh, Manuela Veloso

机构 * JP Morgan AI Research(摩根大通人工智能研究)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);prompting(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.16765 2025-08-26 cs.CR cs.AI cs.CL 79%

Guarding Your Conversations: Privacy Gatekeepers for Secure Interactions with Cloud-Based AI Models

GodsGift Uzor, Hasan Al-Qudah, Ynes Ineza, Abdul Serwadda

机构 * Department of Computer Science, Texas Tech University, Lubbock, Texas(计算机科学系,德克萨斯理工大学,德克萨斯州拉伯克)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments 2025 19th International Conference on Semantic Computing (ICSC)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.16207 2025-08-26 cs.CV 78%

T-MASK: Temporal Masking for Probing Foundation Models across Camera Views in Driver Monitoring

Thinesh Thiyakesan Ponbagavathi, Kunyu Peng, Alina Roitberg

机构 * Institute of Artificial Intelligence(人工智能研究所) University of Stuttgart(斯图加特大学) Institute for Anthropomatics(人机学研究所) Karlsruhe Institute of Technology(卡尔斯鲁厄技术大学)

专题命中 知识编辑与模型理解 :foundation model(title,abstract)

Comments This paper has been accepted by 26th IEEE International Conference on Intelligent Transportation Systems ITSC 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.12964 2025-08-26 cs.CL 77%

Trust Me, I'm Wrong: LLMs Hallucinate with Certainty Despite Knowing the Answer

Adi Simhi, Itay Itzhak, Fazl Barez, Gabriel Stanovsky, Yonatan Belinkov

机构 * Technion – Israel Institute of Technology(技术ion-以色列理工学院) University of Oxford and WhiteBox(牛津大学和WhiteBox) School of Computer Science and Engineering, The Hebrew University of Jerusalem(耶路撒冷希伯来大学计算机科学与工程学院)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.17735 2025-08-26 cs.CL 77%

SMITE: Enhancing Fairness in LLMs through Optimal In-Context Example Selection via Dynamic Validation

Garima Chhikara, Kripabandhu Ghosh, Abhijnan Chakraborty

机构 * Indian Institute of Technology Delhi, India(印度德里理工学院) Delhi Technological University, India(德里技术大学) Indian Institute of Science Education and Research Kolkata, India(印度科学教育与研究学院科希拉分校) Indian Institute of Technology Kharagpur, India(印度理工学院哈里科格分校)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.03001 2025-08-26 cs.CV cs.MM 71%

One Framework to Rule Them All: Unifying Multimodal Tasks with LLM Neural-Tuning

Hao Sun, Yu Song, Jiaqing Liu, Jihong Hu, Yen-Wei Chen, Lanfen Lin

机构 * College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院) College of Information Science and Engineering, Ritsumeikan University(立命馆大学信息科学与工程学院)

专题命中 知识编辑与模型理解 :LLM(title)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.18001 2025-08-26 cs.LG stat.ML 70%

A Novel Framework for Uncertainty Quantification via Proper Scores for Classification and Beyond

Sebastian G. Gruber

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.LG

Comments PhD Thesis (cumulative, spanning 6 peer-reviewed publications)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.17760 2025-08-26 cs.CV cs.CL 70%

CEIDM: A Controlled Entity and Interaction Diffusion Model for Enhanced Text-to-Image Generation

Mingyue Yang, Dianxi Shi, Jialu Zhou, Xinyu Wei, Leqian Li, Shaowu Yang, Chunping Qiu

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.17856 2025-08-26 cs.CR cs.SE 67%

MalLoc: Toward Fine-grained Android Malicious Payload Localization via LLMs

Tiezhu Sun, Marco Alecci, Aleksandr Pilgun, Yewei Song, Xunzhu Tang, Jordan Samhi, Tegawendé F. Bissyandé, Jacques Klein

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

Comments Accepted at ICSME 2025, NIER Track

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.08243 2025-08-26 cs.CL 57%

Jinx: Unlimited LLMs for Probing Alignment Failures

Jiahao Zhao, Liwei Dong

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

Comments https://huggingface.co/Jinx-org

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.12051 2025-08-26 cs.LG cs.CE 57%

GUST: Quantifying Free-Form Geometric Uncertainty of Metamaterials Using Small Data

Jiahui Zheng, Cole Jahnke, Wei "Wayne" Chen

专题命中 知识编辑与模型理解 :pretraining(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.16837 2025-08-26 cs.CL 57%

LLMs Learn Constructions That Humans Do Not Know

Jonathan Dunn, Mai Mohamed Eida

机构 * Department of Linguistics University of Illinois Urbana-Champaign(语言学系伊利诺伊大学厄巴纳-香槟分校)

专题命中 知识编辑与模型理解 :LLM(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏