arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 7608 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 7608 篇

2402.08680 2025-06-13 cs.LG cs.AI cs.CL cs.CV 82%

Mitigating Object Hallucination in Large Vision-Language Models via Image-Grounded Guidance

Linxi Zhao, Yihe Deng, Weitong Zhang, Quanquan Gu

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI、cs.LG

Comments 25 pages, 13 figures, 25 tables

Journal ref In Proceedings of the 42nd International Conference on Machine Learning, Vancouver, Canada. PMLR 267, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.01917 2025-05-23 cs.LG cs.AI cs.CL 82%

Steer LLM Latents for Hallucination Detection

Seongheon Park, Xuefeng Du, Min-Hsuan Yeh, Haobo Wang, Yixuan Li

专题命中 知识编辑与模型理解 :LLM(title,abstract);分类 cs.CL、cs.AI、cs.LG

Comments ICML 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.10378 2025-05-20 cs.CL cs.AI cs.HC cs.LG 82%

Cross-Lingual Consistency of Factual Knowledge in Multilingual Language Models

Jirui Qi, Raquel Fernández, Arianna Bisazza

机构 * Center for Language and Cognition, University of Groningen(格罗宁根大学语言与认知中心) Institute for Logic, Language and Computation, University of Amsterdam(阿姆斯特丹大学逻辑、语言与计算研究所)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI、cs.LG

Comments EMNLP2023 Outstanding Paper (Multilinguality and Linguistic Diversity Track). All code and data are released at https://github.com/Betswish/Cross-Lingual-Consistency

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.17671 2025-05-16 cs.CL cs.AI cs.LG 82%

Data-Driven Calibration of Prediction Sets in Large Vision-Language Models Based on Inductive Conformal Prediction

Yuanchang Ye, Weiyan Wen

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI、cs.LG

Comments Accepted by ICIPCA 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.03757 2025-05-09 cs.CV cs.AI cs.CL cs.LG cs.RO 82%

Lexicon3D: Probing Visual Foundation Models for Complex 3D Scene Understanding

Yunze Man, Shuhong Zheng, Zhipeng Bao, Martial Hebert, Liang-Yan Gui, Yu-Xiong Wang

机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Carnegie Mellon University(卡内基梅隆大学)

专题命中 知识编辑与模型理解 :foundation model(title,abstract);分类 cs.CL、cs.AI、cs.LG

Comments NeurIPS 2024. Project page: https://yunzeman.github.io/lexicon3d Github: https://github.com/YunzeMan/Lexicon3D

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.19647 2025-03-28 cs.LG cs.AI cs.CL 82%

Sparse Feature Circuits: Discovering and Editing Interpretable Causal Graphs in Language Models

Samuel Marks, Can Rager, Eric J. Michaud, Yonatan Belinkov, David Bau, Aaron Mueller

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI、cs.LG

Comments Code and data at https://github.com/saprmarks/feature-circuits. Demonstration at https://feature-circuits.xyz

Journal ref International Conference on Learning Representations, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.13623 2025-03-25 cs.CV 82%

Unsupervised Foundation Model-Agnostic Slide-Level Representation Learning

Tim Lenz, Peter Neidlinger, Marta Ligero, Georg Wölflein, Marko van Treeck, Jakob Nikolas Kather

专题命中 知识编辑与模型理解 :foundation model(title,abstract);pretraining(abstract)

Comments Got accepted at CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.09668 2025-03-21 cs.CV 82%

Vision-Language Models Generate More Homogeneous Stories for Phenotypically Black Individuals

Messi H. J. Lee, Soyeon Jeon

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.08870 2025-03-04 cs.CV 82%

Vista-LLaMA: Reducing Hallucination in Video Language Models via Equal Distance to Visual Tokens

Fan Ma, Xiaojie Jin, Heng Wang, Yuchen Xian, Jiashi Feng, Yi Yang

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.00717 2025-02-04 cs.CV 82%

MINT: Mitigating Hallucinations in Large Vision-Language Models via Token Reduction

Chao Wang, Jianming Yang, Yang Zhou

专题命中 知识编辑与模型理解 :language model(title,abstract);LLM(abstract)

Comments 8 pages, 5 figures, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.02886 2024-12-17 cs.CV 82%

Patchfinder: Leveraging Visual Language Models for Accurate Information Retrieval using Model Uncertainty

Roman Colman, Minh Vu, Manish Bhattarai, Martin Ma, Hari Viswanathan, Daniel O'Malley, Javier E. Santos

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract)

Comments This paper has been accepted to IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.07802 2024-12-12 cs.CV 82%

Language Model as Visual Explainer

Xingyi Yang, Xinchao Wang

专题命中 知识编辑与模型理解 :language model(title,abstract);LLM(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.09894 2024-11-18 cs.CV 82%

Free Lunch in Pathology Foundation Model: Task-specific Model Adaptation with Concept-Guided Feature Enhancement

Yanyan Huang, Weiqin Zhao, Yihang Chen, Yu Fu, Lequan Yu

专题命中 知识编辑与模型理解 :foundation model(title,abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.15814 2024-11-08 cs.CL cs.AI cs.LG 82%

Perceptions of Linguistic Uncertainty by Language Models and Humans

Catarina G Belem, Markelle Kelly, Mark Steyvers, Sameer Singh, Padhraic Smyth

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI、cs.LG

Comments Accepted at EMNLP 2024 (Main)

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.00113 2024-10-31 cs.LG cs.AI cs.CL 82%

Measuring Progress in Dictionary Learning for Language Model Interpretability with Board Game Models

Adam Karvonen, Benjamin Wright, Can Rager, Rico Angell, Jannik Brinkmann, Logan Smith, Claudio Mayrink Verdun, David Bau, Samuel Marks

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI、cs.LG

Comments Accepted as an oral paper (top 5%) at the ICML 2024 Mechanistic Interpretability Workshop and to the NeurIPS 2024 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.03964 2024-10-30 cs.LG cs.AI cs.CL stat.ML 82%

Variational Language Concepts for Interpreting Foundation Language Models

Hengyi Wang, Shiwei Tan, Zhiqing Hong, Desheng Zhang, Hao Wang

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI、cs.LG

Comments Accepted at EMNLP 2024 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.19283 2024-10-08 eess.AS cs.SD 82%

Analyzing and Mitigating Inconsistency in Discrete Audio Tokens for Neural Codec Language Models

Wenrui Liu, Zhifang Guo, Jin Xu, Yuanjun Lv, Yunfei Chu, Zhou Zhao, Junyang Lin

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract)

Comments e.g.: 15 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.15604 2024-09-25 cs.HC 82%

Persona-L has Entered the Chat: Leveraging LLM and Ability-based Framework for Personas of People with Complex Needs

Lipeipei Sun, Tianzi Qin, Anran Hu, Jiale Zhang, Shuojia Lin, Jianyan Chen, Mona Ali, Mirjana Prpa

专题命中 知识编辑与模型理解 :LLM(title);large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.08263 2024-09-11 cs.CL cs.AI cs.LG cs.SI 82%

Relational Prompt-based Pre-trained Language Models for Social Event Detection

Pu Li, Xiaoyan Yu, Hao Peng, Yantuan Xian, Linqin Wang, Li Sun, Jingyun Zhang, Philip S. Yu

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI、cs.LG

Comments ACM TOIS

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.16998 2024-08-19 cs.CL cs.AI cs.LG cs.SD eess.AS 82%

What Do Language Models Hear? Probing for Auditory Representations in Language Models

Jerry Ngo, Yoon Kim

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI、cs.LG

Journal ref 2024.acl-long.297

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.01037 2024-08-12 cs.LG cs.AI cs.CL 82%

Eliciting Latent Knowledge from Quirky Language Models

Alex Mallen, Madeline Brumley, Julia Kharchenko, Nora Belrose

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI、cs.LG

Comments COLM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.11169 2024-08-06 cs.LG cs.AI cs.CL cs.PL 82%

Emergent Representations of Program Semantics in Language Models Trained on Programs

Charles Jin, Martin Rinard

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI、cs.LG

Comments ICML 2024

Journal ref PMLR 235:22160-22184, 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.17100 2024-08-02 cs.CV 82%

Open-Set Video-based Facial Expression Recognition with Human Expression-sensitive Prompting

Yuanyuan Liu, Yuxuan Huang, Shuyang Liu, Yibing Zhan, Zijing Chen, Zhe Chen

专题命中 知识编辑与模型理解 :prompting(title,abstract);language model(abstract)

Comments Accepted by ACM MM2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.16908 2024-07-25 cs.CL cs.AI cs.LG 82%

Generation Constraint Scaling Can Mitigate Hallucination

Georgios Kollias, Payel Das, Subhajit Chaudhury

专题命中 知识编辑与模型理解 :large language model(abstract,comments);language model(abstract,comments);LLM(abstract);分类 cs.CL、cs.AI、cs.LG

Comments 7 pages; accepted at ICML 2024 Workshop on Large Language Models and Cognition

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.08811 2024-07-15 eess.IV cs.CV 82%

CXR-Agent: Vision-language models for chest X-ray interpretation with uncertainty aware radiology reporting

Naman Sharma

专题命中 知识编辑与模型理解 :language model(title,abstract);language agent(abstract)

Comments Supervised by Professor Ben Glocker

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.03118 2024-06-26 cs.CV 82%

LVLM-Interpret: An Interpretability Tool for Large Vision-Language Models

Gabriela Ben Melech Stan, Estelle Aflalo, Raanan Yehezkel Rohekar, Anahita Bhiwandiwalla, Shao-Yen Tseng, Matthew Lyle Olson, Yaniv Gurwicz, Chenfei Wu, Nan Duan, Vasudev Lal

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.08572 2024-06-14 cs.CV 82%

LLM-assisted Concept Discovery: Automatically Identifying and Explaining Neuron Functions

Nhat Hoang-Xuan, Minh Vu, My T. Thai

专题命中 知识编辑与模型理解 :LLM(title);large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.10412 2024-06-10 cs.CL cs.AI cs.LG 82%

Measuring and Reducing LLM Hallucination without Gold-Standard Answers

Jiaheng Wei, Yuanshun Yao, Jean-Francois Ton, Hongyi Guo, Andrew Estornell, Yang Liu

专题命中 知识编辑与模型理解 :LLM(title,abstract);分类 cs.CL、cs.AI、cs.LG

Comments Paper Under Review

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.01345 2024-05-16 cs.CV cs.AI cs.CL cs.LG 82%

Skip \n: A Simple Method to Reduce Hallucination in Large Vision-Language Models

Zongbo Han, Zechen Bai, Haiyang Mei, Qianli Xu, Changqing Zhang, Mike Zheng Shou

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.10136 2024-04-17 cs.CL cs.AI cs.LG 82%

Language Model Cascades: Token-level uncertainty and beyond

Neha Gupta, Harikrishna Narasimhan, Wittawat Jitkrittum, Ankit Singh Rawat, Aditya Krishna Menon, Sanjiv Kumar

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏