arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 7565 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 7565 篇

2506.00074 2025-09-11 cs.CY cs.AI cs.DL cs.IR cs.SI physics.soc-ph 74%

Whose Name Comes Up? Auditing LLM-Based Scholar Recommendations

Daniele Barolo, Chiara Valentin, Fariba Karimi, Luis Galárraga, Gonzalo G. Méndez, Lisette Espín-Noboa

专题命中 知识编辑与模型理解 :LLM(title);分类 cs.AI

Comments 40 pages: 10 main (incl. 9 figures), 3 references, and 27 appendix. Paper under-review

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.13155 2025-06-19 cs.CV cs.CL cs.MM 74%

Bi-VLDoc: Bidirectional Vision-Language Modeling for Visually-Rich Document Understanding

Chuwei Luo, Guozhi Tang, Qi Zheng, Cong Yao, Lianwen Jin, Chenliang Li, Yang Xue, Luo Si

机构 * Alibaba Group(阿里巴巴集团) School of Electronic and Information Engineering(电子与信息工程学院) South China University of Technology(华南理工大学)

专题命中 知识编辑与模型理解 :language model(title);分类 cs.CL

Comments IJDAR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.11021 2025-06-16 cs.SE cs.AI 74%

Eliminating Hallucination-Induced Errors in LLM Code Generation with Functional Clustering

Chaitanya Ravuri, Saman Amarasinghe

机构 * Massachusetts Institute of Technology(麻省理工学院)

专题命中 知识编辑与模型理解 :LLM(title);分类 cs.AI

Comments 9 pages, 1 figure

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.01833 2025-06-03 cs.LG q-bio.GN 74%

SPACE: Your Genomic Profile Predictor is a Powerful DNA Foundation Model

Zhao Yang, Jiwei Zhu, Bing Su

专题命中 知识编辑与模型理解 :foundation model(title);分类 cs.LG

Comments Accepted to ICML 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.04615 2025-05-28 cs.CL 74%

HalluCounter: Reference-free LLM Hallucination Detection in the Wild!

Ashok Urlana, Gopichand Kanumolu, Charaka Vinayak Kumar, Bala Mallikarjunarao Garlapati, Rahul Mishra

机构 * IIIT Hyderabad(IIIT海得拉巴) TCS Research, Hyderabad, India(TCS研究, 海得拉巴, 印度) University of Oslo, Norway(奥斯陆大学)

专题命中 知识编辑与模型理解 :LLM(title);分类 cs.CL

Comments 30 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.13465 2025-04-21 cs.LG 74%

Are you SURE? Enhancing Multimodal Pretraining with Missing Modalities through Uncertainty Estimation

Duy A. Nguyen, Quan Huu Do, Khoa D. Doan, Minh N. Do

专题命中 知识编辑与模型理解 :pretraining(title);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.04772 2025-04-08 cs.LG 74%

Feedback-Enhanced Hallucination-Resistant Vision-Language Model for Real-Time Scene Understanding

Zahir Alsulaimawi

专题命中 知识编辑与模型理解 :language model(title);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21932 2025-03-31 cs.CV cs.CE cs.LG 74%

Multimodal Data Integration for Sustainable Indoor Gardening: Tracking Anyplant with Time Series Foundation Model

Seyed Hamidreza Nabaei, Zeyang Zheng, Dong Chen, Arsalan Heydarian

专题命中 知识编辑与模型理解 :foundation model(title);分类 cs.LG

Comments Accepted at ASCE International Conference on Computing in Civil Engineering (i3ce)

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.01228 2025-03-13 cs.HC cs.AI 74%

The Interaction Layer: An Exploration for Co-Designing User-LLM Interactions in Parental Wellbeing Support Systems

Sruthi Viswanathan, Seray Ibrahim, Ravi Shankar, Reuben Binns, Max Van Kleek, Petr Slovak

专题命中 知识编辑与模型理解 :LLM(title);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.07923 2025-02-28 cs.CL 74%

Word Boundary Information Isn't Useful for Encoder Language Models

Edward Gow-Smith, Dylan Phelps, Harish Tayyar Madabushi, Carolina Scarton, Aline Villavicencio

专题命中 知识编辑与模型理解 :language model(title);分类 cs.CL

Comments 9th Workshop on Representation Learning for NLP

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.05208 2025-02-10 cs.CV cs.CL 74%

Getting More Juice Out of Your Data: Hard Pair Refinement Enhances Visual-Language Models Without Extra Data

Haonan Wang, Minbin Huang, Runhui Huang, Lanqing Hong, Hang Xu, Tianyang Hu, Xiaodan Liang, Zhenguo Li, Hong Cheng, Kenji Kawaguchi

专题命中 知识编辑与模型理解 :language model(title);分类 cs.CL

Comments Accepted to NAACL 2025, main conference. 20 pages, 10 figures, 10 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.06833 2025-02-07 cs.CL 74%

Does Mapo Tofu Contain Coffee? Probing LLMs for Food-related Cultural Knowledge

Li Zhou, Taelin Karidi, Wanlong Liu, Nicolas Garneau, Yong Cao, Wenyu Chen, Haizhou Li, Daniel Hershcovich

专题命中 知识编辑与模型理解 :large language model(abstract,comments);language model(abstract,comments);分类 cs.CL

Comments cultural bias analysis, cultural knowledge probing, large language models, cultural NLP; Accepted by NAACL2025

Journal ref NAACL2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.09352 2025-01-17 cs.LG cs.MM eess.IV 74%

PAL: Prompting Analytic Learning with Missing Modality for Multi-Modal Class-Incremental Learning

Xianghu Yue, Yiming Chen, Xueyi Zhang, Xiaoxue Gao, Mengling Feng, Mingrui Lao, Huiping Zhuang, Haizhou Li

专题命中 知识编辑与模型理解 :prompting(title);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.00847 2024-12-31 cs.DB cs.AI cs.IR 74%

The Design of an LLM-powered Unstructured Analytics System

Eric Anderson, Jonathan Fritz, Austin Lee, Bohou Li, Mark Lindblad, Henry Lindeman, Alex Meyer, Parth Parmar, Tanvi Ranade, Mehul A. Shah, Benjamin Sowell, Dan Tecuci, Vinayak Thapliyal, Matt Welsh

专题命中 知识编辑与模型理解 :LLM(title);分类 cs.AI

Comments Included in the proceedings of The Conference on Innovative Data Systems Research (CIDR) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.08821 2024-12-17 cs.CL 74%

Large Concept Models: Language Modeling in a Sentence Representation Space

LCM team, Loïc Barrault, Paul-Ambroise Duquenne, Maha Elbayad, Artyom Kozhevnikov, Belen Alastruey, Pierre Andrews, Mariano Coria, Guillaume Couairon, Marta R. Costa-jussà, David Dale, Hady Elsahar, Kevin Heffernan, João Maria Janeiro, Tuan Tran, Christophe Ropers, Eduardo Sánchez, Robin San Roman, Alexandre Mourachko, Safiyyah Saleem, Holger Schwenk

专题命中 知识编辑与模型理解 :language model(title);分类 cs.CL

Comments 49 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.03621 2024-11-01 cs.CL 74%

Attend First, Consolidate Later: On the Importance of Attention in Different LLM Layers

Amit Ben-Artzy, Roy Schwartz

专题命中 知识编辑与模型理解 :LLM(title);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.06675 2024-02-13 cs.LG 74%

A Masked language model for multi-source EHR trajectories contextual representation learning

Ali Amirahmadi, Mattias Ohlsson, Kobra Etminani, Olle Melander, Jonas Björk

专题命中 知识编辑与模型理解 :language model(title);分类 cs.LG

Comments Presented at Proceedings of MIE 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.12416 2023-12-20 cs.CV cs.LG 74%

Prompting Hard or Hardly Prompting: Prompt Inversion for Text-to-Image Diffusion Models

Shweta Mahajan, Tanzila Rahman, Kwang Moo Yi, Leonid Sigal

专题命中 知识编辑与模型理解 :prompting(title);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.02298 2023-12-08 cs.SD cs.AI eess.AS 74%

Prompting Audios Using Acoustic Properties For Emotion Representation

Hira Dhamyal, Benjamin Elizalde, Soham Deshmukh, Huaming Wang, Bhiksha Raj, Rita Singh

专题命中 知识编辑与模型理解 :prompting(title);分类 cs.AI

Comments arXiv admin note: substantial text overlap with arXiv:2211.07737

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.18957 2023-10-12 cs.CL 74%

Wave to Syntax: Probing spoken language models for syntax

Gaofei Shen, Afra Alishahi, Arianna Bisazza, Grzegorz Chrupała

专题命中 知识编辑与模型理解 :language model(title);分类 cs.CL

Comments Accepted to Interspeech 2023

Journal ref Proceedings of Interspeech 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2112.02399 2023-08-11 cs.CV cs.CL 74%

VT-CLIP: Enhancing Vision-Language Models with Visual-guided Texts

Longtian Qiu, Renrui Zhang, Ziyu Guo, Ziyao Zeng, Zilu Guo, Yafeng Li, Guangnan Zhang

专题命中 知识编辑与模型理解 :language model(title);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.03588 2023-04-11 cs.SD cs.LG eess.AS 74%

Anomalous Sound Detection using Audio Representation with Machine ID based Contrastive Learning Pretraining

Jian Guan, Feiyang Xiao, Youde Liu, Qiaoxi Zhu, Wenwu Wang

专题命中 知识编辑与模型理解 :pretraining(title);分类 cs.LG

Comments To appear in IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP 2023)

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.13812 2023-02-28 quant-ph cs.CL 74%

Adapting Pre-trained Language Models for Quantum Natural Language Processing

Qiuchi Li, Benyou Wang, Yudong Zhu, Christina Lioma, Qun Liu

专题命中 知识编辑与模型理解 :language model(title);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.04412 2022-12-09 cs.CV cs.LG 74%

Task Bias in Vision-Language Models

Sachit Menon, Ishaan Preetam Chandratreya, Carl Vondrick

专题命中 知识编辑与模型理解 :language model(title);分类 cs.LG

Comments First two authors contributed equally

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.09259 2022-06-22 cs.CL 74%

Can Language Models Capture Graph Semantics? From Graphs to Language Model and Vice-Versa

Tarun Garg, Kaushik Roy, Amit Sheth

专题命中 知识编辑与模型理解 :language model(title);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.06666 2020-10-15 cs.CL 74%

Probing for Multilingual Numerical Understanding in Transformer-Based Language Models

Devin Johnson, Denise Mak, Drew Barker, Lexi Loessberg-Zahl

专题命中 知识编辑与模型理解 :language model(title);分类 cs.CL

Comments BlackboxNLP (EMNLP 2020)

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.18419 2026-08-20 cs.LG cs.AI 新提交 73%

Mechanistic Interpretability of Structure-Aware Numerical Reasoning in LLaMA 3.1 8B

LLaMA 3.1 8B中结构感知数值推理的机制可解释性

Rahul Chowdhury, Timothy A Rupprecht, Senhao Cao, Jiahao Liu, Octavia Camps, David Bau, Pu Zhao, Yanzhi Wang

机构 * Northeastern University(东北大学) EmbodyX Inc.(EmbodyX公司)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本研究从机制可解释性视角探究LLaMA 3.1 8B,通过构建需捕捉结构的数值序列任务,发现其可无监督计算存储一阶差分,还揭示其通过类诱导回路机制完成数值推理。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.18106 2026-08-20 cs.CL cs.AI 新提交 73%

Different Facets of Verbalised Overconfidence: an Interpretability Study

言语化过度自信的不同方面:一项可解释性研究

Davide Mazzaccara, Leonardo Bertolazzi, Raffaella Bernardi

机构 * CIMeC, University of Trento(特伦托大学认知科学与技术跨学科研究中心) DISI, University of Trento(特伦托大学信息工程与计算机科学系) Free University of Bozen-Bolzano(波尔扎诺自由大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 本研究以Qwen3-4B为对象,探究大型语言模型的过度自信现象,通过识别转码器特征揭示其默认机制偏向确定性生成,干预不确定性特征可缓解过度自信错误,且相关特征具有跨场景泛化性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.16353 2026-08-18 cs.CL cs.AI 新提交 73%

HalluTracer: Hallucination Detection via Depth-Averaging Truth Signals

HalluTracer:基于深度平均真值信号的幻觉检测方法

Zhihao Guo, Zonghan Wu, Huan Huo, DaYong Ye, Junwei Zhang, Weiran Yao, Zhiwei Liu, Qingsong Wen, Yilei Shao

机构 * University of Technology Sydney(悉尼科技大学) City University of Macau(澳门城市大学) Meta actAVA AI(actAVA人工智能公司) Microsoft AI(微软人工智能) Squirrel Ai Learning(松鼠人工智能学习公司) East China Normal University(华东师范大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 HalluTracer是一种在模型生成答案前聚合各层真值证据的幻觉检测框架,在六个开源语言模型和五个基准上优于白盒基线,将幻觉检测转化为深度聚合问题。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.14659 2026-08-18 cs.AI cs.LG cs.SE 新提交 73%

When Uncertainty Isn't Enough: An Empirical Study of Self-Correction in Code Generation

当不确定性还不够时:代码生成中自校正的实证研究

Pranav Rakasi, Maanas Lalwani, Arnav Srivastava, Arya Palanivel, Tinuade Adeleke, Ruizhe Li, Sean Wu

机构 * University of Michigan(密歇根大学) New York University(纽约大学) University of Wisconsin-Madison(威斯康星大学麦迪逊分校) Algoverse AI University of Aberdeen(阿伯丁大学) University of Oxford(牛津大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本研究通过实证发现,针对代码生成的不确定性估计方法难以可靠提升生成准确率,仅基于验证的自校正策略可显著提升Pass@1指标,廉价不确定性估计器仅适合作为校正循环的门控信号。

详情

展开后加载摘要…

URL PDF HTML 收藏