arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 7584 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 7584 篇

2407.19200 2025-02-05 cs.CL cs.AI 62%

On Behalf of the Stakeholders: Trends in NLP Model Interpretability in the Era of LLMs

Nitay Calderon, Roi Reichart

专题命中 知识编辑与模型理解 :LLM(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.00305 2025-02-04 cs.CL cs.AI cs.IR 62%

DEUCE: Dual-diversity Enhancement and Uncertainty-awareness for Cold-start Active Learning

Jiaxin Guo, C. L. Philip Chen, Shuzhen Li, Tong Zhang

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI

Comments 18 pages, 3 figures, 12 tables. Accepted manuscript by TACL. For published version by MIT Press, see https://direct.mit.edu/tacl/article/doi/10.1162/tacl_a_00731/125950

Journal ref Transactions of the Association for Computational Linguistics, Vol. 12 (2024), pp. 1736-1754

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.15264 2025-02-03 cs.CL cs.AI 62%

ReXTrust: A Model for Fine-Grained Hallucination Detection in AI-Generated Radiology Reports

Romain Hardy, Sung Eun Kim, Du Hyun Ro, Pranav Rajpurkar

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI

Comments Accepted to AIMedHealth 10 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.00715 2025-01-28 cs.CV cs.AI cs.LG 62%

B-cosification: Transforming Deep Neural Networks to be Inherently Interpretable

Shreyash Arya, Sukrut Rao, Moritz Böhle, Bernt Schiele

专题命中 知识编辑与模型理解 :foundation model(abstract);分类 cs.AI、cs.LG

Comments 31 pages, 9 figures, 12 tables, Neural Information Processing Systems (NeurIPS) 2024; added references, corrected typos

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.16302 2025-01-20 eess.AS cs.CL cs.LG cs.SD 62%

How Redundant Is the Transformer Stack in Speech Representation Models?

Teresa Dorszewski, Albert Kjøller Jacobsen, Lenka Tětková, Lars Kai Hansen

专题命中 知识编辑与模型理解 :post-training(abstract);分类 cs.CL、cs.LG

Comments To appear at ICASSP 2025 (excluding appendix)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.05503 2025-01-13 cs.CL cs.LG 62%

The more polypersonal the better -- a short look on space geometry of fine-tuned layers

Sergei Kudriashov, Veronika Zykova, Angelina Stepanova, Yakov Raskind, Eduard Klyshinsky

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.LG

Comments Neuroinformatics 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.05485 2025-01-13 cs.CL cs.IR cs.LG 62%

S2 Chunking: A Hybrid Framework for Document Segmentation Through Integrated Spatial and Semantic Analysis

Prashant Verma

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.05398 2025-01-10 cs.LG cs.AI 62%

Mechanistic understanding and validation of large AI models with SemanticLens

Maximilian Dreyer, Jim Berend, Tobias Labarta, Johanna Vielhaben, Thomas Wiegand, Sebastian Lapuschkin, Wojciech Samek

专题命中 知识编辑与模型理解 :foundation model(abstract);分类 cs.AI、cs.LG

Comments 74 pages (18 pages manuscript, 7 pages references, 49 pages appendix)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.00053 2025-01-03 eess.IV cs.AI cs.LG 62%

Implementing Trust in Non-Small Cell Lung Cancer Diagnosis with a Conformalized Uncertainty-Aware AI Framework in Whole-Slide Images

Xiaoge Zhang, Tao Wang, Chao Yan, Fedaa Najdawi, Kai Zhou, Yuan Ma, Yiu-ming Cheung, Bradley A. Malin

专题命中 知识编辑与模型理解 :foundation model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.17921 2024-12-25 cs.LG cs.CL 62%

VITRO: Vocabulary Inversion for Time-series Representation Optimization

Filippos Bellos, Nam H. Nguyen, Jason J. Corso

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.LG

Comments Accepted to ICASSP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.02309 2024-12-19 cs.CV cs.AI cs.LG 62%

Semantically Guided Representation Learning For Action Anticipation

Anxhelo Diko, Danilo Avola, Bardh Prenkaj, Federico Fontana, Luigi Cinque

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI、cs.LG

Comments Accepted as a full paper at ECCV'24 with Paper ID #4140

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.08124 2024-12-19 cs.CL cs.AI 62%

Legend: Leveraging Representation Engineering to Annotate Safety Margin for Preference Datasets

Duanyu Feng, Bowen Qin, Chen Huang, Youcheng Huang, Zheng Zhang, Wenqiang Lei

专题命中 知识编辑与模型理解 :LLM(abstract);分类 cs.CL、cs.AI

Comments Our code is available at https://github.com/colfeng/Legend

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.09701 2024-12-16 cs.LG cs.AI 62%

CUAL: Continual Uncertainty-aware Active Learner

Amanda Rios, Ibrahima Ndiour, Parual Datta, Jerry Sydir, Omesh Tickoo, Nilesh Ahuja

专题命中 知识编辑与模型理解 :foundation model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.19370 2024-12-12 cs.LG cs.AI 62%

Emergence of Hidden Capabilities: Exploring Learning Dynamics in Concept Space

Core Francisco Park, Maya Okawa, Andrew Lee, Hidenori Tanaka, Ekdeep Singh Lubana

专题命中 知识编辑与模型理解 :prompting(abstract);分类 cs.AI、cs.LG

Comments NeurIPS 2024 (Spotlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.03178 2024-12-05 cs.AI cs.CV cs.LG 62%

Towards Understanding and Quantifying Uncertainty for Text-to-Image Generation

Gianni Franchi, Dat Nguyen Trong, Nacim Belkhir, Guoxuan Xia, Andrea Pilzer

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI、cs.LG

Comments 28 pages and 22 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.18895 2024-12-02 cs.LG cs.CL 62%

Evaluating Sparse Autoencoders on Targeted Concept Erasure Tasks

Adam Karvonen, Can Rager, Samuel Marks, Neel Nanda

专题命中 知识编辑与模型理解 :LLM(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.14768 2024-11-25 cs.LG cs.AI 62%

Grid and Road Expressions Are Complementary for Trajectory Representation Learning

Silin Zhou, Shuo Shang, Lisi Chen, Peng Han, Christian S. Jensen

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI、cs.LG

Comments This paper is accepted by KDD2025(August Cycle)

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.02124 2024-11-11 cs.LG cs.AI 62%

Adaptive Sparse Allocation with Mutual Choice & Feature Choice Sparse Autoencoders

Kola Ayonrinde

专题命中 知识编辑与模型理解 :foundation model(abstract);分类 cs.AI、cs.LG

Comments 10 pages (18 w/ appendices), 7 figures. Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.04165 2024-11-08 q-bio.BM cs.AI cs.LG 62%

Bio-xLSTM: Generative modeling, representation and in-context learning of biological and chemical sequences

Niklas Schmidinger, Lisa Schneckenreiter, Philipp Seidl, Johannes Schimunek, Pieter-Jan Hoedt, Johannes Brandstetter, Andreas Mayr, Sohvi Luukkonen, Sepp Hochreiter, Günter Klambauer

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.02125 2024-11-05 cs.LG cs.AI cs.CE q-bio.GN 62%

Revisiting K-mer Profile for Effective and Scalable Genome Representation Learning

Abdulkadir Celikkanat, Andres R. Masegosa, Thomas D. Nielsen

专题命中 知识编辑与模型理解 :foundation model(abstract);分类 cs.AI、cs.LG

Comments Accepted to the Thirty-Eighth Annual Conference on Neural Information Processing Systems (NeurIPS 2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.23272 2024-10-31 cs.LG cs.AI 62%

A Monte Carlo Framework for Calibrated Uncertainty Estimation in Sequence Prediction

Qidong Yang, Weicheng Zhu, Joseph Keslin, Laure Zanna, Tim G. J. Rudner, Carlos Fernandez-Granda

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.20526 2024-10-29 cs.LG cs.CL 62%

Llama Scope: Extracting Millions of Features from Llama-3.1-8B with Sparse Autoencoders

Zhengfu He, Wentao Shu, Xuyang Ge, Lingjie Chen, Junxuan Wang, Yunhua Zhou, Frances Liu, Qipeng Guo, Xuanjing Huang, Zuxuan Wu, Yu-Gang Jiang, Xipeng Qiu

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.LG

Comments 22pages, 12 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2108.11574 2024-10-29 cs.CL cs.AI 62%

Multilingual Multi-Aspect Explainability Analyses on Machine Reading Comprehension Models

Yiming Cui, Wei-Nan Zhang, Wanxiang Che, Ting Liu, Zhigang Chen, Shijin Wang

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI

Comments 15 pages

Journal ref iScience 25(5), 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.12982 2024-10-22 cs.LG cs.CL cs.IR 62%

Retrieval-Enhanced Machine Learning: Synthesis and Opportunities

To Eun Kim, Alireza Salemi, Andrew Drozdov, Fernando Diaz, Hamed Zamani

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.04727 2024-10-17 cs.LG cs.AI 62%

Task Aware Modulation using Representation Learning: An Approach for Few Shot Learning in Environmental Systems

Arvind Renganathan, Rahul Ghosh, Ankush Khandelwal, Vipin Kumar

专题命中 知识编辑与模型理解 :pretraining(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.07882 2024-10-15 cs.CL cs.AI cs.HC 62%

Designing a Dashboard for Transparency and Control of Conversational AI

Yida Chen, Aoyu Wu, Trevor DePodesta, Catherine Yeh, Kenneth Li, Nicholas Castillo Marin, Oam Patel, Jan Riecke, Shivam Raval, Olivia Seow, Martin Wattenberg, Fernanda Viégas

专题命中 知识编辑与模型理解 :LLM(abstract);分类 cs.CL、cs.AI

Comments Project page: https://bit.ly/talktuner-project-page, 38 pages, 23 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.10728 2024-10-15 cs.CL cs.AI cs.IT math.IT 62%

Generalized Measures of Anticipation and Responsivity in Online Language Processing

Mario Giulianelli, Andreas Opedal, Ryan Cotterell

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI

Comments Findings of the Association for Computational Linguistics: EMNLP 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.18909 2024-10-11 cs.CL cs.AI 62%

AKEW: Assessing Knowledge Editing in the Wild

Xiaobao Wu, Liangming Pan, William Yang Wang, Anh Tuan Luu

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI

Comments Accepted to EMNLP 2024 main conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.15470 2024-10-08 cs.CL cs.AI cs.SI 62%

Mental Disorder Classification via Temporal Representation of Text

Raja Kumar, Kishan Maharaj, Ashita Saxena, Pushpak Bhattacharyya

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI

Comments RK and KM contributed equally to this work, 15 pages, 5 figures, 9 table

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.03396 2024-10-07 cs.LG cs.AI 62%

GraphCroc: Cross-Correlation Autoencoder for Graph Structural Reconstruction

Shijin Duan, Ruyi Ding, Jiaxing He, Aidong Adam Ding, Yunsi Fei, Xiaolin Xu

专题命中 知识编辑与模型理解 :prompting(abstract);分类 cs.AI、cs.LG

Comments 22 pages, 16 figures. Accepted in NeurIPS 2024

详情

展开后加载摘要…

URL PDF HTML 收藏