arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 7596 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 7596 篇

2403.01373 2024-05-07 cs.CL 79%

Quantity Matters: Towards Assessing and Mitigating Number Hallucination in Large Vision-Language Models

Huixuan Zhang, Junzhe Zhang, Xiaojun Wan

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

Comments 10 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.14467 2024-04-23 cs.HC cs.CL cs.CY 79%

Recourse for reclamation: Chatting with generative language models

Jennifer Chien, Kevin R. McKee, Jackie Kay, William Isaac

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

Comments Extended Abstracts of the CHI Conference on Human Factors in Computing Systems (CHI EA 2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.12686 2024-04-12 cs.AI 79%

On the Computation of Meaning, Language Models and Incomprehensible Horrors

Michael Timothy Bennett

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.AI

Comments Published (and accepted for full oral presentation) at the 16th Conference on Artificial General Intelligence, Stockholm, 2023

Journal ref Proceedings of the 16th International Conference on Artificial General Intelligence. 2023. Lecture Notes in Computer Science, vol 13921. Springer. pp. 32-41

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.07004 2024-04-11 cs.CL 79%

LM Transparency Tool: Interactive Tool for Analyzing Transformer Language Models

Igor Tufanov, Karen Hambardzumyan, Javier Ferrando, Elena Voita

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.00825 2024-04-11 cs.CV cs.AI 79%

SocialCounterfactuals: Probing and Mitigating Intersectional Social Biases in Vision-Language Models with Counterfactual Examples

Phillip Howard, Avinash Madasu, Tiep Le, Gustavo Lujan Moreno, Anahita Bhiwandiwalla, Vasudev Lal

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.AI

Comments Accepted to CVPR 2024. arXiv admin note: text overlap with arXiv:2310.02988

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.17299 2024-03-27 cs.CL q-bio.NC 79%

Decoding Probing: Revealing Internal Linguistic Structures in Neural Language Models using Minimal Pairs

Linyang He, Peili Chen, Ercong Nie, Yuanning Li, Jonathan R. Brennan

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

Comments Accepted by LREC-COLING 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.11025 2024-03-19 cs.CL 79%

Pre-Trained Language Models Represent Some Geographic Populations Better Than Others

Jonathan Dunn, Benjamin Adams, Harish Tayyar Madabushi

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.06204 2024-03-12 cs.CL 79%

Identifying and interpreting non-aligned human conceptual representations using language modeling

Wanqian Bao, Uri Hasson

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

Comments To appear at the ICLR 2024 Workshop on Representational Alignment (Re-Align)

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.05534 2024-03-11 cs.CL 79%

Bayesian Preference Elicitation with Language Models

Kunal Handa, Yarin Gal, Ellie Pavlick, Noah Goodman, Jacob Andreas, Alex Tamkin, Belinda Z. Li

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.00923 2024-03-08 cs.IR cs.CL 79%

An Interpretable Ensemble of Graph and Language Models for Improving Search Relevance in E-Commerce

Nurendra Choudhary, Edward W Huang, Karthik Subbian, Chandan K. Reddy

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

Comments Accepted to The Web Conference 2024 (Industry)

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.02009 2024-03-05 cs.CL 79%

Topic Aware Probing: From Sentence Length Prediction to Idiom Identification how reliant are Neural Language Models on Topic?

Vasudevan Nedumpozhimana, John D. Kelleher

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.08796 2024-02-29 cs.CL 79%

Chinese Spelling Correction as Rephrasing Language Model

Linfeng Liu, Hongqiu Wu, Hai Zhao

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

Comments Accepted by AAAI'2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.17671 2024-02-28 cs.LG 79%

Securing Reliability: A Brief Overview on Enhancing In-Context Learning for Foundation Models

Yunpeng Huang, Yaonan Gu, Jingwei Xu, Zhihong Zhu, Zhaorun Chen, Xiaoxing Ma

专题命中 知识编辑与模型理解 :foundation model(title,abstract);分类 cs.LG

Comments 18 pages, 15 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.01740 2024-02-20 cs.CL 79%

SAC3: Reliable Hallucination Detection in Black-Box Language Models via Semantic-aware Cross-check Consistency

Jiaxin Zhang, Zhuohang Li, Kamalika Das, Bradley A. Malin, Sricharan Kumar

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

Comments EMNLP 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.10239 2024-02-19 hep-ph cs.LG hep-ex 79%

A Language Model for Particle Tracking

Andris Huang, Yash Melkani, Paolo Calafiura, Alina Lazar, Daniel Thomas Murnane, Minh-Tuan Pham, Xiangyang Ju

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.LG

Comments 7 pages, 3 figures, A Proceeding of the Connecting the Dots Workshop (CTD 2023)

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.09124 2024-02-19 cs.CL 79%

Linearity of Relation Decoding in Transformer Language Models

Evan Hernandez, Arnab Sen Sharma, Tal Haklay, Kevin Meng, Martin Wattenberg, Jacob Andreas, Yonatan Belinkov, David Bau

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.00842 2024-01-23 q-bio.QM cs.LG 79%

ESM-NBR: fast and accurate nucleic acid-binding residue prediction via protein language model feature representation and multi-task learning

Wenwu Zeng, Dafeng Lv, Wenjuan Liu, Shaoliang Peng

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.05669 2024-01-12 cs.CL 79%

ConcEPT: Concept-Enhanced Pre-Training for Language Models

Xintao Wang, Zhouhong Gu, Jiaqing Liang, Dakuan Lu, Yanghua Xiao, Wei Wang

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

Comments 12pages. Work completed in 2023.01

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.12976 2023-12-21 cs.CL 79%

Evaluating the Ripple Effects of Knowledge Editing in Language Models

Roi Cohen, Eden Biran, Ori Yoran, Amir Globerson, Mor Geva

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

Comments Accepted for publication in Transactions of the Association for Computational Linguistics (TACL), 2024. Author's final version

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.09670 2023-12-18 cs.CL cs.IR 79%

Probing Pretrained Language Models with Hierarchy Properties

Jesús Lovón-Melgarejo, Jose G. Moreno, Romaric Besançon, Olivier Ferret, Lynda Tamine

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

Comments Accepted at ECIR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.09926 2023-11-28 cs.AI 79%

Estimating Uncertainty in Multimodal Foundation Models using Public Internet Data

Shiladitya Dutta, Hongbo Wei, Lars van der Laan, Ahmed M. Alaa

专题命中 知识编辑与模型理解 :foundation model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.13508 2023-11-23 cs.SE cs.LG 79%

Naturalness of Attention: Revisiting Attention in Code Language Models

Mootez Saad, Tushar Sharma

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.LG

Comments Accepted at ICSE-NIER (2024) track

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.03788 2023-11-08 cs.CL 79%

Language Representation Projection: Can We Transfer Factual Knowledge across Languages in Multilingual Language Models?

Shaoyang Xu, Junzhuo Li, Deyi Xiong

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

Comments Accepted by EMNLP 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.18930 2023-10-31 cs.CL 79%

Retrofitting Light-weight Language Models for Emotions using Supervised Contrastive Learning

Sapan Shah, Sreedhar Reddy, Pushpak Bhattacharyya

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

Comments EMNLP 2023 Camera Ready Version

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.17271 2023-10-27 cs.CL 79%

Understanding the Role of Input Token Characters in Language Models: How Does Information Loss Affect Performance?

Ahmed Alajrami, Katerina Margatina, Nikolaos Aletras

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

Comments To appear at EMNLP 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.16484 2023-10-26 cs.CL 79%

Subspace Chronicles: How Linguistic Information Emerges, Shifts and Interacts during Language Model Training

Max Müller-Eberstein, Rob van der Goot, Barbara Plank, Ivan Titov

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

Comments Accepted at EMNLP 2023 (Findings)

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.15420 2023-10-25 cs.CL 79%

Let the Pretrained Language Models "Imagine" for Short Texts Topic Modeling

Pritom Saha Akash, Jie Huang, Kevin Chen-Chuan Chang

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.15109 2023-10-24 cs.CL 79%

GRENADE: Graph-Centric Language Model for Self-Supervised Representation Learning on Text-Attributed Graphs

Yichuan Li, Kaize Ding, Kyumin Lee

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

Comments Findings of EMNLP 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.12751 2023-10-20 cs.CL 79%

Character-level Chinese Backpack Language Models

Hao Sun, John Hewitt

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

Comments BlackboxNLP 2023 Camera-Ready

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.07929 2023-10-13 cs.CL 79%

Crosslingual Structural Priming and the Pre-Training Dynamics of Bilingual Language Models

Catherine Arnett, Tyler A. Chang, James A. Michaelov, Benjamin K. Bergen

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

Comments Extended abstract accepted to the 3rd Multilingual Representation Learning workshop at EMNLP 2023

详情

展开后加载摘要…

URL PDF HTML 收藏