arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 7596 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 7596 篇

2303.13472 2024-01-23 cs.CV cs.AI 57%

Promptable Game Models: Text-Guided Game Simulation via Masked Diffusion Models

Willi Menapace, Aliaksandr Siarohin, Stéphane Lathuilière, Panos Achlioptas, Vladislav Golyanik, Sergey Tulyakov, Elisa Ricci

专题命中 知识编辑与模型理解 :prompting(abstract);分类 cs.AI

Comments ACM Transactions on Graphics \c{opyright} Copyright is held by the owner/author(s) 2023. This is the author's version of the work. It is posted here for your personal use. Not for redistribution. The definitive Version of Record was published in ACM Transactions on Graphics, http://dx.doi.org/10.1145/3635705

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.09172 2024-01-19 cs.CV cs.LG 57%

Hyperbolic Image-Text Representations

Karan Desai, Maximilian Nickel, Tanmay Rajpurohit, Justin Johnson, Ramakrishna Vedantam

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.LG

Comments ICML 2023 (v3: Add link to code in abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.03642 2024-01-17 cs.CL cs.DL 57%

A Content-Based Novelty Measure for Scholarly Publications: A Proof of Concept

Haining Wang

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

Comments Accepted for publication in the proceedings of iConference2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.05378 2024-01-17 cs.CL cs.SI 57%

Transcending the Attention Paradigm: Representation Learning from Geospatial Social Media Data

Nick DiSanto, Anthony Corso, Benjamin Sanders, Gavin Harding

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.06890 2024-01-17 cs.LG 57%

An Axiomatic Approach to Model-Agnostic Concept Explanations

Zhili Feng, Michal Moshkovitz, Dotan Di Castro, J. Zico Kolter

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.03379 2024-01-17 cs.LG math.OC 57%

Data augmentation for machine learning of chemical process flowsheets

Lukas Schulze Balhorn, Edwin Hirtreiter, Lynn Luderer, Artur M. Schweidtmann

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.LG

Comments Submitted to PROCEEDINGS OF THE 33rd European Symposium on Computer Aided Process Engineering (ESCAPE33), June 18-21, 2023, Athens, Greece

Journal ref Computer Aided Chemical Engineering Volume 52, 2023, Pages 2011-2016

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.01592 2024-01-10 cs.CL 57%

Expand BERT Representation with Visual Information via Grounded Language Learning with Multimodal Partial Alignment

Cong-Duy Nguyen, The-Anh Vu-Le, Thong Nguyen, Tho Quan, Luu Anh Tuan

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.01667 2024-01-04 cs.CL 57%

MLPs Compass: What is learned when MLPs are combined with PLMs?

Li Zhou, Wenyu Chen, Yong Cao, Dingyi Zeng, Wanlong Liu, Hong Qu

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

Comments Accepted by ICASSP 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.14966 2023-12-27 cs.CL 57%

Dynamic Syntax Mapping: A New Approach to Unsupervised Syntax Parsing

Buvarp Gohsh, Woods Ali, Anders Michael

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.13198 2023-12-21 cs.CL 57%

Journey to the Center of the Knowledge Neurons: Discoveries of Language-Independent Knowledge Neurons and Degenerate Knowledge Neurons

Yuheng Chen, Pengfei Cao, Yubo Chen, Kang Liu, Jun Zhao

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

Comments Accepted in the 38th AAAI Conference on Artificial Intelligence (AAAI 2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.09651 2023-12-18 cs.SD cs.CR cs.LG eess.AS 57%

What to Remember: Self-Adaptive Continual Learning for Audio Deepfake Detection

Xiaohui Zhang, Jiangyan Yi, Chenglong Wang, Chuyuan Zhang, Siding Zeng, Jianhua Tao

专题命中 知识编辑与模型理解 :prompting(abstract);分类 cs.LG

Comments Accepted by the main track The 38th Annual AAAI Conference on Artificial Intelligence (AAAI 2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.04281 2023-12-13 cs.CV cs.LG 57%

Local Spatiotemporal Representation Learning for Longitudinally-consistent Neuroimage Analysis

Mengwei Ren, Neel Dey, Martin A. Styner, Kelly Botteron, Guido Gerig

专题命中 知识编辑与模型理解 :pretraining(abstract);分类 cs.LG

Comments Accepted at NeurIPS 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.05092 2023-12-11 cs.SE cs.LG 57%

INSPECT: Intrinsic and Systematic Probing Evaluation for Code Transformers

Anjan Karmakar, Romain Robbes

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.LG

Comments Accepted to IEEE Transactions on Software Engineering. Extension of our previous paper "What do pre-trained code models know about code?" (ASE 2021, arXiv:2108.11308). 21 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.17259 2023-12-05 cs.LG cs.CY 57%

SoUnD Framework: Analyzing (So)cial Representation in (Un)structured (D)ata

Mark Díaz, Sunipa Dev, Emily Reif, Emily Denton, Vinodkumar Prabhakaran

专题命中 知识编辑与模型理解 :foundation model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.00680 2023-12-04 cs.CL 57%

Contextualized word senses: from attention to compositionality

Pablo Gamallo

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

Journal ref Linguistics Vanguard, 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.11603 2023-11-23 cs.CL 57%

Representation Projection Invariance Mitigates Representation Collapse

Anastasia Razdaibiedina, Ashish Khetan, Zohar Karnin, Daniel Khashabi, Vishaal Kapoor, Vivek Madan

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

Comments 41 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.07864 2023-11-15 cs.LG cs.CV 57%

Probing clustering in neural network representations

Thao Nguyen, Simon Kornblith

专题命中 知识编辑与模型理解 :pretraining(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.16607 2023-11-14 cs.CL 57%

On the Interplay between Fairness and Explainability

Stephanie Brandl, Emanuele Bugliarello, Ilias Chalkidis

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

Comments 15 pages (incl Appendix), 4 figures, 8 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.16330 2023-11-10 physics.chem-ph cs.LG 57%

Prompt Engineering for Transformer-based Chemical Similarity Search Identifies Structurally Distinct Functional Analogues

Clayton W. Kosonocky, Aaron L. Feller, Claus O. Wilke, Andrew D. Ellington

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.03320 2023-11-07 cs.CL 57%

Tackling Concept Shift in Text Classification using Entailment-style Modeling

Sumegh Roychowdhury, Karan Gupta, Siva Rajesh Kasa, Prasanna Srinivasa Murthy, Alok Chandra

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

Journal ref NeurIPS 2023 - Workshop on Distribution Shifts

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.02294 2023-11-07 cs.CL cs.CY 57%

LLMs grasp morality in concept

Mark Pock, Andre Ye, Jared Moore

专题命中 知识编辑与模型理解 :LLM(abstract);分类 cs.CL

Comments Presented at NeurIPS 2023 Moral Pyschology and Moral Philosophy workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.01173 2023-11-03 cs.CL 57%

CRUSH4SQL: Collective Retrieval Using Schema Hallucination For Text2SQL

Mayank Kothyari, Dhruva Dhingra, Sunita Sarawagi, Soumen Chakrabarti

专题命中 知识编辑与模型理解 :LLM(abstract);分类 cs.CL

Comments To appear at EMNLP 2023 (Main)

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.15908 2023-11-01 cs.CL 57%

Response Generation in Longitudinal Dialogues: Which Knowledge Representation Helps?

Seyed Mahed Mousavi, Simone Caldarella, Giuseppe Riccardi

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.19127 2023-10-31 cs.CL 57%

Unified Representation for Non-compositional and Compositional Expressions

Ziheng Zeng, Suma Bhat

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

Comments This work is accepted to EMNLP 2023 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.18862 2023-10-31 cs.CL 57%

Counterfactually Probing Language Identity in Multilingual Models

Anirudh Srinivasan, Venkata S Govindarajan, Kyle Mahowald

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

Comments 12 pages, 5 figures, MRL Workshop @ EMNLP 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.18155 2023-10-30 cs.CL 57%

Elevating Code-mixed Text Handling through Auditory Information of Words

Mamta, Zishan Ahmad, Asif Ekbal

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

Comments Accepted to EMNLP 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.17121 2023-10-27 cs.CL 57%

Test-time Augmentation for Factual Probing

Go Kamoda, Benjamin Heinzerling, Keisuke Sakaguchi, Kentaro Inui

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

Comments 12 pages, 4 figures, accepted to EMNLP 2023 Findings (short paper)

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.16303 2023-10-26 cs.CL cs.IR 57%

URL-BERT: Training Webpage Representations via Social Media Engagements

Ayesha Qamar, Chetan Verma, Ahmed El-Kishky, Sumit Binnani, Sneha Mehta, Taylor Berg-Kirkpatrick

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.15316 2023-10-25 cs.CL 57%

Probing Representations for Document-level Event Extraction

Barry Wang, Xinya Du, Claire Cardie

专题命中 知识编辑与模型理解 :LLM(abstract);分类 cs.CL

Comments To appear in EMNLP 2023 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.14981 2023-10-24 cs.CL 57%

Fidelity-Enriched Contrastive Search: Reconciling the Faithfulness-Diversity Trade-Off in Text Generation

Wei-Lin Chen, Cheng-Kuang Wu, Hsin-Hsi Chen, Chung-Chi Chen

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

Comments Accepted as a short paper at EMNLP 2023

详情

展开后加载摘要…

URL PDF HTML 收藏