arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 7584 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 7584 篇

2504.17261 2025-04-25 cs.LG cs.AI 62%

Symbolic Representation for Any-to-Any Generative Tasks

Jiaqi Chen, Xiaoye Zhu, Yue Wang, Tianyang Liu, Xinhui Chen, Ying Chen, Chak Tou Leong, Yifei Ke, Joseph Liu, Yiwen Yuan, Julian McAuley, Li-jia Li

机构 * Stanford University(斯坦福大学) Fellou AI(Fellou人工智能) Fudan University(复旦大学) South China University of Technology(华南理工大学) Cornell University(康奈尔大学) University of California San Diego(加州大学圣地亚哥分校) Wuhan University(武汉大学) University of Illinois at Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Hong Kong Polytechnic University(香港理工大学) University of Southern California(南加州大学) Carnegie Mellon University(卡内基梅隆大学) LiveX AI(LiveX人工智能)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.12022 2025-04-21 cs.CL cs.AI 62%

Understanding Epistemic Language with a Language-augmented Bayesian Theory of Mind

Lance Ying, Tan Zhi-Xuan, Lionel Wong, Vikash Mansinghka, Joshua B. Tenenbaum

专题命中 知识编辑与模型理解 :LLM(abstract);分类 cs.CL、cs.AI

Comments 23 pages; Published at the Transactions of the Association for Computational Linguistics (TACL); Presented at NAACL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.04011 2025-04-18 eess.SY cs.AI cs.LG cs.SY 62%

Predicting and Publishing Accurate Imbalance Prices Using Monte Carlo Tree Search

Fabio Pavirani, Jonas Van Gompel, Seyed Soroush Karimi Madahi, Bert Claessens, Chris Develder

专题命中 知识编辑与模型理解 :prompting(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.09704 2025-04-15 cs.LG cs.AI 62%

Transformer-Based Representation Learning for Robust Gene Expression Modeling and Cancer Prognosis

Shuai Jiang, Saeed Hassanpour

专题命中 知识编辑与模型理解 :pretraining(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.06038 2025-04-11 cs.LG cs.AI cs.CV 62%

Understanding Contrastive Representation Learning from Positive Unlabeled (PU) Data

Anish Acharya, Li Jing, Bhargav Bhushanam, Dhruv Choudhary, Michael Rabbat, Sujay Sanghavi, Inderjit S Dhillon

专题命中 知识编辑与模型理解 :SFT(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.03857 2025-04-08 cs.CV cs.AI cs.LG 62%

Can ChatGPT Learn My Life From a Week of First-Person Video?

Keegan Harris

专题命中 知识编辑与模型理解 :foundation model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.04315 2025-04-03 cs.CL cs.LG 62%

Calibrating Expressions of Certainty

Peiqi Wang, Barbara D. Lam, Yingcheng Liu, Ameneh Asgari-Targhi, Rameswar Panda, William M. Wells, Tina Kapur, Polina Golland

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.LG

Comments International Conference on Learning Representations (ICLR), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.24228 2025-04-01 cs.AI cs.CL cs.MA 62%

PAARS: Persona Aligned Agentic Retail Shoppers

Saab Mansour, Leonardo Perelli, Lorenzo Mainetti, George Davidson, Stefano D'Amato

专题命中 知识编辑与模型理解 :LLM(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.23819 2025-04-01 cs.LG cs.AI cs.CV 62%

Conformal uncertainty quantification to evaluate predictive fairness of foundation AI model for skin lesion classes across patient demographics

Swarnava Bhattacharyya, Umapada Pal, Tapabrata Chakraborti

专题命中 知识编辑与模型理解 :foundation model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.23615 2025-04-01 cs.AI cs.LG 62%

An Organizationally-Oriented Approach to Enhancing Explainability and Control in Multi-Agent Reinforcement Learning

Julien Soulé, Jean-Paul Jamont, Michel Occello, Louis-Marie Traonouez, Paul Théron

专题命中 知识编辑与模型理解 :post-training(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19900 2025-03-26 cs.CV cs.AI cs.CL 62%

CAFe: Unifying Representation and Generation with Contrastive-Autoregressive Finetuning

Hao Yu, Zhuokai Zhao, Shen Yan, Lukasz Korycki, Jianyu Wang, Baosheng He, Jiayi Liu, Lizhu Zhang, Xiangjun Fan, Hanchao Yu

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.19186 2025-03-26 cs.CV cs.CL cs.LG 62%

MetaToken: Detecting Hallucination in Image Descriptions by Meta Classification

Laura Fieback, Jakob Spiegelberg, Hanno Gottschalk

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.18551 2025-03-25 cs.LG cs.AI 62%

Discriminative protein sequence modelling with Latent Space Diffusion

Eoin Quinn, Ghassene Jebali, Maxime Seince, Oliver Bent

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.20899 2025-03-19 cs.AI cs.CL 62%

Faithful and Plausible Natural Language Explanations for Image Classification: A Pipeline Approach

Adam Wojciechowski, Mateusz Lango, Ondrej Dusek

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI

Comments Findings of EMNLP 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.07658 2025-03-18 cs.CV cs.AI cs.LG 62%

TraSCE: Trajectory Steering for Concept Erasure

Anubhav Jain, Yuya Kobayashi, Takashi Shibuya, Yuhta Takida, Nasir Memon, Julian Togelius, Yuki Mitsufuji

专题命中 知识编辑与模型理解 :prompting(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.16073 2025-03-17 cs.LG cs.CL 62%

Challenging Assumptions in Learning Generic Text Style Embeddings

Phil Ostheimer, Marius Kloft, Sophie Fellenz

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.LG

Comments Proceedings of the Sixth Workshop on Insights from Negative Results in NLP at NAACL-HLT

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.08764 2025-03-13 q-bio.BM cs.AI cs.LG 62%

Towards Interpretable Protein Structure Prediction with Sparse Autoencoders

Nithin Parsan, David J. Yang, John J. Yang

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI、cs.LG

Comments Published at the GEMBio ICLR 2025 Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06202 2025-03-11 cs.AI cs.LG 62%

Breaking Free from MMI: A New Frontier in Rationalization by Probing Input Utilization

Wei Liu, Zhiying Deng, Zhongyu Niu, Jun Wang, Haozhao Wang, Zhigang Zeng, Ruixuan Li

专题命中 知识编辑与模型理解 :LLM(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.11283 2025-03-10 cs.CL cs.AI 62%

Zero-resource Hallucination Detection for Text Generation via Graph-based Contextual Knowledge Triples Modeling

Xinyue Fang, Zhen Huang, Zhiliang Tian, Minghui Fang, Ziyi Pan, Quntian Fang, Zhihua Wen, Hengyue Pan, Dongsheng Li

专题命中 知识编辑与模型理解 :LLM(abstract);分类 cs.CL、cs.AI

Comments Accepted by AAAI25

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.04667 2025-03-07 cs.CL cs.IT cs.LG math.IT 62%

An Information-theoretic Multi-task Representation Learning Framework for Natural Language Understanding

Dou Hu, Lingwei Wei, Wei Zhou, Songlin Hu

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.LG

Comments 11 pages, accepted to AAAI 2025 (main conference), the code is available at https://github.com/zerohd4869/InfoMTL

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.19302 2025-03-05 cs.IR cs.AI cs.LG 62%

TEARS: Textual Representations for Scrutable Recommendations

Emiliano Penaloza, Olivier Gouvert, Haolun Wu, Laurent Charlin

专题命中 知识编辑与模型理解 :LLM(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.20613 2025-03-03 cs.CL cs.AI 62%

Continuous Adversarial Text Representation Learning for Affective Recognition

Seungah Son, Andrez Saurez, Dongsoo Har

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI

Comments 6 pages, 3 figures, The 7th International Conference on Artificial Intelligence in Information and Communication (ICAIIC 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.17125 2025-02-25 cs.CL cs.AI 62%

LettuceDetect: A Hallucination Detection Framework for RAG Applications

Ádám Kovács, Gábor Recski

专题命中 知识编辑与模型理解 :LLM(abstract);分类 cs.CL、cs.AI

Comments 6 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.15079 2025-02-24 cs.CV cs.AI cs.CL 62%

Can Hallucination Correction Improve Video-Language Alignment?

Lingjun Zhao, Mingyang Xie, Paola Cascante-Bonilla, Hal Daumé, Kwonjoon Lee

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.11301 2025-02-24 cs.CL cs.AI 62%

Question-to-Question Retrieval for Hallucination-Free Knowledge Access: An Approach for Wikipedia and Wikidata Question Answering

Santhosh Thottingal

专题命中 知识编辑与模型理解 :LLM(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.14486 2025-02-21 cs.CR cs.AI cs.CL 62%

How Jailbreak Defenses Work and Ensemble? A Mechanistic Investigation

Zhuohang Long, Siyuan Wang, Shujun Liu, Yuhang Lai, Xuanjing Huang, Zhongyu Wei

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.14394 2025-02-13 cs.AI cs.CL cs.CV 62%

A Multimodal Automated Interpretability Agent

Tamar Rott Shaham, Sarah Schwettmann, Franklin Wang, Achyuta Rajaram, Evan Hernandez, Jacob Andreas, Antonio Torralba

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI

Comments 25 pages, 13 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.07586 2025-02-12 cs.CL cs.AI 62%

We Can't Understand AI Using our Existing Vocabulary

John Hewitt, Robert Geirhos, Been Kim

专题命中 知识编辑与模型理解 :LLM(abstract);分类 cs.CL、cs.AI

Comments Position paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.06852 2025-02-12 cs.LG cs.AI 62%

EAP-GP: Mitigating Saturation Effect in Gradient-based Automated Circuit Identification

Lin Zhang, Wenshuo Dong, Zhuoran Zhang, Shu Yang, Lijie Hu, Ninghao Liu, Pan Zhou, Di Wang

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.11355 2025-02-12 cs.CL cs.CY cs.LG 62%

A Practical Method for Generating String Counterfactuals

Matan Avitan, Ryan Cotterell, Yoav Goldberg, Shauli Ravfogel

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.LG

Comments Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏