arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12228 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12228 篇

2503.06330 2025-03-11 cs.CL cond-mat.stat-mech cs.AI 76%

States of LLM-generated Texts and Phase Transitions between them

Nikolay Mikhaylovskiy

专题命中 其他LLM :LLM(title);分类 cs.CL、cs.AI

Comments Published as a conference paper at MathAI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.11596 2025-02-18 cs.LG cs.AI 76%

LLM Embeddings for Deep Learning on Tabular Data

Boshko Koloski, Andrei Margeloiu, Xiangjian Jiang, Blaž Škrlj, Nikola Simidjievski, Mateja Jamnik

专题命中 其他LLM :LLM(title);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.18924 2025-02-03 cs.AI cs.CL cs.MA 76%

Language Games as the Pathway to Artificial Superhuman Intelligence

Ying Wen, Ziyu Wan, Shao Zhang

专题命中 其他LLM :large language model(abstract,comments);language model(abstract,comments);分类 cs.CL、cs.AI

Comments This position paper argues that language games provide robust mechanism for achieving superhuman intelligence in large language models

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.11852 2025-01-22 cs.CL cs.CR cs.LG 76%

Cross-Entropy Attacks to Language Models via Rare Event Simulation

Mingze Ni, Yongshun Gong, Wei Liu

专题命中 其他LLM :language model(title);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.14948 2025-01-14 cs.LG cs.AI cs.CV 76%

Effective Backdoor Mitigation in Vision-Language Models Depends on the Pre-training Objective

Sahil Verma, Gantavya Bhatt, Avi Schwarzschild, Soumye Singhal, Arnav Mohanty Das, Chirag Shah, John P Dickerson, Pin-Yu Chen, Jeff Bilmes

专题命中 其他LLM :language model(title);分类 cs.AI、cs.LG

Comments Accepted at TMLR (https://openreview.net/forum?id=Conma3qnaT)

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.18491 2025-01-07 cs.LG cs.CL cs.CR 76%

Publicly-Detectable Watermarking for Language Models

Jaiden Fairoze, Sanjam Garg, Somesh Jha, Saeed Mahloujifar, Mohammad Mahmoody, Mingyuan Wang

专题命中 其他LLM :language model(title);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.01303 2025-01-03 cs.CL cs.AI 76%

Citations and Trust in LLM Generated Responses

Yifan Ding, Matthew Facciani, Amrit Poudel, Ellen Joyce, Salvador Aguinaga, Balaji Veeramani, Sanmitra Bhattacharya, Tim Weninger

专题命中 其他LLM :LLM(title);分类 cs.CL、cs.AI

Comments Accepted to AAAI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.13041 2024-12-18 cs.CL cs.LG 76%

Harnessing Event Sensory Data for Error Pattern Prediction in Vehicles: A Language Model Approach

Hugo Math, Rainer Lienhart, Robin Schön

专题命中 其他LLM :language model(title);分类 cs.CL、cs.LG

Comments 10 pages, 8 figures, accepted to AAAI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.08504 2024-11-15 cs.CL cs.AI 76%

Towards Objective and Unbiased Decision Assessments with LLM-Enhanced Hierarchical Attention Networks

Junhua Liu, Kwan Hui Lim, Roy Ka-Wei Lee

专题命中 其他LLM :LLM(title);分类 cs.CL、cs.AI

Comments Source code is available at: https://github.com/junhua/bgm-han

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.13728 2024-10-25 cs.CL cs.LG stat.ML 76%

Rule Extrapolation in Language Models: A Study of Compositional Generalization on OOD Prompts

Anna Mészáros, Szilvia Ujváry, Wieland Brendel, Patrik Reizinger, Ferenc Huszár

专题命中 其他LLM :language model(title);分类 cs.CL、cs.LG

Comments Accepted as a spotlight poster at NeurIPS2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.14767 2024-10-01 cs.CL cs.AI 76%

I Need Help! Evaluating LLM's Ability to Ask for Users' Support: A Case Study on Text-to-SQL Generation

Cheng-Kuang Wu, Zhi Rui Tam, Chao-Chung Wu, Chieh-Yen Lin, Hung-yi Lee, Yun-Nung Chen

专题命中 其他LLM :LLM(title);分类 cs.CL、cs.AI

Comments Accepted by EMNLP 2024 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.13852 2024-09-24 cs.CL cs.AI 76%

Do language models practice what they preach? Examining language ideologies about gendered language reform encoded in LLMs

Julia Watson, Sophia Lee, Barend Beekhuizen, Suzanne Stevenson

专题命中 其他LLM :language model(title);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.11232 2024-09-23 cs.CL cond-mat.dis-nn cs.AI 76%

Fast Analysis of the OpenAI O1-Preview Model in Solving Random K-SAT Problem: Does the LLM Solve the Problem Itself or Call an External SAT Solver?

Raffaele Marino

专题命中 其他LLM :LLM(title);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.15052 2024-07-02 cs.LG cs.AI 76%

Revisiting MoE and Dense Speed-Accuracy Comparisons for LLM Training

Xianzhi Du, Tom Gunter, Xiang Kong, Mark Lee, Zirui Wang, Aonan Zhang, Nan Du, Ruoming Pang

专题命中 其他LLM :LLM(title);分类 cs.AI、cs.LG

Comments 8 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.13660 2024-06-21 cs.CL cs.AI 76%

Towards Minimal Targeted Updates of Language Models with Targeted Negative Training

Lily H. Zhang, Rajesh Ranganath, Arya Tafvizi

专题命中 其他LLM :language model(title);分类 cs.CL、cs.AI

Comments Published in Transactions of Machine Learning Research

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.12216 2024-06-19 cs.CL cs.AI 76%

Is persona enough for personality? Using ChatGPT to reconstruct an agent's latent personality from simple descriptions

Yongyi Ji, Zhisheng Tang, Mayank Kejriwal

专题命中 其他LLM :large language model(abstract,comments);language model(abstract,comments);分类 cs.CL、cs.AI

Comments Accepted to the ICML 2024 Workshop on Large Language Models and Cognition

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.19192 2024-05-01 cs.CL cs.AI 76%

Mix of Experts Language Model for Named Entity Recognition

Xinwei Chen, Kun Li, Tianyou Song, Jiangjian Guo

专题命中 其他LLM :language model(title);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.13474 2024-04-23 cs.RO cs.AI cs.CV cs.LG 76%

Composing Pre-Trained Object-Centric Representations for Robotics From "What" and "Where" Foundation Models

Junyao Shi, Jianing Qian, Yecheng Jason Ma, Dinesh Jayaraman

专题命中 其他LLM :foundation model(title);分类 cs.AI、cs.LG

Comments ICRA 2024. Project website: https://sites.google.com/view/pocr

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.00824 2024-04-18 cs.CL cs.AI 76%

Information Flow Routes: Automatically Interpreting Language Models at Scale

Javier Ferrando, Elena Voita

专题命中 其他LLM :language model(title);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.05502 2024-04-09 cs.CL cs.AI 76%

PetKaz at SemEval-2024 Task 3: Advancing Emotion Classification with an LLM for Emotion-Cause Pair Extraction in Conversations

Roman Kazakov, Kseniia Petukhova, Ekaterina Kochmar

专题命中 其他LLM :LLM(title);分类 cs.CL、cs.AI

Comments 8 pages, 7 figures, 2 tables, to be published in the Proceedings of the 18th International Workshop on Semantic Evaluation (SemEval-2024), for associated code, see https://github.com/sachertort/petkaz-semeval-ecac

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.05483 2024-04-09 cs.CL cs.AI 76%

PetKaz at SemEval-2024 Task 8: Can Linguistics Capture the Specifics of LLM-generated Text?

Kseniia Petukhova, Roman Kazakov, Ekaterina Kochmar

专题命中 其他LLM :LLM(title);分类 cs.CL、cs.AI

Comments 8 pages, 3 figures, 5 tables, to be published in the Proceedings of the 18th International Workshop on Semantic Evaluation (SemEval-2024), for associated code, see https://github.com/sachertort/petkaz-semeval-m4

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.12809 2024-03-20 cs.CL cs.AI 76%

Comparing Explanation Faithfulness between Multilingual and Monolingual Fine-tuned Language Models

Zhixue Zhao, Nikolaos Aletras

专题命中 其他LLM :language model(title);分类 cs.CL、cs.AI

Comments Accepted at NAACL 2024 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.09706 2023-11-17 cs.AI cs.HC cs.LG 76%

Towards Autonomous Hypothesis Verification via Language Models with Minimal Guidance

Shiro Takagi, Ryutaro Yamauchi, Wataru Kumagai

专题命中 其他LLM :language model(title);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.04707 2023-06-09 cs.CL cs.AI 76%

Improving Open Language Models by Learning from Organic Interactions

Jing Xu, Da Ju, Joshua Lane, Mojtaba Komeili, Eric Michael Smith, Megan Ung, Morteza Behrooz, William Ngan, Rashel Moritz, Sainbayar Sukhbaatar, Y-Lan Boureau, Jason Weston, Kurt Shuster

专题命中 其他LLM :language model(title);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.15206 2023-05-17 cs.CL cs.AI 76%

What Makes Pre-trained Language Models Better Zero-shot Learners?

Jinghui Lu, Dongsheng Zhu, Weidong Han, Rui Zhao, Brian Mac Namee, Fei Tan

专题命中 其他LLM :language model(title);分类 cs.CL、cs.AI

Comments Accepted to ACL2023 main conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.06159 2023-05-11 cs.CL cs.AI 76%

A Review of Vision-Language Models and their Performance on the Hateful Memes Challenge

Bryan Zhao, Andrew Zhang, Blake Watson, Gillian Kearney, Isaac Dale

专题命中 其他LLM :language model(title);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.04790 2023-02-10 cs.CL cs.AI cs.IR 76%

Massively Multilingual Language Models for Cross Lingual Fact Extraction from Low Resource Indian Languages

Bhavyajeet Singh, Pavan Kandru, Anubhav Sharma, Vasudeva Varma

专题命中 其他LLM :language model(title);分类 cs.CL、cs.AI

Comments 5 pages, 2 page Apendix, 3 figures, accepted at 19th International Conference on Natural Language Processing

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.12565 2022-10-25 cs.CL cs.LG 76%

A Visual Tour Of Current Challenges In Multimodal Language Models

Shashank Sonkar, Naiming Liu, Richard G. Baraniuk

专题命中 其他LLM :language model(title);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.01691 2022-08-17 cs.RO cs.CL cs.LG 76%

Do As I Can, Not As I Say: Grounding Language in Robotic Affordances

Michael Ahn, Anthony Brohan, Noah Brown, Yevgen Chebotar, Omar Cortes, Byron David, Chelsea Finn, Chuyuan Fu, Keerthana Gopalakrishnan, Karol Hausman, Alex Herzog, Daniel Ho, Jasmine Hsu, Julian Ibarz, Brian Ichter, Alex Irpan, Eric Jang, Rosario Jauregui Ruano, Kyle Jeffrey, Sally Jesmonth, Nikhil J Joshi, Ryan Julian, Dmitry Kalashnikov, Yuheng Kuang, Kuang-Huei Lee, Sergey Levine, Yao Lu, Linda Luu, Carolina Parada, Peter Pastor, Jornell Quiambao, Kanishka Rao, Jarek Rettinghouse, Diego Reyes, Pierre Sermanet, Nicolas Sievers, Clayton Tan, Alexander Toshev, Vincent Vanhoucke, Fei Xia, Ted Xiao, Peng Xu, Sichun Xu, Mengyuan Yan, Andy Zeng

专题命中 其他LLM :language model(abstract,comments);large language model(abstract);分类 cs.CL、cs.LG;prompting(comments)

Comments See website at https://say-can.github.io/ V1. Initial Upload. V2. Added PaLM results. Added study about new capabilities (drawer manipulation, chain of thought prompting, multilingual instructions). Added an ablation study of language model size. Added an open-source version of \algname on a simulated tabletop environment. Improved readability

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.13947 2022-07-05 cs.LG cs.CL 76%

Long Range Language Modeling via Gated State Spaces

Harsh Mehta, Ankit Gupta, Ashok Cutkosky, Behnam Neyshabur

专题命中 其他LLM :language model(title);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏