arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12193 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12193 篇

2504.08985 2025-04-29 cs.HC cs.AI 79%

Learning from Elders: Making an LLM-powered Chatbot for Retirement Communities more Accessible through User-centered Design

Luna Xingyu Li, Ray-yuan Chung, Feng Chen, Wenyu Zeng, Yein Jeon, Oleg Zaslavsky

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI

Comments Accepted as Research talk for Considering Cultural and Linguistic Diversity in AI Applications workshop at CALD-AI@ASIS&T 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.20268 2025-04-29 cs.LG 79%

Centaur: a foundation model of human cognition

Marcel Binz, Elif Akata, Matthias Bethge, Franziska Brändle, Fred Callaway, Julian Coda-Forno, Peter Dayan, Can Demircan, Maria K. Eckstein, Noémi Éltető, Thomas L. Griffiths, Susanne Haridi, Akshay K. Jagadish, Li Ji-An, Alexander Kipnis, Sreejan Kumar, Tobias Ludwig, Marvin Mathony, Marcelo Mattar, Alireza Modirshanechi, Surabhi S. Nath, Joshua C. Peterson, Milena Rmus, Evan M. Russek, Tankred Saanum, Johannes A. Schubert, Luca M. Schulze Buschoff, Nishad Singhi, Xin Sui, Mirko Thalmann, Fabian Theis, Vuong Truong, Vishaal Udandarao, Konstantinos Voudouris, Robert Wilson, Kristin Witte, Shuchen Wu, Dirk Wulff, Huadong Xiong, Eric Schulz

机构 * Helmholtz Munich(慕尼黑海德堡医学研究院) University of Tuebingen(图宾根大学) University of Oxford(牛津大学) New York University(纽约大学) Max Planck Institute for Biological Cybernetics(生物控制研究所) Google DeepMind(谷歌DeepMind) Princeton University(普林斯顿大学) University of California San Diego(圣地亚哥大学) Boston University(波士顿大学) Georgia Institute of Technology(佐治亚理工学院) University of Basel(巴塞尔大学) Max Planck Institute for Human Development(人类发展研究所) Max Planck School of Cognition(认知研究所) TU Darmstadt(德累斯顿技术大学) University of Cambridge(剑桥大学)

专题命中 其他LLM :foundation model(title);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.07149 2025-04-29 cs.CV cs.LG 79%

Towards Interpreting Visual Information Processing in Vision-Language Models

Clement Neo, Luke Ong, Philip Torr, Mor Geva, David Krueger, Fazl Barez

机构 * Nanyang Technological University(南洋理工大学) University of Oxford(牛津大学) Tel Aviv University(特拉维夫大学) MILA(蒙特利尔人工智能研究院) ERA-Krueger AI Safety Lab(ERA-Krueger人工智能安全实验室)

专题命中 其他LLM :language model(title,abstract);分类 cs.LG

Comments Published at ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.17083 2025-04-25 cs.CL 79%

How Individual Traits and Language Styles Shape Preferences In Open-ended User-LLM Interaction: A Preliminary Study

Rendi Chevi, Kentaro Inui, Thamar Solorio, Alham Fikri Aji

机构 * MBZUAI Abu Dhabi UAE(阿布扎赫 MBZUAI)

专题命中 其他LLM :LLM(title,abstract);分类 cs.CL

Comments Accepted at GenAICHI 2025 @ ACM CHI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.14335 2025-04-22 cs.CV cs.AI 79%

Visual Prompting for One-shot Controllable Video Editing without Inversion

Zhengbo Zhang, Yuxi Zhou, Duo Peng, Joo-Hwee Lim, Zhigang Tu, De Wen Soh, Lin Geng Foo

机构 * Singapore University of Technology and Design(新加坡科技设计大学) Wuhan University(武汉大学) Institute for Infocomm Research, Agency for Science, Technology and Research, Singapore(新加坡资讯与通信研究院)

专题命中 其他LLM :prompting(title,abstract);分类 cs.AI

Comments accepted by cvpr2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.12767 2025-04-18 cs.CL 79%

Out of Sight Out of Mind, Out of Sight Out of Mind: Measuring Bias in Language Models Against Overlooked Marginalized Groups in Regional Contexts

Fatma Elsafoury, David Hartmann

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.10391 2025-04-15 cs.CL 79%

LLM-driven Constrained Copy Generation through Iterative Refinement

Varun Vasudevan, Faezeh Akhavizadegan, Abhinav Prakash, Yokila Arora, Jason Cho, Tanya Mendiratta, Sushant Kumar, Kannan Achan

专题命中 其他LLM :LLM(title,abstract);分类 cs.CL

Comments 10 pages, 2 figures, 7 Tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.04917 2025-04-15 cs.CV cs.AI 79%

Avoid Wasted Annotation Costs in Open-set Active Learning with Pre-trained Vision-Language Model

Jaehyuk Heo, Pilsung Kang

专题命中 其他LLM :language model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19510 2025-03-26 cs.RO cs.AI cs.CV 79%

RoboFlamingo-Plus: Fusion of Depth and RGB Perception with Vision-Language Models for Enhanced Robotic Manipulation

Sheng Wang

专题命中 其他LLM :language model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.18556 2025-03-25 cs.CV cs.CL 79%

Instruction-Aligned Visual Attention for Mitigating Hallucinations in Large Vision-Language Models

Bin Li, Dehong Gao, Yeyuan Wang, Linbo Jin, Shanqing Yu, Xiaoyan Cai, Libin Yang

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

Comments Accepted by ICME2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.07304 2025-03-25 cs.CL 79%

We're Calling an Intervention: Exploring Fundamental Hurdles in Adapting Language Models to Nonstandard Text

Aarohi Srivastava, David Chiang

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

Comments Accepted for publication at W-NUT 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.10419 2025-03-14 cs.RO cs.AI 79%

HiFi-CS: Towards Open Vocabulary Visual Grounding For Robotic Grasping Using Vision-Language Models

Vineet Bhat, Prashanth Krishnamurthy, Ramesh Karri, Farshad Khorrami

专题命中 其他LLM :language model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06137 2025-03-11 cs.CL 79%

Evaluating Discourse Cohesion in Pre-trained Language Models

Jie He, Wanqiu Long, Deyi Xiong

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.04996 2025-03-10 cs.CL 79%

HieroLM: Egyptian Hieroglyph Recovery with Next Word Prediction Language Model

Xuheng Cai, Erica Zhang

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

Comments Accepted at LaTeCH-CLfL 2025 @ NAACL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.03474 2025-03-06 cs.CL 79%

Enhancing Spoken Discourse Modeling in Language Models Using Gestural Cues

Varsha Suresh, M. Hamza Mughal, Christian Theobalt, Vera Demberg

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.16007 2025-03-06 cs.AI 79%

Language Model Probabilities are Not Calibrated in Numeric Contexts

Charles Lovering, Michael Krumdick, Viet Dac Lai, Seth Ebner, Nilesh Kumar, Varshini Reddy, Rik Koncel-Kedziorski, Chris Tanner

专题命中 其他LLM :language model(title,abstract);分类 cs.AI

Comments 8 pages (main), 39 pages (references and appendix), in submission

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.03415 2025-03-05 cs.CL 79%

Surgical, Cheap, and Flexible: Mitigating False Refusal in Language Models via Single Vector Ablation

Xinpeng Wang, Chengzhi Hu, Paul Röttger, Barbara Plank

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

Comments ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.01724 2025-03-04 cs.CL 79%

Syntactic Learnability of Echo State Neural Language Models at Scale

Ryo Ueda, Tatsuki Kuribayashi, Shunsuke Kando, Kentaro Inui

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

Comments 10 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.19788 2025-03-04 cs.CL 79%

Exploring Adversarial Robustness in Classification tasks using DNA Language Models

Hyunwoo Yoo, Haebin Shin, Kaidi Xu, Gail Rosen

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.00224 2025-03-04 cs.CL 79%

À la recherche du sens perdu: your favourite LLM might have more to say than you can understand

K. O. T. Erziev

专题命中 其他LLM :LLM(title,abstract);分类 cs.CL

Comments 22 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.00152 2025-03-04 cs.LG cond-mat.mtrl-sci 79%

Invariant Tokenization of Crystalline Materials for Language Model Enabled Generation

Keqiang Yan, Xiner Li, Hongyi Ling, Kenna Ashen, Carl Edwards, Raymundo Arróyave, Marinka Zitnik, Heng Ji, Xiaofeng Qian, Xiaoning Qian, Shuiwang Ji

专题命中 其他LLM :language model(title,abstract);分类 cs.LG

Comments This paper has been accepted as a NeurIPS 2024 Poster

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.21245 2025-03-03 cs.LG 79%

TimesBERT: A BERT-Style Foundation Model for Time Series Understanding

Haoran Zhang, Yong Liu, Yunzhong Qiu, Haixuan Liu, Zhongyi Pei, Jianmin Wang, Mingsheng Long

专题命中 其他LLM :foundation model(title,abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.14860 2025-02-28 cs.LG 79%

Not All Language Model Features Are One-Dimensionally Linear

Joshua Engels, Eric J. Michaud, Isaac Liao, Wes Gurnee, Max Tegmark

专题命中 其他LLM :language model(title,abstract);分类 cs.LG

Comments Accepted to ICLR 2025. Code and data at https://github.com/JoshEngels/MultiDimensionalFeatures

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.18528 2025-02-27 cs.CR cs.AI cs.RO 79%

ARACNE: An LLM-Based Autonomous Shell Pentesting Agent

Tomas Nieponice, Veronica Valeros, Sebastian Garcia

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI

Comments 7 pages, 2 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.17650 2025-02-26 cs.HC cs.AI 79%

Wearable Meets LLM for Stress Management: A Duoethnographic Study Integrating Wearable-Triggered Stressors and LLM Chatbots for Personalized Interventions

Sameer Neupane, Poorvesh Dongre, Denis Gracanin, Santosh Kumar

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI

Comments In CHI '25 Proceedings of the CHI Conference on Human Factors in Computing Systems Yokohama, Japan

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.13982 2025-02-26 cs.CV cs.LG 79%

Attribute-based Visual Reprogramming for Vision-Language Models

Chengyi Cai, Zesheng Ye, Lei Feng, Jianzhong Qi, Feng Liu

专题命中 其他LLM :language model(title,abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.20777 2025-02-25 cs.CR cs.LG 79%

Black-Box Detection of Language Model Watermarks

Thibaud Gloaguen, Nikola Jovanović, Robin Staab, Martin Vechev

专题命中 其他LLM :language model(title);LLM(abstract);分类 cs.LG

Comments ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.19097 2025-02-25 cs.CL 79%

TEncDM: Understanding the Properties of the Diffusion Model in the Space of Language Model Encodings

Alexander Shabalin, Viacheslav Meshchaninov, Egor Chimbulatov, Vladislav Lapikov, Roman Kim, Grigory Bartosh, Dmitry Molchanov, Sergey Markov, Dmitry Vetrov

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

Comments 15 pages, 13 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.13497 2025-02-21 cs.CL 79%

Repetition Neurons: How Do Language Models Produce Repetitions?

Tatsuya Hiraoka, Kentaro Inui

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

Comments NAACL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.13016 2025-02-19 cs.DB cs.AI 79%

LLM-Powered Proactive Data Systems

Sepanta Zeighami, Yiming Lin, Shreya Shankar, Aditya Parameswaran

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI

Journal ref IEEE Data Engineering Bulletin March 2025

详情

展开后加载摘要…

URL PDF HTML 收藏