arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12287 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12287 篇

2509.04476 2025-09-18 cs.CL cs.AI 62%

Training Text-to-Molecule Models with Context-Aware Tokenization

Seojin Kim, Hyeontae Song, Jaehyun Nam, Jinwoo Shin

机构 * Seoul National University(首尔国立大学) Moloco Inc.(Moloco公司) Korea Advanced Institute of Science and Technology (KAIST)(韩国科学技术院)

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.AI

Comments EMNLP 2025 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.19331 2025-09-17 cs.CV cs.AI cs.CL 62%

Talking to DINO: Bridging Self-Supervised Vision Backbones with Language for Open-Vocabulary Segmentation

Luca Barsellotti, Lorenzo Bianchi, Nicola Messina, Fabio Carrara, Marcella Cornia, Lorenzo Baraldi, Fabrizio Falchi, Rita Cucchiara

机构 * University of Modena and Reggio Emilia(摩德纳和雷吉奥艾米利亚大学) ISTI-CNR(意大利国家研究委员会ISTI) University of Pisa(比萨大学)

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.AI

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11824 2025-09-16 cs.IR cs.AI cs.LG cs.SD 62%

Data-Driven Analysis of Text-Conditioned AI-Generated Music: A Case Study with Suno and Udio

Luca Casini, Laura Cros Vila, David Dalmazzo, Anna-Kaisa Kaila, Bob L. T. Sturm

机构 * KTH Royal Institute of Technology(皇家理工学院)

专题命中 其他LLM :prompting(abstract);分类 cs.AI、cs.LG

Comments Submitted for review to TISMIR Digital Musicology special issue

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.18492 2025-09-12 cs.HC cs.AI cs.CL 62%

VeriSafe Agent: Safeguarding Mobile GUI Agent via Logic-based Action Verification

Jungjae Lee, Dongjae Lee, Chihun Choi, Youngmin Im, Jaeyoung Wi, Kihong Heo, Sangeun Oh, Sunjae Lee, Insik Shin

机构 * School of Computing, KAIST(计算机学院,韩国科学技术院) Korea University(韩国大学) Sungkyunkwan University(成均馆大学)

专题命中 其他LLM :foundation model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.07998 2025-09-11 cs.CL cs.AI 62%

Bilingual Word Level Language Identification for Omotic Languages

Mesay Gemeda Yigezu, Girma Yohannis Bade, Atnafu Lambebo Tonja, Olga Kolesnikova, Grigori Sidorov, Alexander Gelbukh

机构 * Institutetext: Instituto Politécnico Nacional (IPN), Centro de Investigación en Computación (CIC), Mexico City, Mexico(墨西哥城国家理工学院(IPN)计算机研究中心)

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.08715 2025-09-05 eess.AS cs.AI cs.CL eess.SP 62%

MultiGen: Child-Friendly Multilingual Speech Generator with LLMs

Xiaoxue Gao, Huayun Zhang, Nancy F. Chen

机构 * Institute for Infocomm Research, Agency for Science, Technology, and Research (A*STAR), Singapore(信息与通信研究所,科技研究局(A*STAR),新加坡)

专题命中 其他LLM :LLM(abstract);分类 cs.CL、cs.AI

Comments 5 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.10732 2025-08-15 cs.LG cs.AI 62%

APFL: Analytic Personalized Federated Learning via Dual-Stream Least Squares

Kejia Fan, Jianheng Tang, Zhirui Yang, Feijiang Han, Jiaxu Li, Run He, Yajiang Huang, Anfeng Liu, Houbing Herbert Song, Yunhuai Liu, Huiping Zhuang

专题命中 其他LLM :foundation model(abstract);分类 cs.AI、cs.LG

Comments 9 pages, 4 figures, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.18862 2025-08-14 cs.LG cs.AI 62%

One-shot Optimized Steering Vectors Mediate Safety-relevant Behaviors in LLMs

Jacob Dunefsky, Arman Cohan

机构 * Department of Computer Science(计算机科学系) Yale University(耶鲁大学)

专题命中 其他LLM :LLM(abstract);分类 cs.AI、cs.LG

Comments Published at COLM 2025. 30 pages, 7 figures. Code is available at https://github.com/jacobdunefsky/one-shot-steering-repro and https://github.com/jacobdunefsky/one-shot-steering-misalignment

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02622 2025-08-11 cs.AI cs.CL cs.CY 62%

Noosemia: toward a Cognitive and Phenomenological Account of Intentionality Attribution in Human-Generative AI Interaction

Enrico De Santis, Antonello Rizzi

专题命中 其他LLM :LLM(abstract);分类 cs.CL、cs.AI

Comments This version has been extensively revised and revisited in light of feedback and further research. Several sections have been expanded or improved for greater clarity and completeness. Specifically, new clarification on complex system foundation related to Noosemia has been added (Secs. "2.4 and "2.5")

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.05913 2025-08-11 cs.HC cs.AI cs.CL 62%

Do Ethical AI Principles Matter to Users? A Large-Scale Analysis of User Sentiment and Satisfaction

Stefan Pasch, Min Chul Cha

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.13422 2025-08-05 cs.CL cs.AI cs.DB 62%

Towards Question Answering over Large Semi-structured Tables

Yuxiang Wang, Junhao Gan, Jianzhong Qi

专题命中 其他LLM :LLM(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22847 2025-07-31 cs.AI cs.CL cs.CY 62%

The Incomplete Bridge: How AI Research (Mis)Engages with Psychology

Han Jiang, Pengda Wang, Xiaoyuan Yi, Xing Xie, Ziang Xiao

机构 * Department of Computer Science, Johns Hopkins University(约翰霍普金斯大学计算机科学系) Microsoft Research Asia(微软亚洲研究院) Department of Psychological Sciences, Rice University(里士满大学心理学科学系)

专题命中 其他LLM :LLM(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22445 2025-07-31 cs.CL cs.AI 62%

AI-generated stories favour stability over change: homogeneity and cultural stereotyping in narratives generated by gpt-4o-mini

Jill Walker Rettberg, Hermann Wigers

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.AI

Comments This project has received funding from the European Union's Horizon 2020 research and innovation programme under grant agreement number 101142306. The project is also supported by the Center for Digital Narrative, which is funded by the Research Council of Norway through its Centres of Excellence scheme, project number 332643

Journal ref Open Research Europe 2025, 5:202 [version 1; peer review: awaiting peer review]

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.01966 2025-07-24 cs.LG cs.AI 62%

Unified Sparse-Matrix Representations for Diverse Neural Architectures

Yuzhou Zhu

机构 * Dalian University of Technology(大连理工大学)

专题命中 其他LLM :language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.12469 2025-07-18 cs.CC cs.CL cs.LG 62%

Perfect diffusion is $\mathsf{TC}^0$ -- Bad diffusion is Turing-complete

Yuxi Liu

机构 * Berkeley Artificial Intelligence Research Lab, UC Berkeley(伯克利人工智能研究实验室,伯克利大学)

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.LG

Comments 7 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.07906 2025-07-11 cs.LG cs.AI 62%

Agentic Retrieval of Topics and Insights from Earnings Calls

Anant Gupta, Rajarshi Bhowmik, Geoffrey Gunow

机构 * Bloomberg USA(彭博美国)

专题命中 其他LLM :LLM(abstract);分类 cs.AI、cs.LG

Comments The 2nd Workshop on Financial Information Retrieval in the Era of Generative AI, The 48th International ACM SIGIR Conference on Research and Development in Information Retrieval July 13-17, 2025 | Padua, Italy

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.06253 2025-07-10 cs.CR cs.AI cs.CL cs.HC 62%

Emergent misalignment as prompt sensitivity: A research note

Tim Wyse, Twm Stone, Anna Soligo, Daniel Tan

机构 * UCL(伦敦大学学院)

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.AI

Comments 10 pages, 15 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.07616 2025-07-09 cs.CL cs.LG 62%

Tractable Transformers for Flexible Conditional Generation

Anji Liu, Xuejie Liu, Dayuan Zhao, Mathias Niepert, Yitao Liang, Guy Van den Broeck

机构 * Department of Computer Science, University of California, Los Angeles(加州大学洛杉矶分校计算机科学系) Institute for Artificial Intelligence, Peking University(北京大学人工智能研究院) Yuanpei College, Peking University(北京大学元培学院) Institute for Artificial Intelligence, University of Stuttgart(斯图加特大学人工智能研究院) School of Intelligence Science and Technology, Peking University(北京大学智能科学与技术学校)

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.03589 2025-07-08 cs.CV cs.AI cs.CL 62%

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance

Huy Le, Nhat Chung, Tung Kieu, Anh Nguyen, Ngan Le

机构 * FPT Software AI Center(FPT软件AI中心) Aalborg University(奥尔堡大学) Pioneer Centre for AI(先锋人工智能中心) University of Liverpool(利物浦大学) AICV Lab, University of Arkansas(AICV实验室,阿肯色大学)

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.AI

Comments Accepted at ACM MM 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.23596 2025-07-01 cs.LG cs.AI 62%

When Will It Fail?: Anomaly to Prompt for Forecasting Future Anomalies in Time Series

Min-Yeong Park, Won-Jeong Lee, Seong Tae Kim, Gyeong-Moon Park

机构 * Department of Artificial Intelligence, Kyung Hee University, Yongin, Republic of Korea(韩国成均馆大学人工智能系) Department of Artificial Intelligence, Korea University, Seoul, Republic of Korea(韩国大学人工智能系)

专题命中 其他LLM :prompting(abstract);分类 cs.AI、cs.LG

Comments 18 pages, 10 figures, 12 tables, ICML 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.09850 2025-07-01 cs.CV cs.AI cs.LG 62%

Enhancing Diffusion Posterior Sampling for Inverse Problems by Integrating Crafted Measurements

Shijie Zhou, Huaisheng Zhu, Rohan Sharma, Jiayi Chen, Ruiyi Zhang, Kaiyi Ji, Changyou Chen

机构 * University at Buffalo(布法罗大学) Adobe Research(Adobe研究) The Pennsylvania State University(宾夕法尼亚州立大学) Fujian Normal University(福建师范大学)

专题命中 其他LLM :foundation model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.19732 2025-06-25 cs.LG cs.AI 62%

Who Does What in Deep Learning? Multidimensional Game-Theoretic Attribution of Function of Neural Units

Shrey Dixit, Kayson Fakhar, Fatemeh Hadaeghi, Patrick Mineault, Konrad P. Kording, Claus C. Hilgetag

机构 * Institute of Computational Neuroscience(计算神经科学研究所) University Medical Center Eppendorf(埃本多夫大学医学中心) International Max Planck Research School on Cognitive Neuroimaging(认知神经成像国际马克斯·普朗克研究学校) MRC Cognition and Brain Sciences Unit(MRC认知与脑科学单位) Mila - Quebec Artificial Intelligence Institute(魁北克人工智能研究所) Learning in Machines & Brains(机器与大脑学习) Departments of Bioengineering and Neuroscience(生物工程与神经科学系) Department of Health Sciences(健康科学系)

专题命中 其他LLM :LLM(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.13895 2025-06-24 cs.CL cs.LG 62%

RTSUM: Relation Triple-based Interpretable Summarization with Multi-level Salience Visualization

Seonglae Cho, Yonggi Cho, HoonJae Lee, Myungha Jang, Jinyoung Yeo, Dongha Lee

机构 * Yonsei University, Republic of Korea(延世大学,韩国)

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.LG

Comments 8 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.14797 2025-06-19 cs.LG cs.AI 62%

Bound by semanticity: universal laws governing the generalization-identification tradeoff

Marco Nurisso, Jesseba Fernando, Raj Deshpande, Alan Perotti, Raja Marjieh, Steven M. Frankland, Richard L. Lewis, Taylor W. Webb, Declan Campbell, Francesco Vaccarino, Jonathan D. Cohen, Giovanni Petri

机构 * Dipartimento di Scienze Matematiche, Politecnico di Torino(都灵理工大学数学科学系) CENTAI Institute(CENTAI研究院) Network Science Institute, Northeastern University(东北大学网络科学研究所) Institute for Experiential AI, Northeastern University(东北大学体验人工智能研究所) NP Lab, Network Science Institute, Northeastern University London(东北大学伦敦网络科学研究所NP实验室) Department of Psychology, Princeton University(普林斯顿大学心理学系) Program in Cognitive Science, Dartmouth College(达特茅斯学院认知科学项目) Department of Psychology, University of Michigan(密歇根大学心理学系) Microsoft Research(微软研究院) Princeton Neuroscience Institute(普林斯顿神经科学研究所) Department of Physics, Northeastern University(东北大学物理系)

专题命中 其他LLM :language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.08970 2025-06-16 cs.CR cs.AI cs.LG 62%

Self-interpreting Adversarial Images

Tingwei Zhang, Collin Zhang, John X. Morris, Eugene Bagdasarian, Vitaly Shmatikov

机构 * Cornell Tech(康奈尔科技) University of Massachusetts Amherst(马萨诸塞大学阿默斯特分校)

专题命中 其他LLM :language model(abstract);分类 cs.AI、cs.LG

Comments in USENIX Security 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.16586 2025-06-11 cs.HC cs.AI cs.CL cs.CR cs.CY 62%

Big Help or Big Brother? Auditing Tracking, Profiling, and Personalization in Generative AI Assistants

Yash Vekaria, Aurelio Loris Canino, Jonathan Levitsky, Alex Ciechonski, Patricia Callejo, Anna Maria Mandalari, Zubair Shafiq

机构 * UC Davis(加州大学戴维斯分校) UNIRC(意大利国家研究委员会) UCL(伦敦大学学院) UC3M(马德里卡洛斯三世大学)

专题命中 其他LLM :prompting(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.00493 2025-06-11 eess.AS cs.AI cs.CL cs.SD 62%

LLaSE-G1: Incentivizing Generalization Capability for LLaMA-based Speech Enhancement

Boyi Kang, Xinfa Zhu, Zihan Zhang, Zhen Ye, Mingshuai Liu, Ziqian Wang, Yike Zhu, Guobin Ma, Jun Chen, Longshuai Xiao, Chao Weng, Wei Xue, Lei Xie

机构 * Northwestern Polytechnical University(西北工业大学) The Hong Kong University of Science and Technology(香港科技大学) Huawei Technologies Co., Ltd.(华为技术有限公司)

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.AI

Comments ACL2025 main, Codes available at https://github.com/Kevin-naticl/LLaSE-G1

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.08279 2025-06-11 cs.CV cs.AI cs.LG 62%

Seeing Voices: Generating A-Roll Video from Audio with Mirage

Aditi Sundararaman, Amogh Adishesha, Andrew Jaegle, Dan Bigioi, Hyoung-Kyu Song, Jon Kyl, Justin Mao, Kevin Lan, Mojtaba Komeili, ShahRukh Athar, Sheila Babayan, Stanislau Beliasau, William Buchwalter

专题命中 其他LLM :foundation model(abstract);分类 cs.AI、cs.LG

Comments Technical report website: mirage.app/research/seeing-voices, product website: mirage.app

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.07428 2025-06-10 cs.AI cs.LG 62%

HeTa: Relation-wise Heterogeneous Graph Foundation Attack Model

Yuling Wang, Zihui Chen, Pengfei Jiao, Xiao Wang

机构 * School of Cyberspace, Hangzhou Dianzi University(电子科技大学信息学院) Data Security Governance Zhejiang Engineering Research Center, Hangzhou Dianzi University(浙江省数据安全治理工程研究中心) Beihang University(北航)

专题命中 其他LLM :foundation model(abstract);分类 cs.AI、cs.LG

Comments Accepted by IJCAI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.03782 2025-06-06 cs.CL cs.LG 62%

The broader spectrum of in-context learning

Andrew Kyle Lampinen, Stephanie C. Y. Chan, Aaditya K. Singh, Murray Shanahan

机构 * Google DeepMind(谷歌DeepMind) Gatsby Computational Neuroscience Unit, UCL(Gatsby计算神经科学单位,UCL)

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏