arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12266 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12266 篇

2502.05121 2025-02-10 hep-th cs.LG hep-ph 70%

Refining Integration-by-Parts Reduction of Feynman Integrals with Machine Learning

Matt von Hippel, Matthias Wilhelm

机构 * Niels Bohr International Academy(尼尔斯·玻尔国际学院) Niels Bohr Institute(尼尔斯·玻尔研究所) University of Copenhagen(哥本哈根大学) Center for Quantum Mathematics(量子数学中心) University of Southern Denmark(南丹麦大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.LG

Comments 28 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.04218 2025-02-07 cs.CL 70%

Sports and Women's Sports: Gender Bias in Text Generation with Olympic Data

Laura Biester

机构 * Middlebury College(明德学院)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

Comments NAACL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.10646 2025-02-07 cs.AI cs.CY 70%

Ethical ChatGPT: Concerns, Challenges, and Commandments

Jianlong Zhou, Heimo Müller, Andreas Holzinger, Fang Chen

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

Comments 8 pages, 2 figures

Journal ref Electronics, 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.05524 2025-02-07 cs.CL cs.DB 70%

Context-Driven Index Trimming: A Data Quality Perspective to Enhancing Precision of RALMs

Kexin Ma, Ruochun Jin, Xi Wang, Huan Chen, Jing Ren, Yuhua Tang

机构 * National University of Defense Technology(国防科技大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.02893 2025-02-06 cs.CL 70%

Lowering the Barrier of Machine Learning: Achieving Zero Manual Labeling in Review Classification Using LLMs

Yejian Zhang, Shingo Takada

机构 * Keio University(庆应义塾大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

Comments Accepted to 2025 11th International Conference on Computing and Artificial Intelligence (ICCAI 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.10020 2025-02-06 cs.CL cs.MM 70%

Lost in Overlap: Exploring Logit-based Watermark Collision in LLMs

Yiyang Luo, Ke Lin, Chao Gu, Jiahui Hou, Lijie Wen, Ping Luo

机构 * Nanyang Technological University(南洋理工大学) Tsinghua University(清华大学) University of Science and Technology of China(中国科学技术大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

Comments Long Paper, 9 pages, accepted at NAACL 2025 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.14630 2025-02-05 cs.AI 70%

Extracting Problem Structure with LLMs for Optimized SAT Local Search

André Schidler, Stefan Szeider

机构 * TU Wien(维也纳技术大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.16615 2025-01-31 cs.LG 70%

Sparse Autoencoders Trained on the Same Data Learn Different Features

Gonçalo Paulo, Nora Belrose

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.16164 2025-01-28 cs.HC cs.AI cs.ET cs.MM 70%

MetaDecorator: Generating Immersive Virtual Tours through Multimodality

Shuang Xie, Yang Liu, Jeannie S. A. Lee, Haiwei Dong

机构 * Shopify Inc.(Shopify公司) Singapore Institute of Technology(新加坡理工学院) Huawei Canada(华为加拿大公司)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.15922 2025-01-28 cs.SE cs.LG 70%

SkillScope: A Tool to Predict Fine-Grained Skills Needed to Solve Issues on GitHub

Benjamin C. Carter, Jonathan Rivas Contreras, Carlos A. Llanes Villegas, Pawan Acharya, Jack Utzerath, Adonijah O. Farner, Hunter Jenkins, Dylan Johnson, Jacob Penney, Igor Steinmacher, Marco A. Gerosa, Fabio Santos

机构 * Grand Canyon University(大峡谷大学) Northern Arizona University(北亚利桑那大学) Colorado State University(科罗拉多州立大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.15772 2025-01-28 cs.SI cs.AI cs.CY 70%

Analyzing User Characteristics of Hate Speech Spreaders on Social Media

Dominique Geissler, Abdurahman Maarouf, Stefan Feuerriegel

机构 * LMU Munich(慕尼黑大学) Munich Center for Machine Learning(慕尼黑机器学习中心)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.17094 2025-01-27 cs.SE cs.AI 70%

Analysis on LLMs Performance for Code Summarization

Md. Ahnaf Akib, Md. Muktadir Mazumder, Salman Ahsan

机构 * Department of Computer Science and Engineering(计算机科学与工程系)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.12793 2025-01-23 cs.SE cs.AI 70%

Revisit Self-Debugging with Self-Generated Tests for Code Generation

Xiancai Chen, Zhengwei Tao, Kechi Zhang, Changzhi Zhou, Wanli Gu, Yuanpeng He, Mengdi Zhang, Xunliang Cai, Haiyan Zhao, Zhi Jin

机构 * Peking University(北京大学) Beijing Institute of Technology(北京理工大学) Meituan(美团)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

Comments Work in Progress

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.12033 2025-01-22 cs.NI cs.AI 70%

Harnessing Generative Pre-Trained Transformer for Datacenter Packet Trace Generation

Chen Griner

机构 * Ben-Gurion University of the Negev(内盖夫本-古里安大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.10909 2025-01-22 cs.AI cs.HC 70%

Fine-Grained Appropriate Reliance: Human-AI Collaboration with a Multi-Step Transparent Decision Workflow for Complex Task Decomposition

Gaole He, Patrick Hemmer, Michael Vössing, Max Schemmer, Ujwal Gadiraju

机构 * Delft University of Technology(代尔夫特理工大学) Karlsruhe Institute of Technology(卡尔斯鲁厄理工学院)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

Comments Work in progress

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.10011 2025-01-20 cs.CV cs.AI 70%

Mitigating Hallucinations on Object Attributes using Multiview Images and Negative Instructions

Zhijie Tan, Yuzhi Li, Shengwei Meng, Xiang Yuan, Weiping Li, Tong Mo, Bingce Wang, Xu Chu

机构 * Peking University(北京大学) Beijing University of Posts and Telecommunications(北京邮电大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

Comments 2025 IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.09950 2025-01-20 cs.SI cs.CL 70%

Sympathy over Polarization: A Computational Discourse Analysis of Social Media Posts about the July 2024 Trump Assassination Attempt

Qingcheng Zeng, Guanhong Liu, Zhaoqian Xue, Diego Ford, Rob Voigt, Loni Hagen, Lingyao Li

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.17308 2025-01-20 cs.LG math.ST stat.TH 70%

Consistent estimation of generative model representations in the data kernel perspective space

Aranyak Acharyya, Michael W. Trosset, Carey E. Priebe, Hayden S. Helm

机构 * Johns Hopkins University(约翰斯·霍普金斯大学) Indiana University(印第安纳大学) Helivan Research(海利万研究院)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.18550 2025-01-20 cs.IR cs.AI 70%

Search Still Matters: Information Retrieval in the Era of Generative AI

William R. Hersh

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

Comments 7 pages, no figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.09219 2025-01-17 cs.CL 70%

A Simple Graph Contrastive Learning Framework for Short Text Classification

Yonghao Liu, Fausto Giunchiglia, Lan Huang, Ximing Li, Xiaoyue Feng, Renchu Guan

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

Comments AAAI2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.06276 2025-01-14 cs.SD cs.CL eess.AS 70%

PROEMO: Prompt-Driven Text-to-Speech Synthesis Based on Emotion and Intensity Control

Shaozuo Zhang, Ambuj Mehrish, Yingting Li, Soujanya Poria

机构 * Singapore University of Technology and Design(新加坡科技设计大学) Beijing University of Posts and Telecommunications(北京邮电大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.16527 2025-01-14 cs.CL 70%

Profiling Bias in LLMs: Stereotype Dimensions in Contextual Word Embeddings

Carolin M. Schuster, Maria-Alexandra Dinisor, Shashwat Ghatiwala, Georg Groh

机构 * TUM School of Computation, Information and Technology(慕尼黑工业大学计算、信息与技术学院) Technical University of Munich(慕尼黑工业大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

Comments Accepted to NoDaLiDa/Baltic-HLT 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.04242 2025-01-14 cs.LG 70%

Multimodal Structure-Aware Quantum Data Processing

Hala Hawashin, Mehrnoosh Sadrzadeh

机构 * University College London(伦敦大学学院)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.LG

Comments 10 Pages, 16 Figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.04289 2025-01-14 cs.CL 70%

What Languages are Easy to Language-Model? A Perspective from Learning Probabilistic Regular Languages

Nadav Borenstein, Anej Svete, Robin Chan, Josef Valvoda, Franz Nowak, Isabelle Augenstein, Eleanor Chodroff, Ryan Cotterell

机构 * Københavns Universitet(哥本哈根大学) ETH Zürich(苏黎世联邦理工学院) Universität Zürich(苏黎世大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

Comments Accepted to ACL 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.00039 2025-01-13 q-bio.NC cs.CL cs.HC 70%

Decoding moral judgement from text: a pilot study

Diana E. Gherman, Thorsten O. Zander

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

Comments 7 pages, 2 figures, conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.05423 2025-01-10 cs.SI cs.LG 70%

Using LLMs to Infer Non-Binary COVID-19 Sentiments of Chinese Micro-bloggers

Jerry Chongyi Hu, Mohammed Shahid Modi, Boleslaw K. Szymanski

机构 * Rensselaer Polytechnic Institute(伦斯勒理工学院)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.LG

Comments 11 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.04835 2025-01-10 cs.SE cs.AI 70%

Do Code LLMs Understand Design Patterns?

Zhenyu Pan, Xuefeng Song, Yunkun Wang, Rongyu Cao, Binhua Li, Yongbin Li, Han Liu

机构 * Northwestern University(西北大学) Zhejiang University(浙江大学) Alibaba Group(阿里巴巴集团)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

Comments accpeted by llm4code workshop in ICSE 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.12239 2025-01-10 cs.CL 70%

What makes a language easy to deep-learn? Deep neural networks and humans similarly benefit from compositional structure

Lukas Galke, Yoav Ram, Limor Raviv

机构 * University of Southern Denmark(南丹麦大学) Max Planck Institute for Psycholinguistics(马克斯·普朗克心理语言学研究所) Tel Aviv University(特拉维夫大学) University of Glasgow(格拉斯哥大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

Comments 20 pages + supplementary material

Journal ref Nature Communications, vol. 15, article 10816, 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.04302 2025-01-09 cs.CV cs.AI 70%

H-MBA: Hierarchical MamBa Adaptation for Multi-Modal Video Understanding in Autonomous Driving

Siran Chen, Yuxiao Luo, Yue Ma, Yu Qiao, Yali Wang

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

Comments 7 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.17624 2025-01-08 q-fin.RM cs.CL q-fin.GN 70%

Forecasting Credit Ratings: A Case Study where Traditional Methods Outperform Generative LLMs

Felix Drinkall, Janet B. Pierrehumbert, Stefan Zohren

机构 * University of Oxford(牛津大学) The Alan Turing Institute(艾伦·图灵研究所)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏