arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-10-15 至 2025-10-15 共收录 213 信号源:cs.CL, cs.AI, cs.LG

1. 评测与基准 54 篇

2510.12722 2025-10-15 cs.CL 57%

Which Word Orders Facilitate Length Generalization in LMs? An Investigation with GCG-Based Artificial Languages

Nadine El-Naggar, Tatsuki Kuribayashi, Ted Briscoe

机构 * Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)

专题命中 评测与基准 :language model(abstract);分类 cs.CL

Comments EMNLP 2025 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12548 2025-10-15 cs.CL cs.CV 57%

VISaGE: Understanding Visual Generics and Exceptions

Stella Frank, Emily Allaway

机构 * University of Copenhagen(哥本哈根大学) University of Edinburgh(爱丁堡大学)

专题命中 评测与基准 :language model(abstract);分类 cs.CL

Comments EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11883 2025-10-15 cs.CV cs.AI 57%

MammoDINO: Anatomically Aware Self-Supervision for Mammographic Images

Sicheng Zhou, Lei Wu, Cao Xiao, Parminder Bhatia, Taha Kass-Hout

专题命中 评测与基准 :pretraining(abstract);分类 cs.AI

Comments 5 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22337 2025-10-15 cs.CL cs.IR 57%

A Comprehensive Taxonomy of Negation for NLP and Neural Retrievers

Roxana Petcu, Samarth Bhargav, Maarten de Rijke, Evangelos Kanoulas

机构 * University of Amsterdam(阿姆斯特丹大学)

专题命中 评测与基准 :LLM(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.07421 2025-10-15 cs.CL 57%

AgentAda: Skill-Adaptive Data Analytics for Tailored Insight Discovery

Amirhossein Abaskohi, Amrutha Varshini Ramesh, Shailesh Nanisetty, Chirag Goel, David Vazquez, Christopher Pal, Spandana Gella, Giuseppe Carenini, Issam H. Laradji

机构 * ServiceNow Research(ServiceNow 研究所) University of British Columbia(不列颠哥伦比亚大学) University of Toronto(多伦多大学) University of Montreal(蒙特利尔大学) Mila(Mila 研究所) CIFAR AI Chair(CIFAR 人工智能主席职位) Polytechnique Montréal(蒙特利尔理工学院)

专题命中 评测与基准 :LLM(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.11866 2025-10-15 cs.LG cs.CY 57%

Evaluating multiple models using labeled and unlabeled data

Divya Shanmugam, Shuvom Sadhuka, Manish Raghavan, John Guttag, Bonnie Berger, Emma Pierson

机构 * Department of Computer Science, Cornell University(康奈尔大学计算机科学系) Department of Electrical Engineering & Computer Science, MIT(麻省理工学院电子工程与计算机科学系) Sloan School of Management, MIT(麻省理工学院斯隆管理学院) Department of Electrical Engineering & Computer Sciences, UC Berkeley(伯克利大学电子工程与计算机科学系)

专题命中 评测与基准 :language model(abstract);分类 cs.LG

Comments To appear at NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12565 2025-10-15 cs.CV 50%

MMOT: The First Challenging Benchmark for Drone-based Multispectral Multi-Object Tracking

Tianhao Li, Tingfa Xu, Ying Wang, Haolin Qin, Xu Lin, Jianan Li

机构 * Beijing Institute of Technology(北京理工大学) Beijing Institute of Technology Chongqing Innovation Center(北京理工大学重庆创新中心)

专题命中 评测与基准 :pretraining(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12444 2025-10-15 cs.CV 50%

A Review of Longitudinal Radiology Report Generation: Dataset Composition, Methods, and Performance Evaluation

Shaoyang Zhou, Yingshu Li, Yunyi Liu, Lingqiao Liu, Lei Wang, Luping Zhou

机构 * School of Electrical and Computer Engineering, The University of Sydney(电气与计算机工程学院,悉尼大学) School of Computer Science, The University of Adelaide(计算机科学学院,阿德莱德大学) School of Computing and Information Technology, University of Wollongong(计算与信息科技学院,沃伦冈大学)

专题命中 评测与基准 :language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12326 2025-10-15 eess.AS 50%

DeePAQ: A Perceptual Audio Quality Metric Based On Foundational Models and Weakly Supervised Learning

Guanxin Jiang, Andreas Brendel, Pablo M. Delgado, Jürgen Herre

专题命中 评测与基准 :foundation model(abstract)

Comments 5 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12119 2025-10-15 cs.CV 50%

ImageSentinel: Protecting Visual Datasets from Unauthorized Retrieval-Augmented Image Generation

Ziyuan Luo, Yangyi Zhao, Ka Chun Cheung, Simon See, Renjie Wan

机构 * Department of Computer Science, Hong Kong Baptist University(香港 Baptist 大学计算机科学系) NVIDIA AI Technology Center, NVIDIA(NVIDIA 人工智能技术中心)

专题命中 评测与基准 :language model(abstract)

Comments Accepted at NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.16815 2025-10-15 cs.CV cs.RO 50%

Image Quality Assessment for Embodied AI

Chunyi Li, Jiaohao Xiao, Jianbo Zhang, Farong Wen, Zicheng Zhang, Yuan Tian, Xiangyang Zhu, Xiaohong Liu, Zhengxue Cheng, Weisi Lin, Guangtao Zhai

机构 * Shanghai Jiao Tong University(上海交通大学) Shanghai AI Lab(上海人工智能实验室) Nanyang Technological University(南洋理工大学)

专题命中 评测与基准 :language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19764 2025-10-15 cs.CV cs.RO 50%

OpenLex3D: A Tiered Evaluation Benchmark for Open-Vocabulary 3D Scene Representations

Christina Kassab, Sacha Morin, Martin Büchner, Matías Mattamala, Kumaraditya Gupta, Abhinav Valada, Liam Paull, Maurice Fallon

专题命中 评测与基准 :language model(abstract)

Comments NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 效率与部署 36 篇

2510.12637 2025-10-15 cs.CL 92%

COSTAR-A: A prompting framework for enhancing Large Language Model performance on Point-of-View questions

Nzubechukwu C. Ohalete, Kevin B. Gittner, Lauren M. Matheny

专题命中 效率与部署 :large language model(title,abstract);language model(title,abstract);prompting(title,abstract);分类 cs.CL

Comments 20 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11837 2025-10-15 cs.CR cs.AI 90%

Countermind: A Multi-Layered Security Architecture for Large Language Models

Dominik Schwarz

机构 * Independent Researcher(独立研究者)

专题命中 效率与部署 :large language model(title,abstract);language model(title,abstract);LLM(abstract,comments);分类 cs.AI

Comments 33 pages, 3 figures, 6 tables. Keywords: LLM security; defense-in-depth; prompt injection; activation steering; multimodal sandbox; threat modeling

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12015 2025-10-15 cs.AI 89%

Asking Clarifying Questions for Preference Elicitation With Large Language Models

Ali Montazeralghaem, Guy Tennenholtz, Craig Boutilier, Ofer Meshi

机构 * Google(谷歌)

专题命中 效率与部署 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12633 2025-10-15 cs.LG cs.AI cs.DC 86%

Laminar: A Scalable Asynchronous RL Post-Training Framework

Guangming Sheng, Yuxuan Tong, Borui Wan, Wang Zhang, Chaobo Jia, Xibin Wu, Yuqi Wu, Xiang Li, Chi Zhang, Yanghua Peng, Haibin Lin, Xin Liu, Chuan Wu

机构 * The University of Hong Kong(香港大学) ByteDance Seed(字节跳动种子)

专题命中 效率与部署 :post-training(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11813 2025-10-15 cs.SE cs.CL cs.DB 86%

Task-Aware Reduction for Scalable LLM-Database Systems

Marcus Emmanuel Barnes, Taher A. Ghaleb, Safwat Hassan

机构 * Department of Computer Science Trent University Peterborough, Canada(计算机科学系特伦特大学彼得伯勒,加拿大) Faculty of Information University of Toronto Toronto, Canada(信息学院多伦多大学多伦多,加拿大)

专题命中 效率与部署 :LLM(title,abstract);language model(abstract,comments);large language model(abstract);分类 cs.CL

Comments Preprint. Accepted for presentation at the Workshop on Language Models and Databases (LMD), co-located with CASCON 2025 (IEEE). The final version will appear in IEEE Xplore

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12721 2025-10-15 cs.LG 85%

CARVQ: Corrective Adaptor with Group Residual Vector Quantization for LLM Embedding Compression

Dayin Gou, Sanghyun Byun, Nilesh Malpeddi, Gabrielle De Micheli, Prathamesh Vaste, Jacob Song, Woo Seong Chung

机构 * LG Electronics USA(LG电子美国公司)

专题命中 效率与部署 :LLM(title);large language model(abstract);language model(abstract);post-training(abstract)

Comments Accepted at EMNLP Findings 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12343 2025-10-15 cs.LG cs.CR 85%

Traveling Salesman-Based Token Ordering Improves Stability in Homomorphically Encrypted Language Models

Donghwan Rho, Sieun Seo, Hyewon Sung, Chohong Min, Ernest K. Ryu

专题命中 效率与部署 :language model(title,abstract);LLM(abstract);large language model(abstract);分类 cs.LG

Comments 34 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12023 2025-10-15 cs.CL 85%

Information Extraction from Conversation Transcripts: Neuro-Symbolic vs. LLM

Alice Saebom Kwak, Maria Alexeeva, Gus Hahn-Powell, Keith Alcock, Kevin McLaughlin, Doug McCorkle, Gabe McNunn, Mihai Surdeanu

机构 * Department of Linguistics, University of Arizona(亚利桑那大学语言学系) Lum AI Eocene Environmental Group(Eocene环境集团)

专题命中 效率与部署 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL

Comments 15 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.04075 2025-10-15 cs.CL 85%

From Rational Answers to Emotional Resonance: The Role of Controllable Emotion Generation in Language Models

Yurui Dong, Luozhijie Jin, Yao Yang, Bingjie Lu, Jiaxi Yang, Zhi Liu

机构 * State Key Laboratory on Technologies for Chinese Medicine Pharmaceutical Process Control and Intelligent Manufacture(中药制药过程控制与智能制造技术国家重点实验室) Zhejiang Lab(浙江实验室)

专题命中 效率与部署 :language model(title,abstract);LLM(abstract);large language model(abstract);分类 cs.CL

Comments 43 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.00299 2025-10-15 cs.CL 85%

ChunkKV: Semantic-Preserving KV Cache Compression for Efficient Long-Context LLM Inference

Xiang Liu, Zhenheng Tang, Peijie Dong, Zeyu Li, Yue Liu, Bo Li, Xuming Hu, Xiaowen Chu

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) CSE, The Hong Kong University of Science and Technology(香港科学与技术大学计算机科学与工程系) Guangzhou HKUST Fok Ying Tung Research Institute(广州HKUST福 Ying Tung研究 institute) Terminus Technologies

专题命中 效率与部署 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL

Comments NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11915 2025-10-15 cs.CR 85%

Robust ML-based Detection of Conventional, LLM-Generated, and Adversarial Phishing Emails Using Advanced Text Preprocessing

Deeksha Hareesha Kulal, Chidozie Princewill Arannonu, Afsah Anwar, Nidhi Rastogi, Quamar Niyaz

专题命中 效率与部署 :LLM(title,abstract);large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12389 2025-10-15 cs.CL cs.AI 84%

Tokenization Disparities as Infrastructure Bias: How Subword Systems Create Inequities in LLM Access and Efficiency

Hailay Kidu Teklehaymanot, Wolfgang Nejdl

机构 * L3S Research Center(L3S研究所以) Leibniz University Hannover(莱比锡大学汉诺威分校)

专题命中 效率与部署 :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments 6 pages 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11962 2025-10-15 cs.LG cs.CV 83%

MosaicDiff: Training-free Structural Pruning for Diffusion Model Acceleration Reflecting Pretraining Dynamics

Bowei Guo, Shengkun Tang, Cong Zeng, Zhiqiang Shen

机构 * Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)

专题命中 效率与部署 :pretraining(title,abstract);post-training(abstract);分类 cs.LG

Comments International Conference on Computer Vision, ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.02613 2025-10-15 cs.LG cs.AI 81%

ACCO: Accumulate While You Communicate for Communication-Overlapped Sharded LLM Training

Adel Nabli, Louis Fournier, Pierre Erbacher, Louis Serrano, Eugene Belilovsky, Edouard Oyallon

专题命中 效率与部署 :LLM(title,abstract);分类 cs.AI、cs.LG

Journal ref The Thirty-ninth Annual Conference on Neural Information Processing Systems, Dec 2025, San diego (Californie), United States

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.07961 2025-10-15 cs.RO cs.AI 79%

BridgeVLA: Input-Output Alignment for Efficient 3D Manipulation Learning with Vision-Language Models

Peiyan Li, Yixiang Chen, Hongtao Wu, Xiao Ma, Xiangnan Wu, Yan Huang, Liang Wang, Tao Kong, Tieniu Tan

机构 * CASIA(中国科学院自动化研究所) ByteDance Seed(字节跳动种子实验室) UCAS(中国科学院大学) FiveAges NJU(南京大学)

专题命中 效率与部署 :language model(title,abstract);分类 cs.AI

Comments NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12160 2025-10-15 cs.CV 78%

State Space Prompting via Gathering and Spreading Spatio-Temporal Information for Video Understanding

Jiahuan Zhou, Kai Zhu, Zhenyu Cui, Zichen Liu, Xu Zou, Gang Hua

机构 * Wangxuan Institute of Computer Technology, Peking University(北京大学王轩计算机技术研究所) the Huazhong University of Science and Technology(华中科技大学) Amazon.com, Inc(亚马逊公司)

专题命中 效率与部署 :prompting(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.02631 2025-10-15 cs.LG cs.AI cs.CL cs.CV 78%

ParetoQ: Improving Scaling Laws in Extremely Low-bit LLM Quantization

Zechun Liu, Changsheng Zhao, Hanxian Huang, Sijia Chen, Jing Zhang, Jiawei Zhao, Scott Roy, Lisa Jin, Yunyang Xiong, Yangyang Shi, Lin Xiao, Yuandong Tian, Bilge Soran, Raghuraman Krishnamoorthi, Tijmen Blankevoort, Vikas Chandra

机构 * Meta AI

专题命中 效率与部署 :LLM(title);分类 cs.CL、cs.AI、cs.LG

Comments NeurIPS 2025. Model weights are available at https://huggingface.co/collections/facebook/mobilellm-6722be18cb86c20ebe113e95

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12742 2025-10-15 cs.AI cs.IR 77%

CTRL-Rec: Controlling Recommender Systems With Natural Language

Micah Carroll, Adeline Foote, Kevin Feng, Marcus Williams, Anca Dragan, W. Bradley Knox, Smitha Milli

机构 * MATS University of Washington(华盛顿大学) UC Berkeley(加州大学伯克利分校) UT Austin(得克萨斯大学奥斯汀分校) FAIR at Meta(Meta的FAIR)

专题命中 效率与部署 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏