arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-07-31 至 2025-07-31 共收录 120 信号源:cs.CL, cs.AI, cs.LG

1. 评测与基准 36 篇

2507.22076 2025-07-31 cs.LG 70%

Test-time Prompt Refinement for Text-to-Image Models

Mohammad Abdul Hafeez Khan, Yash Jain, Siddhartha Bhattacharyya, Vibhav Vineet

机构 * Florida Institute of Technology(佛罗里达理工学院) Microsoft Research(微软研究院)

专题命中 评测与基准 :large language model(abstract);language model(abstract);分类 cs.LG

Comments Accepted to ICCV 2025, MARS2 Workshop. Total 14 pages, 12 figures and 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22346 2025-07-31 cs.CV 67%

DeltaVLM: Interactive Remote Sensing Image Change Analysis via Instruction-guided Difference Perception

Pei Deng, Wenqian Zhou, Hanlin Wu

机构 * School of Information Science and Technology, Beijing Foreign Studies University(信息科学与技术学院,北京外国语大学)

专题命中 评测与基准 :large language model(abstract);language model(abstract)

Comments 12 pages, 5 figures. Submitted to IEEE Transactions on Geoscience and Remote Sensing (TGRS). Code and dataset are available at https://github.com/hanlinwu/DeltaVLM

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22410 2025-07-31 cs.CL cs.AI 62%

Question Generation for Assessing Early Literacy Reading Comprehension

Xiaocheng Yang, Sumuk Shashidhar, Dilek Hakkani-Tur

专题命中 评测与基准 :language model(abstract);分类 cs.CL、cs.AI

Comments 2 pages, 1 figure, accepted by SLaTE 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.07631 2025-07-31 cs.LG cs.CL 62%

OWLViz: An Open-World Benchmark for Visual Question Answering

Thuy Nguyen, Dang Nguyen, Hoang Nguyen, Thuan Luong, Long Hoang Dang, Viet Dac Lai

机构 * Posts and Telecommunications Institute of Technology, Viet Nam(越南电信技术研究所) Adobe Research, USA(Adobe研究) University of Maryland, USA(美国马里兰大学)

专题命中 评测与基准 :language model(abstract);分类 cs.CL、cs.LG

Comments 8 pages + appendix

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22189 2025-07-31 cs.LG cs.AI 62%

Measuring Time-Series Dataset Similarity using Wasserstein Distance

Hongjie Chen, Akshay Mehra, Josh Kimball, Ryan A. Rossi

机构 * Dolby Labs(多利贝实验室) Adobe Research(Adobe研究)

专题命中 评测与基准 :foundation model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.11257 2025-07-31 cs.HC cs.CL cs.CV 57%

UI-E2I-Synth: Advancing GUI Grounding with Large-Scale Instruction Synthesis

Xinyi Liu, Xiaoyi Zhang, Ziyun Zhang, Yan Lu

专题命中 评测与基准 :language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22075 2025-07-31 cs.LG 57%

Prototype-Guided Pseudo-Labeling with Neighborhood-Aware Consistency for Unsupervised Adaptation

Eman Ali, Chetan Arora, Muhammad Haris Khan

机构 * Mohamed Bin Zayed University of Artificial Intelligence(莫扎德·本·扎耶德人工智能大学) Indian Institute of Technology Delhi(印度德里理工学院) Alexandria University(亚历山大大学)

专题命中 评测与基准 :language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22617 2025-07-31 cs.CR cs.CV 50%

Hate in Plain Sight: On the Risks of Moderating AI-Generated Hateful Illusions

Yiting Qu, Ziqing Yang, Yihan Ma, Michael Backes, Savvas Zannettou, Yang Zhang

机构 * CISPA Helmholtz Center for Information Security(CISPA 欧洲信息安全研究中心) TU Delft(代尔夫特理工大学)

专题命中 评测与基准 :language model(abstract)

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22194 2025-07-31 cs.CV cs.RO 50%

Temporally Consistent Unsupervised Segmentation for Mobile Robot Perception

Christian Ellis, Maggie Wigness, Craig Lennon, Lance Fiondella

机构 * Oden Institute for Computational Engineering & Sciences, University of Texas at Austin(德纳学院,德克萨斯大学奥斯汀分校) DEVCOM Army Research Laboratory, Adelphi, MD, United States(陆军研究实验室,阿德菲,马里兰州,美国) Department of Electrical and Computer Engineering, University of Massachusetts Dartmouth(电气与计算机工程系,马萨诸塞大学达特茅斯分校)

专题命中 评测与基准 :foundation model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22101 2025-07-31 cs.CV 50%

AI in Agriculture: A Survey of Deep Learning Techniques for Crops, Fisheries and Livestock

Umair Nawaz, Muhammad Zaigham Zaheer, Fahad Shahbaz Khan, Hisham Cholakkal, Salman Khan, Rao Muhammad Anwer

机构 * MBZ University of AI(人工智能大学) CECS, Australian National University(计算机科学与工程系,澳大利亚国立大学) Computer Vision Laboratory, Linköping University(链接öping大学计算机视觉实验室)

专题命中 评测与基准 :foundation model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 效率与部署 23 篇

2507.22478 2025-07-31 cs.CL 93%

SLM-SQL: An Exploration of Small Language Models for Text-to-SQL

Lei Sheng, Shuai-Shuai Xu

机构 * Wuhan University of Technology(武汉理工大学) University of Science and Technology of China(中国科学技术大学)

专题命中 效率与部署 :language model(title,abstract);small language model(title,abstract);SLM(title,abstract);large language model(abstract)

Comments 16 pages, 2 figures, work in progress

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22711 2025-07-31 cs.NI cs.AI 89%

OFCnetLLM: Large Language Model for Network Monitoring and Alertness

Hong-Jun Yoon, Mariam Kiran, Danial Ebling, Joe Breen

机构 * Oak Ridge National Laboratory(奥克兰国家实验室) Utah Education Network(犹他州教育网络) University of Utah(犹他大学)

专题命中 效率与部署 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.19442 2025-07-31 cs.AI cs.DC 89%

A Survey on Large Language Model Acceleration based on KV Cache Management

Haoyang Li, Yiming Li, Anxin Tian, Tianhao Tang, Zhanchao Xu, Xuejia Chen, Nicole Hu, Wei Dong, Qing Li, Lei Chen

机构 * Department of Computing, The Hong Kong Polytechnic University(计算系,香港理工大学) Department of Computer Science and Engineering(计算机科学与工程系) Department of Computer Science and Technology(计算机科学与技术系) Department of Computing and Data Science(计算与数据科学系)

专题命中 效率与部署 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

Comments Accepted to TMLR 2025. The revised version incorporates more papers and has been further polished

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20984 2025-07-31 cs.LG cs.AI 88%

SmallThinker: A Family of Efficient Large Language Models Natively Trained for Local Deployment

Yixin Song, Zhenliang Xue, Dongliang Wei, Feiyang Chen, Jianxiang Gao, Junchen Liu, Hangyu Liang, Guangshuo Qin, Chengrong Tian, Bo Wen, Longyu Zhao, Xinrui Zheng, Zeyu Mi, Haibo Chen

机构 * Institute of Parallel and Distributed Systems, Shanghai Jiao Tong University(并行与分布式系统研究所,上海交通大学) School of Artificial Intelligence, Shanghai Jiao Tong University(人工智能学院,上海交通大学) Zenergize AI

专题命中 效率与部署 :large language model(title,abstract);language model(title,abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.13932 2025-07-31 cs.LG cs.CL 88%

Enhancing Ultra-Low-Bit Quantization of Large Language Models Through Saliency-Aware Partial Retraining

Deyu Cao, Samin Aref

机构 * The University of Tokyo(东京大学) Department of Mechanical and Industrial Engineering, University of Toronto(多伦多大学机械与工业工程系)

专题命中 效率与部署 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.LG

Comments This is a post-peer-review accepted manuscript from the proceedings of the 22nd International Conference on Modeling Decisions for Artificial Intelligence (MDAI'25). The publisher authenticated version and full citation details are available on Springer's website (LNAI 15957). https://doi.org/10.1007/978-3-032-00891-6_28

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22065 2025-07-31 cs.SE cs.AI cs.CR cs.PL 88%

Fuzzing: Randomness? Reasoning! Efficient Directed Fuzzing via Large Language Models

Xiaotao Feng, Xiaogang Zhu, Kun Hu, Jincheng Wang, Yingjie Cao, Guang Gong, Jianfeng Pan

机构 * Security Technology Inc.(360安全技术公司) School of Science Edith Cowan University(科学学院,埃迪斯科文大学)

专题命中 效率与部署 :large language model(title,abstract);language model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22326 2025-07-31 cs.AI 85%

An Explainable Emotion Alignment Framework for LLM-Empowered Agent in Metaverse Service Ecosystem

Qun Ma, Xiao Xue, Ming Zhang, Yifan Shen, Zihan Zhao

机构 * College of Intelligence and Computing(智能与计算学院) Tianjin University(天津大学) Tianjin Key Laboratory of Healhy Habitat and Smart Technology(天津健康人居环境与智能技术重点实验室) Laboratory of Computation and Analytics of Complex Management Systems(复杂管理系统计算与分析实验室) Faculty of Environment, Science and Economy(环境、科学与经济学院)

专题命中 效率与部署 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.07052 2025-07-31 cs.CL cs.AI cs.LG 85%

Towards the Law of Capacity Gap in Distilling Language Models

Chen Zhang, Qiuchi Li, Dawei Song, Zheyu Ye, Yan Gao, Yan Hu

机构 * Beijing Institute of Technology(北京理工大学) University of Copenhagen(哥本哈根大学) The Open University(开放大学) Xiaohongshu(小红书)

专题命中 效率与部署 :language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI、cs.LG

Comments 32 pages, 10 figures, 15 tables, accepted to ACL 2025. Code and checkpoints are available at https://github.com/GeneZC/MiniMA

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.06680 2025-07-31 cs.CV cs.AI cs.LG cs.RO 84%

Application of Vision-Language Model to Pedestrians Behavior and Scene Understanding in Autonomous Driving

Haoxiang Gao, Li Zhang, Yu Zhao, Zhou Yang, Jinghan Cao

机构 * ECE Department(电子工程系) Carnegie Mellon University(卡内基梅隆大学) Computer Science Department(计算机科学系) Columbia University(哥伦比亚大学) Rotman School of Management(罗特曼管理学院) University of Toronto(多伦多大学) Department of Statistics(统计学系) George Washington University(乔治华盛顿大学) Department of Computer Science(计算机科学系) San Francisco State University(旧金山州立大学)

专题命中 效率与部署 :language model(title,abstract);foundation model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22448 2025-07-31 cs.CL 83%

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance

Jingwei Zuo, Maksim Velikanov, Ilyas Chahed, Younes Belkada, Dhia Eddine Rhayem, Guillaume Kunsch, Hakim Hacid, Hamza Yous, Brahim Farhat, Ibrahim Khadraoui, Mugariya Farooq, Giulia Campesan, Ruxandra Cojocaru, Yasser Djilali, Shi Hu, Iheb Chaabane, Puneesh Khanna, Mohamed El Amine Seddik, Ngoc Dung Huynh, Phuc Le Khac, Leen AlQadi, Billel Mokeddem, Mohamed Chami, Abdalgader Abubaker, Mikhail Lubinets, Kacper Piskorski, Slim Frikha

机构 * Falcon LLM Team(Falcon LLM团队)

专题命中 效率与部署 :language model(title,abstract);large language model(abstract);分类 cs.CL

Comments Technical report of Falcon-H1 model series

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22086 2025-07-31 cs.SE cs.AI cs.PL 83%

TypyBench: Evaluating LLM Type Inference for Untyped Python Repositories

Honghua Dong, Jiacheng Yang, Xun Deng, Yuhe Jiang, Gennady Pekhimenko, Fan Long, Xujie Si

机构 * University of Toronto(多伦多大学) Vector Institute(向量研究所) CIFAR AI Chair(CIFAR人工智能主席)

专题命中 效率与部署 :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

Journal ref Proceedings of the 42nd International Conference on Machine Learning, Vancouver, Canada. PMLR 267, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22352 2025-07-31 cs.HC 82%

Mitigating Response Delays in Free-Form Conversations with LLM-powered Intelligent Virtual Agents

Mykola Maslych, Mohammadreza Katebi, Christopher Lee, Yahya Hmaiti, Amirpouya Ghasemaghaei, Christian Pumarada, Janneese Palmer, Esteban Segarra Martinez, Marco Emporio, Warren Snipes, Ryan P. McMahan, Joseph J. LaViola

专题命中 效率与部署 :LLM(title);large language model(abstract);language model(abstract)

Comments 15 pages, 8 figures. Published at the 7th ACM Conference on Conversational User Interfaces (CUI '25), July 8-10, 2025, Waterloo, Canada. Open-source code available at https://github.com/ISUE/iva-cui

Journal ref Proceedings of the 7th ACM Conference on Conversational User Interfaces (CUI '25), 2025, Article 49, 1-15

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.12920 2025-07-31 cs.LG stat.ML 79%

Lightweight Online Adaption for Time Series Foundation Model Forecasts

Thomas L. Lee, William Toner, Rajkarn Singh, Artjom Joosen, Martin Asenov

机构 * intern(实习生) Huawei SIR Lab, Edinburgh Research Centre, UK(华为SIR实验室,爱丁堡研究中心,英国) School of Informatics, University of Edinburgh, UK(信息学院,爱丁堡大学,英国)

专题命中 效率与部署 :foundation model(title,abstract);分类 cs.LG

Comments 9 pages, Published at ICML 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.09509 2025-07-31 cs.CV 78%

ViM-VQ: Efficient Post-Training Vector Quantization for Visual Mamba

Juncan Deng, Shuaiting Li, Zeyu Wang, Kedong Xu, Hong Gu, Kejie Huang

机构 * Zhejiang University(浙江大学) vivo Mobile Communication Co., Ltd(vivo移动通信有限公司)

专题命中 效率与部署 :post-training(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22853 2025-07-31 cs.SE cs.AI 77%

Repair-R1: Better Test Before Repair

Haichuan Hu, Xiaochen Xie, Quanjun Zhang

机构 * Nanjing University of Science and Technology(南京理工大学)

专题命中 效率与部署 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22565 2025-07-31 cs.LG cs.AI cs.CL 75%

Efficient Differentially Private Fine-Tuning of LLMs via Reinforcement Learning

Afshin Khadangi, Amir Sartipi, Igor Tchappi, Ramin Bahmani, Gilbert Fridgen

机构 * SnT, University of Luxembourg(卢森堡大学SnT学院)

专题命中 效率与部署 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.08771 2025-07-31 cs.LG cs.CL 73%

BlockFFN: Towards End-Side Acceleration-Friendly Mixture-of-Experts with Chunk-Level Activation Sparsity

Chenyang Song, Weilin Zhao, Xu Han, Chaojun Xiao, Yingfa Chen, Yuxuan Li, Zhiyuan Liu, Maosong Sun

机构 * Dept. of Comp. Sci. & Tech., Institute for AI, Tsinghua University, Beijing, China(计算机科学与技术系,人工智能研究院,清华大学,北京,中国)

专题命中 效率与部署 :large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

Comments 21 pages, 7 figures, 15 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22358 2025-07-31 cs.AI cs.HC 70%

Magentic-UI: Towards Human-in-the-loop Agentic Systems

Hussein Mozannar, Gagan Bansal, Cheng Tan, Adam Fourney, Victor Dibia, Jingya Chen, Jack Gerrits, Tyler Payne, Matheus Kunzler Maldaner, Madeleine Grunde-McLaughlin, Eric Zhu, Griffin Bassman, Jacob Alber, Peter Chang, Ricky Loynd, Friederike Niedtner, Ece Kamar, Maya Murad, Rafah Hosn, Saleema Amershi

机构 * Microsoft Research AI Frontiers(微软研究院人工智能前沿)

专题命中 效率与部署 :large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19741 2025-07-31 cs.CL 70%

Basic Reading Distillation

Zhi Zhou, Sirui Miao, Xiangyu Duan, Hao Yang, Min Zhang

机构 * School of Computer Science and Technology, Soochow University(苏州大学计算机科学与技术学院) Huawei Translation Services Center(华为翻译服务中心)

专题命中 效率与部署 :large language model(abstract);language model(abstract);分类 cs.CL

Comments Accepted by ACL2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.14136 2025-07-31 cs.LG cs.AI 62%

Local Mixtures of Experts: Essentially Free Test-Time Training via Model Merging

Ryo Bertolissi, Jonas Hübotter, Ido Hakimi, Andreas Krause

机构 * ETH Zürich(苏黎世联邦理工学院)

专题命中 效率与部署 :language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏