arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-09-09 至 2025-09-09 共收录 262 信号源:cs.CL, cs.AI, cs.LG

1. 评测与基准 62 篇

2312.08334 2025-09-09 cs.CV 67%

LD-SDM: Language-Driven Hierarchical Species Distribution Modeling

Srikumar Sastry, Xin Xing, Aayush Dhakal, Subash Khanal, Adeel Ahmad, Nathan Jacobs

机构 * Washington University in St. Louis(圣路易斯华盛顿大学) University of Nebraska Omaha(内布拉斯加大学奥马哈分校)

专题命中 评测与基准 :large language model(abstract);language model(abstract)

Comments Accepted at Computer Vision for Ecology (CV4E) Workshop, ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.05440 2025-09-09 cs.CL cs.AI cs.LG 67%

Direct-Scoring NLG Evaluators Can Use Pairwise Comparisons Too

Logan Lawrence, Ashton Williamson, Alexander Shelton

专题命中 评测与基准 :language model(abstract);分类 cs.CL、cs.AI、cs.LG

Comments 12 pages, 18 tables, 1 figure

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.06550 2025-09-09 cs.LG cs.AI cs.CR cs.NI 62%

Contrastive Self-Supervised Network Intrusion Detection using Augmented Negative Pairs

Jack Wilkie, Hanan Hindy, Christos Tachtatzis, Robert Atkinson

机构 * University of Strathclyde(斯特拉思克莱德大学) Ain Shams University(爱因夏姆大学)

专题命中 评测与基准 :pretraining(abstract);分类 cs.AI、cs.LG

Comments Published in: Proceedings of IEEE Conference on Cyber Security and Resilience (CSR), 2025. Official version: https://doi.org/10.1109/CSR64739.2025.11129979 Code: https://github.com/jackwilkie/CLAN

Journal ref 2025 IEEE International Conference on Cyber Security and Resilience (CSR), Chania, Crete, Greece, 2025, pp. 206-213

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.08828 2025-09-09 cs.CL cs.AI cs.CY 62%

Human-AI Collaboration or Academic Misconduct? Measuring AI Use in Student Writing Through Stylometric Evidence

Eduardo Araujo Oliveira, Madhavi Mohoni, Sonsoles López-Pernas, Mohammed Saqr

专题命中 评测与基准 :LLM(abstract);分类 cs.CL、cs.AI

Comments 19 pages, 10 figures, 11 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.02100 2025-09-09 cs.HC cs.CL 57%

E-THER: A Multimodal Dataset for Empathic AI -- Towards Emotional Mismatch Awareness

Sharjeel Tahir, Judith Johnson, Jumana Abu-Khalaf, Syed Afaq Ali Shah

机构 * Centre for AI and ML, Edith Cowan University(人工智能与机器学习中心,埃德温·科温大学) University of Manchester(曼彻斯特大学)

专题命中 评测与基准 :language model(abstract);分类 cs.CL

Comments 15 pages, 4 figures. Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.18658 2025-09-09 cs.HC cs.AI cs.SE 57%

Assistance or Disruption? Exploring and Evaluating the Design and Trade-offs of Proactive AI Programming Support

Kevin Pu, Daniel Lazaro, Ian Arawjo, Haijun Xia, Ziang Xiao, Tovi Grossman, Yan Chen

机构 * University of Toronto(多伦多大学) University of California San Diego(加州大学圣地亚哥分校) Johns Hopkins University(约翰霍普金斯大学)

专题命中 评测与基准 :LLM(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.05566 2025-09-09 cs.CL cs.CY 57%

Ad hoc conventions generalize to new referents

Anya Ji, Claire Augusta Bergey, Ron Eliav, Yoav Artzi, Robert D. Hawkins

专题命中 评测与基准 :language agent(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.13394 2025-09-09 cs.LG math.ST stat.ML stat.TH 57%

Flow-based generative models as iterative algorithms in probability space

Yao Xie, Xiuyuan Cheng

机构 * H. Milton Stewart School of Industrial and Systems Engineering (ISyE) at the Georgia Institute of Technology(佐治亚理工学院H.米尔顿·斯图尔特工业与系统工程学院) Mathematics Department at Duke University(杜克大学数学系)

专题命中 评测与基准 :language model(abstract);分类 cs.LG

Comments IEEE Signal Processing Magazine, Special Issue on The Mathematics of Deep Learning, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.11163 2025-09-09 cs.CV cs.CL 57%

AI Sees Your Location, But With A Bias Toward The Wealthy World

Jingyuan Huang, Jen-tse Huang, Ziyi Liu, Xiaoyuan Liu, Wenxuan Wang, Jieyu Zhao

机构 * University of Southern California(南加州大学) University of Georgia(佐治亚大学) University of California, Los Angeles(加州大学洛杉矶分校)

专题命中 评测与基准 :language model(abstract);分类 cs.CL

Comments Accepted to EMNLP 2025 (Main)

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.06056 2025-09-09 cs.CL 57%

Synth-SBDH: A Synthetic Dataset of Social and Behavioral Determinants of Health for Clinical Text

Avijit Mitra, Zhichao Yang, Emily Druhl, Raelene Goodwin, Hong Yu

机构 * Manning College of Information and Computer Sciences, University of Massachusetts Amherst(马萨诸塞大学阿姆赫斯特分校信息与计算机科学学院) U.S. Department of Veterans Affairs(美国退伍军人事务部) Department of Medicine, University of Massachusetts Chan Medical School(马萨诸塞大学查恩医学院医学部) Miner School of Computer and Information Sciences, University of Massachusetts Lowell(马萨诸塞大学洛厄尔分校米纳尔计算机与信息科学学院)

专题命中 评测与基准 :LLM(abstract);分类 cs.CL

Comments Accepted at EMNLP 2025 (main) Github: https://github.com/avipartho/Synth-SBDH

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.10993 2025-09-09 eess.IV cs.CV 50%

Content Generation Models in Computational Pathology: A Comprehensive Survey on Methods, Applications, and Challenges

Yuan Zhang, Xinfeng Zhang, Xiaoming Qi, Xinyu Wu, Feng Chen, Guanyu Yang, Huazhu Fu

机构 * Key Laboratory of New Generation Artificial Intelligence Technology and Its Interdisciplinary Applications (Southeast University), Ministry of Education, Nanjing, China(新型一代人工智能技术及其交叉应用重点实验室(东南大学),教育部,南京,中国) School of Biomedical Engineering, Tsinghua University(生物医学工程学院,清华大学) Department of Biomedical Engineering and Department of Electrical and Computer Engineering, National University of Singapore(生物医学工程系和电子与计算机工程系,新加坡国立大学) School of Information Science and Engineering, Southeast University(信息科学与工程学院,东南大学) Department of Biostatistics, Center for Global Health, School of Public Health, Nanjing Medical University(流行病学系,全球健康中心,公共卫生学院,南京医科大学) Institute of High-Performance Computing, Agency for Science, Technology and Research, Singapore(高性能计算研究所,科技研究局,新加坡)

专题命中 评测与基准 :language model(abstract)

Comments 20 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.12654 2025-09-09 quant-ph 50%

Gaussian decomposition of magic states for matchgate computations

Joshua Cudby, Sergii Strelchuk

专题命中 评测与基准 :prompting(abstract)

Comments See also related works by Dias and Koenig and by Reardon-Smith et al. appearing in the same arXiv listing

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.05539 2025-09-09 q-bio.GN 50%

Investigating DNA words and their distributions across the tree of life

Charalampos Koilakos, Kimonas Provatas, Michail Patsakis, Aris Karatzikos, Ilias Georgakopoulos-Soares

专题命中 评测与基准 :language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.17771 2025-09-09 cs.CV 50%

Diagram-Driven Course Questions Generation

Xinyu Zhang, Lingling Zhang, Yanrui Wu, Muye Huang, Wenjun Wu, Bo Li, Shaowei Wang, Basura Fernando, Jun Liu

机构 * School of Computer Science and Technology, Xi’an Jiaotong University(西安交通大学计算机科学与技术学院) Ministry of Education Key Laboratory of Intelligent Networks and Network Security(教育部智能网络与网络安全重点实验室) Shaanxi Province Key Laboratory of Big Data Knowledge Engineering(陕西省大数据知识工程重点实验室) IHPC, Agency for Science, Technology and Research(科技研究局IHPC) College of Computing and Data Science, Nanyang Technological University(南洋理工大学计算与数据科学学院)

专题命中 评测与基准 :language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.04180 2025-09-09 cs.CV 50%

Slice-100K: A Multimodal Dataset for Extrusion-based 3D Printing

Anushrut Jignasu, Kelly O. Marshall, Ankush Kumar Mishra, Lucas Nerone Rillo, Baskar Ganapathysubramanian, Aditya Balu, Chinmay Hegde, Adarsh Krishnamurthy

机构 * Iowa State University(爱荷华州立大学) New York University(纽约大学)

专题命中 评测与基准 :foundation model(abstract)

Comments Accepted to NeurIPS 2024. For codebase, see https://github.com/idealab-isu/Slice-100K

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 效率与部署 39 篇

2509.05320 2025-09-09 cs.CR cs.LG 89%

Privacy-Preserving Offloading for Large Language Models in 6G Vehicular Networks

Ikhlasse Badidi, Nouhaila El Khiyaoui, Aya Riany, Badr Ben Elallid, Amine Abouaomar

机构 * School of Science and Engineering, Al Akhawayn University in Ifrane(科学与工程学院,阿尔阿赫韦因大学) Department of Electrical and Computer Engineering, Université du Québec à Trois-Rivières(电气与计算机工程系,魁北克大学Trois-Rivières分校)

专题命中 效率与部署 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.LG

Comments 7 pages, 6 figures, 1 algorithm, 5 equations

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.07831 2025-09-09 cs.LG 88%

ALPS: Improved Optimization for Highly Sparse One-Shot Pruning for Large Language Models

Xiang Meng, Kayhan Behdin, Haoyue Wang, Rahul Mazumder

专题命中 效率与部署 :large language model(title,abstract);language model(title,abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.05852 2025-09-09 stat.ML cs.LG math.ST stat.TH 88%

Fisher Random Walk: Automatic Debiasing Contextual Preference Inference for Large Language Model Evaluation

Yichi Zhang, Alexander Belloni, Ethan X. Fang, Junwei Lu, Xiaoan Xu

机构 * Department of Statistics, Indiana University Bloomington(印第安纳大学布卢明顿分校统计学系) Fuqua School of Business, Duke University(杜克大学福克商学院) Department of Biostatistics & Bioinformatics, Duke University(杜克大学生物统计学与生物信息学系) Department of Biostatistics, Harvard T.H. Chan School of Public Health(哈佛大学T.H. Chan公共卫生学院生物统计学系)

专题命中 效率与部署 :large language model(title,abstract);language model(title,abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.01190 2025-09-09 cs.CL 88%

Efficient Large Language Models with Zero-Shot Adjustable Acceleration

Sajjad Kachuee, Mohammad Sharifkhani

专题命中 效率与部署 :large language model(title,abstract);language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.05936 2025-09-09 cs.NI cs.LG 85%

ALPHA: LLM-Enabled Active Learning for Human-Free Network Anomaly Detection

Xuanhao Luo, Shivesh Madan Nath Jha, Akruti Sinha, Zhizhen Li, Yuchen Liu

机构 * North Carolina State University, USA(北卡罗来纳州立大学)

专题命中 效率与部署 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.LG

Comments Accepted at 44th IEEE International Performance Computing and Communications Conference (IPCCC 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.15006 2025-09-09 cs.AR cs.AI cs.DC cs.ET cs.PF 85%

Scaling Intelligence: Designing Data Centers for Next-Gen Language Models

Jesmin Jahan Tithi, Hanjiang Wu, Avishaii Abuhatzera, Fabrizio Petrini

机构 * Georgia Institute of Technology(佐治亚理工学院)

专题命中 效率与部署 :language model(title,abstract);LLM(abstract);large language model(abstract);分类 cs.AI

Comments 14 pages, submitted to SC25 for review

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.05263 2025-09-09 cs.AI cs.CV cs.LG 85%

LatticeWorld: A Multimodal Large Language Model-Empowered Framework for Interactive Complex World Generation

Yinglin Duan, Zhengxia Zou, Tongwei Gu, Wei Jia, Zhan Zhao, Luyi Xu, Xinzhu Liu, Yenan Lin, Hao Jiang, Kang Chen, Shuang Qiu

机构 * NetEase, Inc.(网易公司) Beihang University(北京航空航天大学) Tsinghua University(清华大学) City University of Hong Kong(香港城市大学) Independent Researcher & Technical Artists(独立研究者及技术艺术家)

专题命中 效率与部署 :large language model(title);language model(title);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.05584 2025-09-09 cs.LG cs.CV cs.PF 84%

ProfilingAgent: Profiling-Guided Agentic Reasoning for Adaptive Model Optimization

Sadegh Jafari, Aishwarya Sarkar, Mohiuddin Bilwal, Ali Jannesari

机构 * Department of Computer Science(计算机科学系) Iowa State University(爱荷华州立大学)

专题命中 效率与部署 :LLM(abstract);large language model(abstract);language model(abstract);foundation model(abstract)

Comments 13 pages, 3 figures, 5 tables, 1 algorithm

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.06768 2025-09-09 cs.RO 82%

Embodied Hazard Mitigation using Vision-Language Models for Autonomous Mobile Robots

Oluwadamilola Sotomi, Devika Kodi, Kiruthiga Chandra Shekar, Aliasghar Arab

机构 * Department of Mechanical and Aerospace Engineering, Tandon School of Engineering, New York University(机械与航空航天工程系,坦顿工程学院,纽约大学) GenAuto.ai by General Autonomy Inc.(General Autonomy Inc. 的 GenAuto.ai)

专题命中 效率与部署 :language model(title,abstract);large language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.06321 2025-09-09 cs.CV 82%

Text4Seg++: Advancing Image Segmentation via Generative Language Modeling

Mengcheng Lan, Chaofeng Chen, Jiaxing Xu, Zongrui Li, Yiping Ke, Xudong Jiang, Yingchen Yu, Yunqing Zhao, Song Bai

机构 * College of Computing and Data Science, Nanyang Technological University(computing and Data Science学院,南洋理工大学) School of Electrical and Electronic Engineering, Nanyang Technological University(Electrical and Electronic Engineering学院,南洋理工大学) School of Artificial Intelligence, Wuhan University(Artificial Intelligence学院,武汉大学) ByteDance(字节跳动)

专题命中 效率与部署 :language model(title,abstract);large language model(abstract)

Comments Extended version of our conference paper arXiv:2410.09855

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.05962 2025-09-09 cs.HC 82%

The Reel Deal: Designing and Evaluating LLM-Generated Short-Form Educational Videos

Lazaros Stavrinou, Argyris Constantinides, Marios Belk, Vasos Vassiliou, Fotis Liarokapis, Marios Constantinides

专题命中 效率与部署 :LLM(title);large language model(abstract);language model(abstract)

Comments 9 pages, 3 figures, 1 table

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.05925 2025-09-09 cs.CV cs.IT math.IT 82%

Compression Beyond Pixels: Semantic Compression with Multimodal Foundation Models

Ruiqi Shen, Haotian Wu, Wenjing Zhang, Jiangjing Hu, Deniz Gunduz

机构 * Department of Electrical and Electronic Engineering, Imperial College London(帝国理工学院电子与电气工程系)

专题命中 效率与部署 :foundation model(title,abstract);pretraining(abstract)

Comments Published as a conference paper at IEEE 35th Workshop on Machine Learning for Signal Processing (MLSP)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.05355 2025-09-09 cs.RO cs.MA 82%

Human-LLM Synergy in Context-Aware Adaptive Architecture for Scalable Drone Swarm Operation

Ahmed R. Sadik, Muhammad Ashfaq, Niko Mäkitalo, Tommi Mikkonen

机构 * Honda Research Institute Europe, Germany(本田欧洲研究机构)

专题命中 效率与部署 :LLM(title);large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.06631 2025-09-09 cs.CL 81%

Guided Decoding and Its Critical Role in Retrieval-Augmented Generation

Özgür Uğur, Musa Yılmaz, Esra Şavirdi, Özay Ezerceli, Mahmut El Huseyni, Selva Taş, Reyhan Bayraktar

机构 * Newmind AI Istanbul, Türkiye(新 mind AI 伊斯坦布尔,土耳其)

专题命中 效率与部署 :LLM(abstract);large language model(abstract);language model(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.06518 2025-09-09 cs.CL cs.AI 81%

Crown, Frame, Reverse: Layer-Wise Scaling Variants for LLM Pre-Training

Andrei Baroian, Kasper Notebomer

机构 * LIACS, Leiden University, The Netherlands(LIACS,莱顿大学,荷兰)

专题命中 效率与部署 :LLM(title);language model(abstract);分类 cs.CL、cs.AI

Comments The reported results are skewed due to a data type mismatch. The dataset was saved with int32, but the data loader interpreted it as uint16. As a result, each 32-bit token was incorrectly split into two 16-bit tokens. Outcome: a consistent artifact where every other token is zero

详情

展开后加载摘要…

URL PDF HTML 收藏