arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12266 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12266 篇

2507.19929 2025-12-16 physics.soc-ph cs.AI 70%

DynamiX: Large-Scale Dynamic Social Network Simulator

DynamiX: 大规模动态社交网络模拟器

Yanhui Sun, Wu Liu, Wentao Wang, Hantao Yao, Jiebo Luo, Yongdong Zhang

机构 * School of Information Science and Technology, University of Science and Technology of China(信息科学与技术学院,中国科学技术大学) Department of Computer Science, University of Rochester(计算机科学系,罗切斯特大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 DynamiX通过动态层级模块和不同用户类型的社交关系建模策略,提升了大规模动态社交网络模拟的准确性与实用性。

Comments Social and Information Networks

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.12921 2025-12-16 cs.CR cs.AI 70%

Cisco Integrated AI Security and Safety Framework Report

思科集成AI安全与安全框架报告

Amy Chang, Tiffany Saade, Sanket Mendapara, Adam Swanda, Ankit Garg

机构 * Cisco AI Threat and Security Research(思科人工智能威胁与安全研究)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出思科集成AI安全与安全框架,旨在统一分类和操作化AI风险,涵盖安全与安全,适用于威胁识别、风险优先级排序等,具有全面性和扩展性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.08464 2025-12-12 cs.AI cond-mat.mtrl-sci 70%

A Generation Framework with Strict Constraints for Crystal Materials Design

具有严格约束的晶体材料设计生成框架

Chao Huang, Jiahui Chen, Chen Chen, Chen Chen, Chunyan Chen, Renjie Su, Shiyu Du

机构 * Institute of Computing Technology(计算技术研究所) Chinese Academy of Science(中国科学院) Ningbo Institute of Artificial Intelligence Industry(宁波人工智能产业研究所) Hangzhou Institute for Advanced Study(杭州高级研究所) UCAS China University of Petroleum (East China)(中国石油大学(华东)) Ningbo Institute of Materials Technology and Engineering(宁波材料技术与工程研究所)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出一种具有严格约束的晶体材料设计生成框架,通过约束生成器和结构生成器生成满足特定化学和物理性质的晶体结构,提高生成效率并确保化学组成准确性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.17863 2025-12-11 cs.LG cs.NE 70%

The emergence of sparse attention: impact of data distribution and benefits of repetition

稀疏注意力的出现:数据分布的影响与重复性的益处

Nicolas Zucchet, Francesco d'Angelo, Andrew K. Lampinen, Stephanie C. Y. Chan

机构 * ETH Zürich(苏黎世联邦理工学院) EPFL(苏黎世联邦理工学院) Google DeepMind(谷歌DeepMind)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.LG

AI总结 本文研究了稀疏注意力在训练过程中的涌现机制,揭示其与任务结构、架构和优化器选择的关系,并发现重复能加速这一过程。

Comments NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.15771 2025-12-10 cs.CY cs.AI econ.GN q-fin.EC 70%

Left Leaning Models: How AI Evaluates Economic Policy?

左倾模型:人工智能如何评估经济政策?

Maxim Chupilkin

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文研究了AI在多因素约束下对经济政策的偏好,发现LLMs普遍偏好高增长、低失业和低不平等,而非传统宏观经济目标。

Comments 16 pages, 2 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.03295 2025-12-10 cs.AI cs.RO cs.SE 70%

Capability-Driven Skill Generation with LLMs: A RAG-Based Approach for Reusing Existing Libraries and Interfaces

基于能力驱动的技能生成:一种基于RAG的方法,用于重用现有库和接口

Luis Miguel Vieira da Silva, Aljosha Köcher, Nicolas König, Felix Gehlhoff, Alexander Fay

机构 * Ruhr University, Bochum, Germany(鲁尔大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出一种基于RAG的方法,利用大语言模型生成可执行代码,通过整合现有库和接口实现能力驱动的技能生成。

Comments \c{opyright} 2025 IEEE. Personal use of this material is permitted. Permission from IEEE must be obtained for all other uses, in any current or future media, including reprinting/republishing this material for advertising or promotional purposes, creating new collective works, for resale or redistribution to servers or lists, or reuse of any copyrighted component of this work in other works

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06042 2025-12-09 cs.SE cs.AI 70%

Auto-SPT: Automating Semantic Preserving Transformations for Code

Auto-SPT:用于代码的自动化语义保持变换

Ashish Hooda, Mihai Christodorescu, Chuangang Ren, Aaron Wilson, Kassem Fawaz, Somesh Jha

机构 * Google(谷歌) Google, U. Wisconsin–Madison(谷歌,威斯康星大学麦迪逊分校) U. Wisconsin–Madison, Google(威斯康星大学麦迪逊分校,谷歌)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 Auto-SPT通过生成多样化的语义保持变换来提升代码克隆检测模型的鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04987 2025-12-05 cs.CL 70%

Nex-N1: Agentic Models Trained via a Unified Ecosystem for Large-Scale Environment Construction

Nex-N1: 通过统一生态系统训练的代理模型用于大规模环境构建

AGI Team, Yuxuan Cai, Lu Chen, Qiaoling Chen, Yuyang Ding, Liwen Fan, Wenjie Fu, Yufei Gao, Honglin Guo, Pinxue Guo, Zhenhua Han, Zhengfu He, Hanglei Hu, Kai Hu, Shengjia Hua, Tianyu Huai, Baodai Huang, Li Ji, Zhen Jiang, Zhikai Lei, Bufan Li, Jiahang Lin, Lizhi Lin, Jinxiu Liu, Shichun Liu, Ziming Liu, Yuchen Ni, Pengfang Qian, Yujiong Shen, Qingyun Shi, Wentao Shu, Peng Sun, Yiran Suo, Tian Tang, Boyu Tian, Guoteng Wang, Junzhe Wang, Peixin Wang, Zhiheng Xi, Hang Yan, Jie Yang, Zhixiong Yang, Tianchu Yao, Guangze Ye, Qianxi Yu, Shuo Zhang, Xinyue Zhang, Yiqi Zhang, Jiarong Zhao, Miao Zheng, Rui Zheng, Enyu Zhou, Jiazheng Zhou, Maosen Zhou, Yuhao Zhou, Tao Gui, Yining Zheng, Xinchi Chen, Jie Zhou, Siyuan Feng, Qin Chen, Liang He, Qi Zhang, Xuanjing Huang, Xipeng Qiu

机构 * AGI Team(AGI团队)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 Nex-N1通过统一生态系统训练的代理模型,实现了大规模环境构建,展示了在复杂代理任务中的优越性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.19550 2025-12-05 cs.AI 70%

Turing Test 2.0: The General Intelligence Threshold

Turing Test 2.0:通用智能阈值

Georgios Mappouras

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出图灵测试2.0框架,通过设定通用智能阈值来检测是否达到人工通用智能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.00673 2025-12-02 cs.CL 70%

A Comparison of Human and ChatGPT Classification Performance on Complex Social Media Data

对复杂社交媒体数据中人类与ChatGPT分类性能的比较

Breanna E. Green, Ashley L. Shea, Pengfei Zhao, Drew B. Margolin

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文比较了ChatGPT和人类在复杂社交媒体数据分类中的性能,发现GPT-4在处理微妙语言时存在困难,需谨慎使用。

Comments About 15 pages, draft version of accepted conference full paper. Published paper to follow

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.00392 2025-12-02 cs.CL 70%

A Taxonomy of Errors in English as she is spoke: Toward an AI-Based Method of Error Analysis for EFL Writing Instruction

英语口语中的错误分类:一种基于AI的EFL写作教学错误分析方法

Damian Heywood, Joseph Andrew Carrier, Kyu-Hong Hwang

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本研究提出一种基于AI的英语写作错误分析系统,利用大型语言模型对写作错误进行分类和纠正,展示AI在EFL教学中的潜力。

Comments Metadata at "Replication Data for: A Taxonomy of Errors in English as she is spoke: An AI-Based System for Error Analysis for EFL Writing Instruction", https://doi.org/10.7910/DVN/N5O7C4, Harvard Dataverse, V1

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.22777 2025-12-01 cs.RO cs.AI 70%

Improving Robotic Manipulation Robustness via NICE Scene Surgery

通过NICE场景手术提升机器人操作的鲁棒性

Sajjad Pakdamansavoji, Mozhgan Pourkeshavarz, Adam Sigal, Zhiyuan Li, Rui Heng Yang, Amir Rasouli

机构 * Huawei Technologies Canada(华为技术加拿大公司)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 NICE通过增强场景上下文减少分布外差距,提升机器人操作在复杂环境中的鲁棒性和安全性。

Comments 11 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.14522 2025-12-01 cs.HC cs.AI 70%

Biased by Design: Leveraging AI Biases to Enhance Critical Thinking of News Readers

设计偏见:利用AI偏见增强新闻读者的批判性思维

Liudmila Zavolokina, Kilian Sprenkamp, Zoya Katashinskaya, Daniel Gordon Jones

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出利用AI偏见提升新闻读者批判性思维的设计策略,通过用户选择和个性化方法增强信息辨别能力。

Comments European Conference on Information Systems (ECIS)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.22341 2025-12-01 cs.CV cs.LG 70%

Unexplored flaws in multiple-choice VQA evaluations

多选式视觉问答评估中的未探索缺陷

Fabio Rosenthal, Sebastian Schmidt, Thorsten Graf, Thorsten Bagodonat, Stephan Günnemann, Leo Schwinn

机构 * Technical University of Munich(慕尼黑技术大学) Volkswagen AG(大众集团) Munich Data Science Institute(慕尼黑数据科学研究所)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.LG

AI总结 本文揭示了多选式视觉问答评估中因提示格式变化导致的未探索偏差,指出其对MLLM评估结果的显著影响,并表明现有缓解策略无法有效应对这些新发现的偏见。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.21802 2025-12-01 cs.GT cs.AI econ.GN q-fin.EC 70%

Tacit Bidder-Side Collusion: Artificial Intelligence in Dynamic Auctions

隐性投标方合谋:动态拍卖中的人工智能

Sriram Tolety

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文研究了大型语言模型在动态拍卖中通过隐性协调实现投标方合谋的可能性,并展示了市场结构对缓解合谋的影响。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.20965 2025-11-27 cs.CV cs.CL 70%

TrafficLens: Multi-Camera Traffic Video Analysis Using LLMs

TrafficLens: 多摄像头交通视频分析使用LLMs

Md Adnan Arefeen, Biplob Debnath, Srimat Chakradhar

机构 * NEC Laboratories America(NEC美国实验室) University of Missouri–Kansas City (UMKC)(密苏里大学-堪萨斯城分校)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 TrafficLens通过利用多摄像头重叠区域和对象相似性检测,高效地将视频转换为文本,减少处理时间并提升交通视频分析效率。

Comments 2024 IEEE 27th International Conference on Intelligent Transportation Systems (ITSC)

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.17846 2025-11-27 cs.RO cs.AI 70%

Safety Control of Service Robots with LLMs and Embodied Knowledge Graphs

服务机器人中结合大语言模型和具身知识图谱的安全控制

Yong Qi, Gabriel Kyebambo, Siyuan Xie, Wei Shen, Shenghui Wang, Bitao Xie, Bin He, Zhipeng Wang, Shuo Jiang

机构 * School of Electronic Information(电子信息学院) Artificial Intelligence, Shaanxi University of Science(人工智能,陕西科技大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出结合大语言模型与具身知识图谱的方法,提升服务机器人安全控制,通过预定义指令和知识库验证增强安全实践。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00763 2025-11-25 cs.AI 70%

How Focused Are LLMs? A Quantitative Study via Repetitive Deterministic Prediction Tasks

LLM的聚焦性如何?通过重复确定性预测任务的定量研究

Wanda Hou, Leon Zhou, Hong-Ye Hu, Yubei Chen, Yi-Zhuang You, Xiao-Liang Qi

机构 * EdenCode Inc.(EdenCode公司) Stevenson School(斯蒂文森学校) Harvard University(哈佛大学) UC Davis(加州大学戴维斯分校) Path Integral Technology, Inc.(路径积分技术公司)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 研究通过重复确定性预测任务揭示LLM在长序列中的准确性下降现象,提出统计物理模型解释内部干扰与外部条件的竞争,建立误差积累机制。

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.18337 2025-11-25 eess.AS cs.AI cs.SD 70%

Warm Chat: Diffuse Emotion-aware Interactive Talking Head Avatar with Tree-Structured Guidance

Warm Chat: 一种具有树状引导的情感感知交互式对话头虚拟形象

Haijie Yang, Zhenyu Zhang, Hao Tang, Jianjun Qian, Jian Yang

机构 * Nanjing University of Science and Technology(南京理工大学) Nanjing University(南京大学) School of Computer Science, Peking University(北京大学计算机学院)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 Warm Chat提出一种基于树状结构引导的情感感知对话头生成框架,利用大语言模型生成时序一致且情感丰富的虚拟形象,提升双向交互的自然度与情感适应性。

Comments The submission is withdrawn at the request of the authors due to internal reasons within the research team

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.03283 2025-11-24 cs.SE cs.AI 70%

CATCODER: Repository-Level Code Generation with Relevant Code and Type Context

CATCODER: 基于相关代码和类型上下文的仓库级代码生成

Zhiyuan Pan, Xing Hu, Xin Xia, Xiaohu Yang

机构 * The State Key Laboratory of Blockchain and Data Security, Zhejiang University(区块链与数据安全国家重点实验室,浙江大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 CatCoder通过整合相关代码和类型上下文,提升仓库级代码生成的性能和可扩展性。

Comments Revised and extended version; To be published in ACM Transactions on Software Engineering and Methodology

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.17534 2025-11-21 cs.CV cs.CL cs.MM 70%

Co-Reinforcement Learning for Unified Multimodal Understanding and Generation

协同强化学习用于统一多模态理解和生成

Jingjing Jiang, Chongjie Si, Jun Luo, Hanwang Zhang, Chao Ma

机构 * Shanghai Jiao Tong University(上海交通大学) Nanyang Technological University(南洋理工大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文提出CoRL框架,通过协同强化学习提升多模态大语言模型在生成与理解任务上的性能。

Comments NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.15512 2025-11-20 cs.CL 70%

Standardising the NLP Workflow: A Framework for Reproducible Linguistic Analysis

标准化NLP工作流程:可重复语言分析的框架

Yves Pauli, Jan-Bernard Marsman, Finn Rabe, Victoria Edkins, Roya Hüppi, Silvia Ciampelli, Akhil Ratan Misra, Nils Lang, Wolfram Hinzen, Iris Sommer, Philipp Homan

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文提出LPDS数据结构和pelican nlp框架,旨在通过标准化流程提升语言数据处理的可重复性和透明度。

Comments 26 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.14930 2025-11-20 stat.AP cs.AI econ.GN q-fin.EC 70%

Fifty Shades of Greenwashing: The Political Economy of Climate Change Advertising on Social Media

Robert Kubinec, Aseem Mahajan

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

Comments Supplementary information can be downloaded at https://www.icloud.com/iclouddrive/00eqjqGpFLQ86sPSGGWRPuuhw#mahajan%5Fkubinec%5FSI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.02505 2025-11-20 cs.CV cs.AI 70%

ESA: Energy-Based Shot Assembly Optimization for Automatic Video Editing

Yaosen Chen, Wei Wang, Tianheng Zheng, Xuming Wen, Han Yang, Yanru Zhang

机构 * Sobey Media Intelligence Laboratory(索贝媒体智能实验室) University of Electronic Science and Technology of China(电子科学与技术大学) SiChuan University(四川大学) Qinghai Normal University(青海师范大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.12347 2025-11-18 eess.AS cs.CL cs.SD 70%

VoiceCraft-X: Unifying Multilingual, Voice-Cloning Speech Synthesis and Speech Editing

Zhisheng Zheng, Puyuan Peng, Anuj Diwan, Cong Phuoc Huynh, Xiaohang Sun, Zhu Liu, Vimal Bhat, David Harwath

机构 * University of Texas at Austin(德克萨斯大学奥斯汀分校) Amazon(亚马逊)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

Comments EMNLP 2025. Demo and code are available at https://zhishengzheng.com/voicecraft-x/

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25292 2025-11-18 cs.CY cs.AI 70%

A Measurement Study of Model Context Protocol Ecosystem

Hechuan Guo, Yongle Hao, Yue Zhang, Minghui Xu, Peizhuo Lv, Jiezhi Chen, Xiuzhen Cheng

机构 * Shandong University(山东大学) Nanyang Technological University(南洋理工大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.22162 2025-11-18 cs.CY cs.CL 70%

Surface Reading LLMs: Synthetic Text and its Styles

Hannes Bajohr

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

Comments 12 pages, 1 figure

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.11380 2025-11-17 cs.LG 70%

When Genes Speak: A Semantic-Guided Framework for Spatially Resolved Transcriptomics Data Clustering

Jiangkai Long, Yanran Zhu, Chang Tang, Kun Sun, Yuanyuan Liu, Xuesong Yan

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.LG

Comments AAAI'2026 poster paper. 12 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.11108 2025-11-17 cs.CL cs.CY 70%

Analysing Personal Attacks in U.S. Presidential Debates

Ruban Goyal, Rohitash Chandra, Sonit Singh

机构 * School of Computer Science and Engineering, University of New South Wales(计算机科学与工程学院,新南威尔士大学) School of Mathematics and Statistics, University of New South Wales(数学与统计学学院,新南威尔士大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

Comments 13 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.19238 2025-11-17 cs.AI cs.CY 70%

Designing AI-Agents with Personalities: A Psychometric Approach

Muhua Huang, Xijuan Zhang, Christopher Soto, James Evans

机构 * Stanford University(斯坦福大学) University of Chicago Knowledge Lab(芝加哥大学知识实验室) Chicago Center for Computational Social Science(芝加哥计算社会科学中心) York University(约克大学) Colby College(科尔比学院)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏