arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-12-25 至 2025-12-25 共收录 113 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 11 篇

2512.20997 2025-12-25 cs.NI 85%

LLM-Empowered Agentic AI for QoE-Aware Network Slicing Management in Industrial IoT

基于大语言模型的代理AI用于工业物联网中面向QoE的网络切片管理

Xudong Wang, Lei Feng, Ruichen Zhang, Fanqin Zhou, Hongyang Du, Wenjing Li, Dusit Niyato, Abbas Jamalipour, Ping Zhang

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract)

AI总结 本文提出一种基于大语言模型的代理AI方法,用于工业物联网中面向服务质量的网络切片管理,通过整合推理、规划和适应能力,提升网络切片的延迟、可靠性和成本效率。

Comments 8 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.01061 2025-12-25 cs.CY cs.AI cs.HC 81%

Epitome: Pioneering an Experimental Platform for AI-Social Science Integration

Epitome:开创人工智能与社会科学整合的实验平台

Jingjing Qu, Kejia Hu, Jun Zhu, Yulei Ye, Wenhao Li, Teng Wang, Zhiyun Chen, Chaochao Lu, Aimin Zhou, Xiangfeng Wang, Xia Hu, James Evans

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);foundation model(abstract)

AI总结 Epitome通过矩阵式社会世界,开创人工智能与社会科学整合的实验平台,研究人类与AI交互的社会动态及涌现特性。

Comments 25 pages, 6figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.11783 2025-12-25 eess.SP cs.AI cs.LG q-bio.NC 81%

EEG Foundation Models: A Critical Review of Current Progress and Future Directions

EEG基础模型:对当前进展和未来方向的批判性回顾

Gayal Kuruppu, Neeraj Wagh, Vaclav Kremen, Sandipan Pati, Gregory Worrell, Yogatheesan Varatharajah

机构 * Department of Computer Science & Engineering, University of Minnesota Twin Cities(计算机科学与工程系,明尼苏达大学双城分校) Department of Bioengineering, University of Illinois at Urbana-Champaign(生物工程系,伊利诺伊大学厄巴纳-香槟分校) Department of Neurology, Mayo Clinic(神经病学系,梅奥诊所) Department of Neurology, University of Minnesota Twin Cities(神经病学系,明尼苏达大学双城分校)

专题命中 领域大模型 :foundation model(title,abstract);分类 cs.AI、cs.LG

AI总结 本文回顾了EEG基础模型的当前进展,分析了其在自监督建模中的方法和评估策略,指出未来需加强可扩展性和可信度以提升实际应用价值。

Comments 22 pages (main), 5 figures (main), 4 tables (main + supplement)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.00903 2025-12-25 cs.CL cs.AI cs.CY cs.SI 81%

Embracing Dialectic Intersubjectivity: Coordination of Different Perspectives in Content Analysis with LLM Persona Simulation

拥抱辩证的主体间性:利用LLM人格模拟协调内容分析中的不同视角

Taewoo Kang, Kjerstin Thorson, Tai-Quan Peng, Dan Hiaeshutter-Rice, Sanguk Lee, Stuart Soroka

机构 * Department of Media and Information(媒体与信息系) Michigan State University(密歇根州立大学) College of Liberal Arts(人文学院) Colorado State University(科罗拉多州立大学) Department of Communication(传播系) Department of Advertising and Public Relations(广告与公共关系系) Department of Communication Studies(传播学系) Texas Christian University(德克萨斯 Christian 大学) Departments of Communication and Political Science(传播与政治学系) University of California, Los Angeles(加州大学洛杉矶分校)

专题命中 领域大模型 :LLM(title,abstract);分类 cs.CL、cs.AI

AI总结 本文通过LLM人格模拟协调内容分析中的不同视角,探讨了党派偏见对编码结果的影响,并提升了AI驱动的社会科学研究的严谨性。

Journal ref Social Science Computer Review, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.07201 2025-12-25 cs.CL cs.AI 81%

A Review on the Applications of Transformer-based language models for Nucleotide Sequence Analysis

基于变换器语言模型在核苷酸序列分析中应用的综述

Nimisha Ghosh, Daniele Santoni, Indrajit Saha, Giovanni Felici

机构 * Department of Computer Science and Information Technology, Institute of Technical Education and Research(计算机科学与信息技术系,技术教育与研究学院)

专题命中 领域大模型 :language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本文综述了基于变换器语言模型在核苷酸序列分析中的应用,分析了其主要特征和不同定制方法,并为初学者提供了变换器的工作原理描述。

Journal ref Computational and Structural Biotechnology Journal, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.20959 2025-12-25 cs.LG cs.AI stat.ME 73%

Can Agentic AI Match the Performance of Human Data Scientists?

代理AI能否匹配数据科学家的性能?

An Luo, Jin Du, Fangqiao Tian, Xun Xian, Robert Specht, Ganghua Wang, Xuan Bi, Charles Fleming, Jayanth Srinivasa, Ashish Kundu, Mingyi Hong, Jie Ding

机构 * School of Statistics, University of Minnesota(统计学系,明尼苏达大学) Department of Electrical and Computer Engineering, University of Minnesota(电气与计算机工程系,明尼苏达大学) Data Science Institute, University of Chicago(数据科学研究所,芝加哥大学) Carlson School of Management, University of Minnesota(卡尔森管理学院,明尼苏达大学) Cisco Research, San Jose, CA, USA(Cisco研究,加州圣何塞)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 研究探讨代理AI在处理隐藏潜在变量任务时的表现,发现其在缺乏领域知识时无法匹敌人类数据科学家。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.21701 2025-12-25 cs.CL cs.LG 73%

47B Mixture-of-Experts Beats 671B Dense Models on Chinese Medical Examinations

47B混合专家模型在中文医学考试中超越671B密集模型

Chiung-Yi Tseng, Danyang Zhang, Tianyang Wang, Hongying Luo, Lu Chen, Junming Huang, Jibin Guan, Junfeng Hao, Junhao Song, Xinyuan Song, Ziqian Bi

机构 * AI Agent Lab, Vokram Group(AI代理实验室,Vokram集团) Purdue University(普渡大学) University of Minnesota(明尼苏达大学) Imperial College London(伦敦帝国理工学院)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

AI总结 本文评估了多种大语言模型在中文医学考试中的表现,发现47B混合专家模型在准确率上超越了671B密集模型,揭示了模型大小与性能无直接关联,并指出了不同医学专科间性能差异及模型泛化能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.21021 2025-12-25 cs.IR cs.LG 57%

Towards Better Search with Domain-Aware Text Embeddings for C2C Marketplaces

面向C2C市场向更好搜索的领域感知文本嵌入

Andre Rusli, Miao Cao, Shoma Ishimoto, Sho Akiyama, Max Frenzel

专题命中 领域大模型 :LLM(abstract);分类 cs.LG

AI总结 本文提出领域感知的文本嵌入方法,通过优化C2C市场places的搜索质量,提升相关性和效率,为更丰富的LLM时代搜索体验奠定基础。

Comments 5 pages, AAAI 2026 Workshop on New Frontiers in Information Retrieval

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.20892 2025-12-25 cs.CV 50%

Beyond Weight Adaptation: Feature-Space Domain Injection for Cross-Modal Ship Re-Identification

超越权重适应:基于特征空间的域注入用于跨模态船舶重识别

Tingfeng Xian, Wenlve Zhou, Zhiheng Zhou, Zhelin Li

机构 * School of Electronic and Information Engineering, South China University of Technology(华南理工大学电子与信息学院) Key Laboratory of Big Data and Intelligent Robot, Ministry of Education, South China University of Technology(大数据与智能机器人教育部重点实验室)

专题命中 领域大模型 :foundation model(abstract)

AI总结 本文提出基于特征空间的域注入方法,解决跨模态船舶重识别中的模态差异问题,通过轻量级模型提升性能,实现SOTA效果。

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 知识编辑与模型理解 7 篇

2511.17129 2025-12-25 cs.CL cs.AI 90%

Learning to Compress: Unlocking the Potential of Large Language Models for Text Representation

学习压缩:解锁大型语言模型在文本表示中的潜力

Yeqin Zhang, Yizheng Zhao, Chen Hu, Binxing Jiao, Daxin Jiang, Ruihang Miao, Cam-Tu Nguyen

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI

AI总结 本文提出通过上下文压缩作为预训练任务来提升大型语言模型的文本表示能力,实验表明其在多种任务中优于传统方法。

Comments Accepted by AAAI'26

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.20949 2025-12-25 cs.CL cs.AI 88%

Neural Probe-Based Hallucination Detection for Large Language Models

基于神经探针的大型语言模型幻觉检测

Shize Liang, Hongzhi Wang

机构 * Faculty of Computing, Harbin Institute of Technology(计算机学院,哈尔滨工业大学)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本文提出基于神经网络的token级幻觉检测框架,通过冻结语言模型参数并使用MLP探针进行非线性建模,结合多目标损失函数和贝叶斯优化,实现对LLMs幻觉内容的高效检测。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.20794 2025-12-25 cs.CL 84%

Investigating Model Editing for Unlearning in Large Language Models

探究大型语言模型中去学习的模型编辑

Shariqah Hossain, Lalana Kagal

机构 * Massachusetts Institute of Technology(麻省理工学院)

专题命中 知识编辑与模型理解 :large language model(title);language model(title);分类 cs.CL

AI总结 本研究探讨了大型语言模型中去学习的模型编辑方法,通过设计新的编辑目标,展示了模型编辑在去学习任务中的优势。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.21024 2025-12-25 cs.GT cs.AI 77%

Policy-Conditioned Policies for Multi-Agent Task Solving

基于策略条件的多智能体任务解决

Yue Lin, Shuhui Zhu, Wenhao Li, Ang Li, Dan Qiao, Pascal Poupart, Hongyuan Zha, Baoxiang Wang

机构 * The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳)) University of Waterloo(滑铁卢大学) Tongji University(同济大学) Vector Institute(向量研究所)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出通过程序化表示和大型语言模型实现多智能体任务解决,引入程序均衡概念并提出PIBR算法,有效解决协调博弈和合作环境问题。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.21004 2025-12-25 cs.CV 67%

Learning from Next-Frame Prediction: Autoregressive Video Modeling Encodes Effective Representations

从下一帧预测学习:自回归视频建模编码有效表示

Jinghan Li, Yang Jin, Hao Jiang, Yadong Mu, Yang Song, Kun Xu

机构 * Peking University(北京大学)

专题命中 知识编辑与模型理解 :foundation model(abstract);pretraining(abstract)

AI总结 NExT-Vid通过掩码下一帧预测提出自回归视频预训练框架,提升视频生成质量和语义表示能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.20796 2025-12-25 cs.CL 57%

Measuring Mechanistic Independence: Can Bias Be Removed Without Erasing Demographics?

测量机制独立性:在不擦除人口统计学信息的情况下,偏见能否被移除?

Zhengyang Shan, Aaron Mueller

机构 * Boston University(波士顿大学)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

AI总结 本研究通过多任务评估发现,语言模型中的偏见源于任务特定机制,通过针对性干预可实现去偏见而不影响核心能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.19347 2025-12-25 physics.optics cs.GR eess.IV 50%

High contrast holography through dual modulation

高对比度全息成像通过双调制

Leyla Kabuli, Oliver Cossairt, Florian Schiffers, Nathan Matsuda, Grace Kuo

专题命中 知识编辑与模型理解 :SLM(abstract)

AI总结 本文提出通过双调制技术提升全息显示对比度,实验显示对比度显著提高,为高对比度全息显示提供新设计思路。

Comments 24 pages, 17 figures

Journal ref Nature Scientific Reports 15, 17615 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 其他LLM 7 篇

2512.09769 2025-12-25 cs.CR 89%

Defining Cost Function of Steganography with Large Language Models

利用大语言模型定义隐写术的成本函数

Hanzhou Wu, Yige Wang

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract)

AI总结 本文利用大语言模型提出隐写术成本函数的定义方法,通过两阶段策略结合程序合成与进化搜索,设计出更高效的隐写术成本函数。

Comments Some minor typo errors are corrected, https://scholar.google.com/citations?hl=en&user=IdiF7M0AAAAJ

Journal ref IS&T Electronic Imaging, Media Watermarking, Security, and Forensics (2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.20973 2025-12-25 cs.MA 75%

DAO-Agent: Zero Knowledge-Verified Incentives for Decentralized Multi-Agent Coordination

DAO-Agent: 零知识验证的去中心化多智能体协调激励机制

Yihan Xia, Taotao Wang, Wenxin Xu, Shengli Zhang

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 DAO-Agent通过零知识证明和混合架构,在去中心化环境中实现可审计的多智能体协调与公平激励分配,显著降低链上计算成本。

Comments 10 pages, 1 figure

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.24331 2025-12-25 cs.CL 74%

Context-Aware Sentiment Forecasting via LLM-based Multi-Perspective Role-Playing Agents

基于LLM的多视角角色扮演代理的上下文感知情感预测

Fanhang Man, Huandong Wang, Jianjie Fang, Zhaoyi Deng, Baining Zhao, Xinlei Chen, Yong Li

机构 * Shenzhen International Graduate School Tsinghua University(清华大学深圳国际研究生院) Department of Electronic Engineering Tsinghua University(清华大学电子工程系) Northeastern University at Qinghuangdao(清华大学秦皇岛分校) Department of Computer Science University of California Irvine(加州大学 Irvine 分校计算机科学系)

专题命中 其他LLM :LLM(title);分类 cs.CL

AI总结 本文提出基于LLM的多视角角色扮演框架,用于社交媒体中基于上下文的情感预测,通过提取情感相关特征提升预测性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.20649 2025-12-25 cs.AI cs.CR 70%

AIAuditTrack: A Framework for AI Security system

AIAuditTrack:AI安全系统的框架

Zixun Luo, Yuhang Fan, Yufei Li, Youzhi Zhang, Hengyu Lin, Ziqi Wang

机构 * organization= Huazhong University of Science organization= Lingnan University organization= Centre for Artificial Intelligence Robotics (CAIR) Hong Kong Institute of Science \& Innovation, Chinese Academy of Sciences Tsinghua University Fujian Jiangxia University

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 AIAuditTrack通过区块链技术实现AI交互轨迹记录与治理,提供可扩展的AI安全审计与责任追溯方案。

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.02080 2025-12-25 cs.CR cs.AI 70%

Evolving Security in LLMs: A Study of Jailbreak Attacks and Defenses

LLM安全性的演变:对劫持攻击及防御的研究

Zhengchun Shang, Wenlan Wei, Weiheng Bai

机构 * Cornell University Ithaca, NY Department of Computer Science \& Engineering University of Minnesota -- Twin Cities Minneapolis, MN

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文研究了LLM安全性的演变,分析了劫持攻击的检测技术,并探讨了模型版本、大小及多防御策略对安全性的综合影响。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.20996 2025-12-25 cs.AI 57%

TrafficSimAgent: A Hierarchical Agent Framework for Autonomous Traffic Simulation with MCP Control

TrafficSimAgent: 一个用于具有MCP控制的自动驾驶交通模拟的分层代理框架

Yuwei Du, Jun Zhang, Jie Feng, Zhicheng Liu, Jian Yuan, Yong Li

机构 * Department of Electronic Engineering, BNRist, Tsinghua University, Beijing, China(电子工程系、BNRist、清华大学、北京、中国) AMAP, Alibaba Group, Beijing, China(AMAP、阿里巴巴集团、北京、中国)

专题命中 其他LLM :LLM(abstract);分类 cs.AI

AI总结 TrafficSimAgent通过基于大语言模型的分层代理框架,解决交通模拟中实验设计和决策优化的挑战,实现高效且准确的自动驾驶交通模拟。

Comments The code will be available at: https://github.com/tsinghua-fib-lab/TrafficSimAgent

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.14352 2025-12-25 gr-qc hep-th 50%

Weak Gravity Conjecture Validation with Photon Spheres of Quantum Corrected AdS-Reissner-Nordstrom Black Holes in Kiselev Spacetime

用量子修正的AdS-Reissner-Nordstrom黑洞在Kiselev时空中的光子球验证弱引力猜想

Mohammad Reza Alipour, Mohammad Ali S. Afshar, Saeed Noori Gashti, Jafar Sadeghi

专题命中 其他LLM :prompting(abstract)

AI总结 该研究通过分析量子修正的AdS-RN黑洞在Kiselev时空中的光子球,验证了弱引力猜想在特定参数下的适用性。

Comments 13 pages, 6 figures, 1 Table

Journal ref The European Physical Journal C 85.2 (2025): 138

详情

展开后加载摘要…

URL PDF HTML 收藏