arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-12-15 至 2025-12-15 共收录 137 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 15 篇

2512.10867 2025-12-15 cs.CV 78%

From Macro to Micro: Benchmarking Microscopic Spatial Intelligence on Molecules via Vision-Language Models

从宏观到微观:通过视觉-语言模型在分子上基准测试微观空间智能

Zongzhao Li, Xiangzhe Kong, Jiahui Su, Zongyang Ma, Mingze Li, Songyou Li, Yuelin Zhang, Yu Rong, Tingyang Xu, Deli Zhao, Wenbing Huang

机构 * Gaoling School of Artificial Intelligence, Renmin University of China(中国人民大学北京校区人工智能学院) Beijing Key Laboratory of Research on Large Models and Intelligent Governance(北京大型模型与智能治理研究重点实验室) Engineering Research Center of Next-Generation Intelligent Search and Recommendation, MOE(下一代智能搜索与推荐工程研究中心) Dept. of Comp. Sci. & Tech., Tsinghua University(清华大学计算机科学与技术系) Institute for AI Industry Research (AIR), Tsinghua University(清华大学人工智能产业研究院) SKL-ESPC & SEPKL-AERM, College of Environmental Sciences and Engineering, Peking University(环境科学与工程学院,北京大学) MAIS, Institute of Automation, Chinese Academy of Sciences(自动化研究所,中国科学院) DAMO Academy, Alibaba Group, Hangzhou, China(阿里巴巴集团 DAMO Academy,中国杭州) Hupan Lab, Hangzhou, China(杭州实验室)

专题命中 领域大模型 :language model(title,abstract)

AI总结 本文提出MiSI-Bench基准测试框架,评估视觉-语言模型在分子微观空间智能任务中的表现,发现微调模型在空间转换任务上超越人类,但需整合领域知识以推动科学AGI发展。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.11474 2025-12-15 cs.AI cs.CY cs.IR 77%

General-purpose AI models can generate actionable knowledge on agroecological crop protection

通用人工智能模型可为生态农业作物保护生成可操作的知识

Kris A. G. Wyckhuys

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 通用人工智能模型在生态农业作物保护中生成可操作知识,通过对比DeepSeek与ChatGPT的性能,展示其在科学知识生成和决策支持中的潜力。

Comments 33 pages, 3 figures, 3 tables, 1 supplementary table

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.10991 2025-12-15 cs.LG cs.AI physics.chem-ph q-bio.QM 62%

MolSculpt: Sculpting 3D Molecular Geometries from Chemical Syntax

MolSculpt:从化学语法雕刻3D分子几何

Zhanpeng Chen, Weihao Gao, Shunyu Wang, Yanan Zhu, Hong Meng, Yuexian Zou

机构 * AI for Science (AI4S)-Preferred Program, Peking University Shenzhen Graduate School, China(人工智能科学(AI4S)优选计划,北京大学深圳研究生院,中国) Faculty of Materials Science, Shenzhen MSU-BIT University, Shenzhen, China(材料科学学院,深圳MSU-BIT大学,中国)

专题命中 领域大模型 :foundation model(abstract);分类 cs.AI、cs.LG

AI总结 MolSculpt通过整合1D化学知识与3D生成过程,实现了从化学语法到精确3D分子几何的高效生成。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.11251 2025-12-15 cs.LG 57%

Insight Miner: A Time Series Analysis Dataset for Cross-Domain Alignment with Natural Language

Insight Miner: 一个用于跨领域对齐的时间序列分析数据集与自然语言

Yunkai Zhang, Yawen Zhang, Ming Zheng, Kezhen Chen, Chongyang Gao, Ruian Ge, Siyuan Teng, Amine Jelloul, Jinmeng Rao, Xiaoyuan Guo, Chiang-Wei Fang, Zeyu Zheng, Jie Yang

机构 * UC Berkeley(加州大学伯克利分校) Mineral(矿石) Northwestern University(西北大学)

专题命中 领域大模型 :instruction tuning(abstract);分类 cs.LG

AI总结 Insight Miner 是一个大规模多模态模型,通过代理工作流生成高质量时间序列描述,提升跨领域时间序列分析能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.16979 2025-12-15 cs.CV cs.AI 57%

The Finer the Better: Towards Granular-aware Open-set Domain Generalization

更细的更好:面向粒度感知的开集域泛化

Yunyun Wang, Zheng Duan, Xinyue Liao, Ke-Jia Chen, Songcan Chen

专题命中 领域大模型 :language model(abstract);分类 cs.AI

AI总结 SeeCLIP通过细粒度语义增强解决开集域泛化中的结构风险与开放空间风险困境,提升模型对难区分未知类别的判别能力。

Comments 9 pages,3 figures,aaai2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.11183 2025-12-15 cs.LG 57%

Progress over Points: Reframing LM Benchmarks Around Scientific Objectives

基于点的进展:围绕科学目标重新构建语言模型基准测试

Alwin Jin, Sean M. Hendryx, Vaskar Nath

机构 * Georgia Institute of Technology(佐治亚理工学院) Scale AI

专题命中 领域大模型 :language model(abstract);分类 cs.LG

AI总结 本文提出以科学目标为导向的语言模型基准测试环境,通过标准化数据集和训练框架,推动语言模型领域的科学进步。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.11548 2025-12-15 cs.CV 50%

SSL-MedSAM2: A Semi-supervised Medical Image Segmentation Framework Powered by Few-shot Learning of SAM2

SSL-MedSAM2: 一种由SAM2少样本学习驱动的半监督医学图像分割框架

Zhendi Gong, Xin Chen

机构 * School of Computer Science, University of Nottingham, UK(Nottingham大学计算机科学学院)

专题命中 领域大模型 :foundation model(abstract)

AI总结 SSL-MedSAM2通过SAM2的少样本学习和nnUNet的迭代监督学习,在肝脏分割任务中实现了高精度的半监督医学图像分割。

Comments Accepted by MICCAI 2025 CARE Challenge, waiting for publication

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 知识编辑与模型理解 5 篇

2503.07020 2025-12-15 cs.RO cs.AI cs.LG 86%

Driving Through Uncertainty: Risk-Averse Control with LLM Commonsense for Autonomous Driving under Perception Deficits

在不确定性中驾驶:利用大语言模型常识实现风险厌恶控制以应对自动驾驶中的感知缺陷

Yuting Hu, Chenhui Xu, Ruiyang Qin, Dancheng Liu, Amir Nassereldine, Yiyu Shi, Jinjun Xiong

机构 * University at Buffalo(布法罗大学) University of Notre Dame(诺特尔大学)

专题命中 知识编辑与模型理解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本文提出LLM-RCO框架,利用大语言模型整合人类驾驶常识,以提升自动驾驶在感知缺陷下的风险规避与适应性控制能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.11104 2025-12-15 cs.CV 78%

Information-driven Fusion of Pathology Foundation Models for Enhanced Disease Characterization

基于信息驱动的病理基础模型融合以增强疾病表征

Brennan Flannery, Thomas DeSilvio, Jane Nguyen, Satish E. Viswanath

机构 * Case Western Reserve University(凯斯西储大学) Department of Biomedical Engineering(生物医学工程系) Cleveland Clinic(克利夫兰诊所) Department of Pathology(病理学系) Emory University(埃默里大学) Department of Pediatrics(儿科学系) Louis Stokes VA Cleveland Medical Center(路易斯·斯托克斯退伍军人医疗中心)

专题命中 知识编辑与模型理解 :foundation model(title,abstract)

AI总结 本研究提出基于信息驱动的病理基础模型融合方法,通过智能融合提升癌症分级和分期的预测性能与可解释性。

Comments 29 Pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.03803 2025-12-15 cs.CL 57%

Enhancing Instruction-Following Capabilities in Seq2Seq Models: DoLA Adaptations for T5

增强序列到序列模型的指令遵循能力:T5的DoLA适应

Huey Sun, Anabel Yong, Lorenzo Gilly, Felipe Jin

机构 * University College London(伦敦大学学院)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

AI总结 通过梯度激活引导方法改进T5模型的指令遵循能力,显著提升MemoTrap性能

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.26005 2025-12-15 nlin.CD 50%

Regime identification and control of extremes in the non-autonomous Lorenz model with chaos and intransitivity

非自治洛伦兹模型中混沌与不稳定性下的极端现象识别与控制

Moyan Liu, Qin Huang, Upmanu Lall

专题命中 知识编辑与模型理解 :foundation model(abstract)

AI总结 本文提出了一种基于非均匀隐马尔可夫模型和局部李雅普诺夫指数的自适应混沌控制策略,用于控制季节性驱动和噪声扰动的洛伦兹84模型中的极端现象。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.04552 2025-12-15 eess.SP 50%

DAS-MAE: A self-supervised pre-training framework for universal and high-performance representation learning of distributed fiber-optic acoustic sensing

DAS-MAE:一种用于分布式光纤声学传感通用性和高性能表示学习的自监督预训练框架

Junyi Duan, Jiageng Chen, Zuyuan He

专题命中 知识编辑与模型理解 :foundation model(abstract)

AI总结 DAS-MAE通过自监督预训练框架实现对分布式光纤声学传感信号的高效表示学习,提升少样本分类性能和实际应用中的识别精度。

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 其他LLM 5 篇

2512.11270 2025-12-15 cs.AI 86%

A-LAMP: Agentic LLM-Based Framework for Automated MDP Modeling and Policy Generation

A-LAMP:基于代理的大型语言模型框架用于自动马尔可夫决策过程建模与策略生成

Hong Je-Gal, Chan-Bin Yi, Hyun-Suk Lee

专题命中 其他LLM :LLM(title,abstract);large language model(abstract,comments);language model(abstract,comments);分类 cs.AI

AI总结 A-LAMP通过基于代理的大型语言模型自动将自然语言任务描述转换为马尔可夫决策过程并生成策略,提升了自动化建模和策略生成的效率与准确性。

Comments NeurIPS 2025 Workshop: Multi-Turn Interactions in Large Language Models. 26 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07462 2025-12-15 cs.MA cs.AI cs.GT cs.LG math.DS 86%

Understanding LLM Agent Behaviours via Game Theory: Strategy Recognition, Biases and Multi-Agent Dynamics

通过博弈论理解LLM代理行为:策略识别、偏差与多代理动态

Trung-Kiet Huynh, Duy-Minh Dao-Sy, Thanh-Bang Cao, Phong-Hao Le, Hong-Dan Nguyen, Phu-Quy Nguyen-Lam, Minh-Luan Nguyen-Vo, Hong-Phat Pham, Phu-Hoa Pham, Thien-Kim Than, Chi-Nguyen Tran, Huy Tran, Gia-Thoai Tran-Le, Alessio Buscemi, Le Hong Trang, The Anh Han

机构 * Faculty of Information and Technology, Ho Chi Minh City University of Science (HCMUS), Vietnam(信息科技学院,胡志明市科学大学(HCMUS)) Vietnam National University - Ho Chi Minh City (VNU-HCM), Vietnam(越南国家大学-胡志明市(VNU-HCM)) Faculty of Computer Science and Engineering, Ho Chi Minh City University of Technology (HCMUT), Vietnam(计算机科学与工程学院,胡志明市技术大学(HCMUT))

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本文通过扩展FAIRGAME框架,系统评估LLM在重复社会困境中的行为,揭示其策略识别、偏差及多代理动态,为AI治理和安全多代理系统设计提供方法论基础。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.16396 2025-12-15 cs.LG 79%

GoalLadder: Incremental Goal Discovery with Vision-Language Models

GoalLadder: 基于视觉语言模型的增量目标发现

Alexey Zakharov, Shimon Whiteson

机构 * University of Oxford(牛津大学)

专题命中 其他LLM :language model(title,abstract);分类 cs.LG

AI总结 GoalLadder利用视觉语言模型在视觉环境中通过增量目标发现提升RL智能体性能,实现高成功率。

Comments NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.15126 2025-12-15 cs.CV cs.AI 77%

Text-Derived Relational Graph-Enhanced Network for Skeleton-Based Action Segmentation

基于文本的图增强网络用于基于骨架的动作分割

Haoyu Ji, Bowen Chen, Weihong Ren, Wenze Huang, Zhihao Yang, Zhiyong Wang, Honghai Liu

机构 * State Key Laboratory of Robotics and Systems, Harbin Institute of Technology Shenzhen(机器人系统国家重点实验室,哈尔滨工业大学深圳)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 TRG-Net通过文本生成的图增强方法,提升基于骨架的动作分割性能,实现更精确的动作识别与分类。

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.17480 2025-12-15 cs.GR cs.AR eess.IV eess.SP physics.optics 50%

Random-phase Wave Splatting of Translucent Primitives for Computer-generated Holography

随机相位波溅射法生成透明原语的全息图用于计算机生成全息图

Brian Chao, Jacqueline Yang, Suyeon Choi, Manu Gopakumar, Ryota Koiso, Gordon Wetzstein

专题命中 其他LLM :SLM(abstract)

AI总结 随机相位波溅射法通过统一的波光学框架,实现任意透明原语的全息图生成,提升带宽利用率并实现逼真的3D全息效果。

详情

展开后加载摘要…

URL PDF HTML 收藏