arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-12-08 至 2025-12-08 共收录 139 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 12 篇

2507.08355 2025-12-08 cs.LG 57%

scE2TM improves single-cell embedding interpretability and reveals cellular perturbation signatures

scE2TM提升了单细胞嵌入的可解释性并揭示了细胞扰动特征

Hegang Chen, Yuyin Lu, Yifan Zhao, Zhiming Dai, Fu Lee Wang, Qing Li, Yanghui Rao, Yue Li

机构 * School of Computer Science and Engineering, Sun Yat-sen University, Guangzhou, China(中山大学计算机科学与工程学院) School of Computer Science, McGill University, Montreal, Canada(麦吉尔大学计算机科学学院) Department of Biomedical Informatics, Harvard Medical School, Boston, USA(哈佛医学院生物医学信息学系) School of Science and Technology, Hong Kong Metropolitan University, Hong Kong, China(香港 metropolitan 大学科学与技术学院) Department of Computing, The Hong Kong Polytechnic University, Hong Kong, China(香港理工大学计算系)

专题命中 知识编辑与模型理解 :foundation model(abstract);分类 cs.LG

AI总结 scE2TM通过外部知识引导的嵌入式主题模型提升单细胞嵌入的可解释性,揭示细胞扰动特征和生物通路一致性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05677 2025-12-08 stat.ME math.PR stat.ML 50%

Empirical Decision Theory

经验决策理论

Christoph Jansen, Georg Schollmeyer, Thomas Augustin, Julian Rodemann

专题命中 知识编辑与模型理解 :prompting(abstract)

AI总结 本文提出了一种经验决策模型,通过协议中的观察行动-后果对来处理决策问题,无需显式指定世界状态,提供了三种推断保证方法。

Comments Christoph Jansen and Georg Schollmeyer contributed equally

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05482 2025-12-08 cs.CV 50%

Concept-based Explainable Data Mining with VLM for 3D Detection

基于概念的可解释数据挖掘与VLM用于3D检测

Mai Tsujimoto

机构 * The University of Tokyo(东京大学)

专题命中 知识编辑与模型理解 :language model(abstract)

AI总结 本文提出基于概念的可解释数据挖掘方法,利用VLMs识别稀有物体以提升3D检测性能,减少标注负担并提高模型效果。

Comments 28 pages including appendix. Code: https://github.com/mm1129/concept_based_rare_detector_2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05349 2025-12-08 cond-mat.mtrl-sci 50%

Platonic representation of foundation machine learning interatomic potentials

柏拉图表示法用于基础机器学习互原子势

Zhenzhu Li, Aron Walsh

专题命中 知识编辑与模型理解 :foundation model(abstract)

AI总结 本研究提出柏拉图表示法,通过统一不同MLIPs的潜在空间,实现跨模型最优传输和可解释的嵌入运算,揭示潜在空间中几何失真与物理预测失败的关系。

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 其他LLM 15 篇

2512.05525 2025-12-08 cs.DB cs.LG 89%

Poodle: Seamlessly Scaling Down Large Language Models with Just-in-Time Model Replacement

Poodle:通过即时模型替换无缝缩放大型语言模型

Nils Strassenburg, Boris Glavic, Tilmann Rabl

机构 * Hasso Plattner Institute, Uni Potsdam(霍普夫-普朗特研究所,波茨坦大学) University of Illinois Chicago(伊利诺伊大学芝加哥分校)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.LG

AI总结 Poodle通过即时模型替换技术,在无需用户干预的情况下,自动替换大型语言模型为更经济的替代模型,以降低资源和能源消耗,同时保持模型的易用性和性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05331 2025-12-08 cs.CL cs.LG 86%

Exposing Pink Slime Journalism: Linguistic Signatures and Robust Detection Against LLM-Generated Threats

揭露粉红 slime 纪实:语言特征与对抗 LLM 生成威胁的鲁棒检测

Sadat Shahriar, Navid Ayoobi, Arjun Mukherjee, Mostafa Musharrat, Sai Vishnu Vamsi

机构 * University of Houston, Texas, USA(德克萨斯大学)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

AI总结 本文提出了一种基于语言特征的鲁棒检测框架,以应对LLM生成的粉红 slime 纪实威胁,提升了检测性能27%。

Comments Published in RANLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05243 2025-12-08 cs.CL cs.CY 85%

Decoding the Black Box: Discerning AI Rhetorics About and Through Poetic Prompting

解码黑箱:通过诗意提示 discerning AI 的修辞

P. D. Edgar, Alia Hall

专题命中 其他LLM :prompting(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 通过诗意提示评估AI模型对诗人作品的描述和评价,探讨模型在受众面前适应和重写创意作品的能力。

Comments Late-Breaking Paper accepted to IEEE SSCI 2025 NLP & Social Media Track as extended abstract and presented in Trondheim, Norway 17-20 March 2025 as Poster Presentation

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05309 2025-12-08 cs.SE 85%

Engagement in Code Review: Emotional, Behavioral, and Cognitive Dimensions in Peer vs. LLM Interactions

代码审查中的参与:同行与LLM互动中的情感、行为和认知维度

Adam Alami, Nathan Cassee, Thiago Rocha Silva, Elda Paja, Neil A. Ernst

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract)

AI总结 研究探讨了同行与LLM辅助代码审查中参与的差异,揭示了情感自我调节与行为参与的关系,以及AI作为支持性伙伴对减少认知和情绪负荷的作用。

Comments Submitted to TOSEM

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05138 2025-12-08 cond-mat.soft cond-mat.mtrl-sci 85%

polyRETRO: a Language Model Approach to predict Polymerization Class and Monomer(s) for a Target Polymer

polyRETRO: 一种语言模型方法用于预测目标聚合物的聚合类别和单体

Sakshi Agarwal, Wei Xiong, Rampi Ramprasad

专题命中 其他LLM :language model(title,abstract);LLM(abstract);large language model(abstract)

AI总结 polyRETRO通过语言模型预测聚合物的聚合类别和单体,为连接计算设计与实验合成提供新方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05162 2025-12-08 stat.ML cs.AI cs.LG math.DS math.PR 84%

How to Tame Your LLM: Semantic Collapse in Continuous Systems

如何驯服你的大语言模型:连续系统中的语义崩溃

C. M. Wyss

机构 * Exolytica AI

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本文提出通过连续状态机理论解释大语言模型中离散符号语义的生成,揭示了语义坍缩与逻辑可解释性的统一。

Comments 35 pages, 1 figure. Exolytica AI Technical Report XTR-2025-01

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.02221 2025-12-08 cs.CL 81%

GDC Cohort Copilot: An AI Copilot for Curating Cohorts from the Genomic Data Commons

GDC队列助手:一种用于从基因组数据共同体中整理队列的AI助手

Steven Song, Anirudh Subramanyam, Zhenyu Zhang, Aarti Venkat, Robert L. Grossman

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);prompting(abstract)

AI总结 GDC队列助手通过自然语言处理技术帮助用户从基因组数据共同体中高效整理队列,采用本地开源模型优于GPT-4o。

Comments 12 pages, 1 figure, 7 tables. v2 updated to reflect migration to HF Spaces

Journal ref Bioinformatics Advances 5(1), vbaf295 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05222 2025-12-08 cs.LG 79%

Mitigating the Antigenic Data Bottleneck: Semi-supervised Learning with Protein Language Models for Influenza A Surveillance

缓解抗原数据瓶颈:基于蛋白质语言模型的半监督学习用于甲型流感A病毒监测

Yanhua Xu

机构 * Department of Computer Science, University of Liverpool, UK(利物浦大学计算机科学系)

专题命中 其他LLM :language model(title,abstract);分类 cs.LG

AI总结 本研究利用蛋白质语言模型与半监督学习缓解抗原数据瓶颈,提升流感A病毒监测的预测准确性。

Comments V0: initial draft uploaded

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05374 2025-12-08 cs.CR cs.AI cs.DB 77%

Please Don't Kill My Vibe: Empowering Agents with Data Flow Control

请不要破坏我的氛围:通过数据流控制赋能代理

Charlie Summers, Haneen Mohammed, Eugene Wu

机构 * Columbia University(哥伦比亚大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出通过数据流控制(DFC)机制提升代理的安全性和可控性,旨在解决LLM代理在执行复杂任务时面临的风险问题。

Comments 7 pages, 7 figures, CIDR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05925 2025-12-08 cs.AI cs.CL 76%

To Err Is Human: Systematic Quantification of Errors in Published AI Papers via LLM Analysis

出错是人之常情:通过LLM分析系统性量化已发表AI论文中的错误

Federico Bianchi, Yongchan Kwon, Zachary Izzo, Linjun Zhang, James Zou

机构 * Together AI NEC Labs America(NEC美国实验室) Rutgers University(罗格斯大学) Stanford University(斯坦福大学)

专题命中 其他LLM :LLM(title);分类 cs.CL、cs.AI

AI总结 通过LLM分析发现已发表AI论文中存在大量客观错误,错误数量随时间增加,AI检查器能有效识别并纠正大部分错误,提升文献的准确性和可重复性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05399 2025-12-08 cs.DB 75%

Featurized-Decomposition Join: Low-Cost Semantic Joins with Guarantees

特征分解连接:具有保证的低成本语义连接

Sepanta Zeighami, Shreya Shankar, Aditya Parameswaran

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 Featurized-Decomposition Join通过自动提取特征并结合逻辑表达式,实现低成本高质量的语义连接。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05389 2025-12-08 cs.HC 75%

CLIO: A Tour Guide Robot with Co-speech Actions for Visual Attention Guidance and Enhanced User Engagement

CLIO:一种带有同步动作的导览机器人,用于视觉注意力引导和增强用户参与度

Yuxuan Chen, Ian Leong Ting Lo, Bao Guo, Netitorn Kawmali, Chun Kit Chan, Ruoyu Wang, Jia Pan, Lei Yang

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 CLIO是一种通过同步动作引导游客视觉注意力并提升用户参与度的导览机器人,利用大型语言模型协调动作与叙述脚本,实验证明其在增强游客互动方面的有效性。

Comments 10 pages, 7 figures, human-robot interaction

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05666 2025-12-08 cs.LG cs.AI cs.SE 73%

Feasibility of AI-Assisted Programming for End-User Development

面向终用户开发的AI辅助编程可行性

Irene Weber

机构 * University of Applied Sciences Kempten Faculty of Mechanical Engineering(应用科学大学克雷滕学院机械工程系)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本文研究了AI辅助编程在终用户开发中的可行性,通过案例研究发现非程序员能通过AI助手快速开发应用程序,可能替代传统低代码平台。

Comments 12 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05718 2025-12-08 q-bio.NC 67%

Emergence of Language in the Developing Brain

发育大脑中语言的涌现

Linnea Evanson, Christine Bulteau, Mathilde Chipaux, Georg Dorfmüller, Sarah Ferrand-Sorbets, Emmanuel Raffo, Sarah Rosenberg, Pierre Bourdillon, Jean-Rémi King

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 研究揭示了发育大脑中语言表示的成熟过程,并表明现代AI系统能有效建模语言习得的神经基础。

Comments *Equal contribution

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05144 2025-12-08 econ.GN cs.CY cs.SI q-fin.EC stat.AP 67%

Job Satisfaction Through the Lens of Social Media: Rural--Urban Patterns in the U.S

通过社交媒体视角审视就业满意度:美国城乡差异模式

Stefano M Iacus, Giuseppe Porro

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本研究通过分析社交媒体数据,发现城乡就业满意度差异受劳动力市场松弛影响,而非单纯收入差距。

详情

展开后加载摘要…

URL PDF HTML 收藏