arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Transactions on Machine Learning Research · 期刊 · Machine Learning

2026-01-19 至 2026-01-19 共收录 4
2507.12979 2026-01-19 cs.LG cs.AI

A Distributed Generative AI Approach for Heterogeneous Multi-Domain Environments under Data Sharing constraints

在数据共享约束下用于异构多域环境的分布式生成AI方法

Youssef Tawfilis, Hossam Amer, Minar El-Aasser, Tallal Elshabrawy

AI总结 本研究提出了一种在数据共享约束下用于异构多域环境的分布式生成AI方法,结合KLD加权聚类联邦学习和异构U型分割学习,提升多域生成模型的训练效率和性能。

Comments Accepted and published in Transactions on Machine Learning Research (TMLR), 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.06125 2026-01-19 math.OC cs.IT cs.LG math.IT math.PR

Convergence of linear programming hierarchies for Gibbs states of spin systems

线性规划层次在自旋系统吉布斯态下的收敛性

Hamza Fawzi, Omar Fawzi

机构 * DAMTP, University of Cambridge, United Kingdom(剑桥大学 DAMTP 实验室,英国) Univ Lyon, Inria, ENS Lyon, UCBL, LIP, France(里昂大学,法国国家信息与自动化技术研究院,里昂高等师范学校, UCBL,LIP,法国)

AI总结 本文研究了两种线性规划层次在自旋系统吉布斯态下的收敛性,证明了在空间混合和马尔可夫链快速混合条件下,能够高效近似局部期望值并提供严格上下界。

Comments 11 pages

Journal ref Transactions on Machine Learning Research (11/2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.01386 2026-01-19 cs.LG

ThinkEval: Practical Evaluation of Knowledge Leakage in LLM Editing using Thought-based Knowledge Graphs

ThinkEval: 基于思维的知识图谱对 LLM 编辑中知识泄漏的实用评估

Manit Baser, Dinil Mon Divakaran, Mohan Gurusamy

机构 * Electrical and Computer Engineering National University of Singapore(新加坡国立大学电子与计算机工程系) A*STAR Institute for Infocomm Research (A*STAR I 2 R), Singapore(新加坡A*STAR信息与通信研究机构)

AI总结 ThinkEval通过构建专门知识图谱评估LLM编辑中的间接知识泄漏,揭示了现有编辑技术在平衡知识抑制与保留方面的不足。

Comments Accepted to TMLR

Journal ref Transactions on Machine Learning Research (01/2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.07128 2026-01-19 cs.CL

DeepSeek-R1 Thoughtology: Let's think about LLM Reasoning

DeepSeek-R1 思维学:让我们思考LLM推理

Sara Vera Marjanović, Arkil Patel, Vaibhav Adlakha, Milad Aghajohari, Parishad BehnamGhader, Mehar Bhatia, Aditi Khandelwal, Austin Kraft, Benno Krojer, Xing Han Lù, Nicholas Meade, Dongchan Shin, Amirhossein Kazemnejad, Gaurav Kamath, Marius Mosbach, Karolina Stańczak, Siva Reddy

AI总结 DeepSeek-R1通过多步骤推理链实现复杂问题解决,研究揭示其推理性能的'甜点'及潜在安全漏洞。

Comments 135 pages, Published to TMLR

详情

展开后加载摘要…

URL PDF HTML 收藏