arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Transactions on Machine Learning Research · 期刊 · Machine Learning

2026-05-25 至 2026-05-25 共收录 5
2508.13663 2026-05-25 cs.AI cs.LG

Interactive Query Answering on Knowledge Graphs with Soft Entity Constraints

具有软实体约束的知识图谱交互式查询回答

Daniel Daza, Alberto Bernardi, Luca Costabello, Christophe Gueret, Masoud Mansoury, Michael Cochez, Martijn Schut

机构 * Translational AI Laboratory, Department of Laboratory Medicine(转化人工智能实验室,实验室医学系) Amsterdam University Medical Center, Vrije Universiteit Amsterdam(阿姆斯特丹大学医学中心,伏里埃大学阿姆斯特丹) Accenture Labs(埃森哲实验室) Delft University of Technology(代尔夫特理工大学) ELLIS Institute Finland & Abo Akademi University, Turku, Finland & Elsevier Discovery Lab, Amsterdam(芬兰ELLIS研究所 & 阿博阿卡迪米大学,图尔库,芬兰 & 埃西弗尔发现实验室,阿姆斯特丹)

AI总结 针对知识图谱查询中存在的模糊或上下文依赖的软约束问题,提出两种轻量级方法,通过调整查询答案分数来融入软约束,保持原始排序结构,并在扩展基准上验证了性能。

Comments Accepted in Transactions on Machine Learning Research (2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.18000 2026-05-25 cs.LG cs.AI q-bio.PE

Reward Engineering for Spatial Epidemic Simulations: A Reinforcement Learning Platform for Individual Behavioral Learning

空间流行病模拟中的奖励工程:个体行为学习的强化学习平台

Radman Rakhshandehroo, Daniel Coombs

机构 * Department of Computer Science University of British Columbia(计算机科学系,不列颠哥伦比亚大学) Department of Mathematics and Institute of Applied Mathematics University of British Columbia(数学系和应用数学研究所,不列颠哥伦比亚大学)

AI总结 提出ContagionRL平台,通过奖励函数设计系统评估空间流行病模拟中个体行为学习策略,发现势场奖励方法能有效提升非药物干预依从性和空间规避策略。

Comments 38 pages, 15 figures and 18 tables; Accepted to TMLR. OpenReview: https://openreview.net/forum?id=yPEASsx3hk

Journal ref Transactions on Machine Learning Research, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.03508 2026-05-25 cs.LG

D2 Actor Critic: Diffusion Actor Meets Distributional Critic

D2 Actor Critic: 扩散演员遇上分布式评论家

Lunjun Zhang, Shuo Han, Hanrui Lyu, Bradly C Stadie

机构 * Department of Computer Science, University of Toronto(计算机科学系,多伦多大学) Department of Statistics, Northwestern University(统计学系,西北大学)

AI总结 提出D2AC算法,通过融合扩散策略与分布式评论家,实现无模型强化学习中扩散策略的在线高效训练,在18个困难任务上达到最先进性能。

Comments Accepted to TMLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.15105 2026-05-25 cs.LG

Super-Linear: A Lightweight Pretrained Mixture of Linear Experts for Time Series Forecasting

Super-Linear: 一种轻量级预训练线性专家混合模型用于时间序列预测

Liran Nochumsohn, Raz Marshanski, Hedi Zisling, Omri Azencot

机构 * Faculty of Computer and Information Science, Ben-Gurion University(计算机与信息科学学院,本·古里安大学)

AI总结 提出Super-Linear,一种轻量级可扩展的混合专家模型,用频率特化的线性专家替代深度架构,通过光谱门控机制实现高效准确的时间序列预测。

Journal ref Transactions on Machine Learning Research (TMLR), 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14311 2026-05-25 cs.LG cs.AI

Online Learning with Multiple Fairness Regularizers via Graph-Structured Feedback

通过图结构反馈进行多重公平正则化器的在线学习

Quan Zhou, Jakub Marecek, Robert Shorten

机构 * Department of Mathematics, National University of Singapore(新加坡国立大学数学系) Department of Computer Science, Czech Technical University(捷克技术大学计算机科学系) Dyson School of Design Engineering, Imperial College London(伦敦帝国理工学院设计工程戴森学院) Imperial College London(伦敦帝国理工学院)

AI总结 本文针对在线决策中多重公平约束的权重自适应问题,提出了一种基于图结构反馈的赌博机算法,能够在不预先知道权重的情况下在线学习并平衡多个公平性目标。

Comments Published in Transactions on Machine Learning Research (TMLR), 2026. OpenReview: https://openreview.net/forum?id=y8iWuDZtEw

Journal ref Transactions on Machine Learning Research (TMLR), 2026

详情

展开后加载摘要…

URL PDF HTML 收藏