arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

International Conference on Machine Learning · 会议 · Machine Learning

2025-12-23 至 2025-12-23 共收录 3
2512.19527 2025-12-23 cs.LG cs.DM

Deep Learning for Unrelated-Machines Scheduling: Handling Variable Dimensions

深度学习用于无关机器调度:处理可变维度

Diego Hitzges, Guillaume Sagnol

AI总结 本文提出一种深度学习方法,用于处理无关机器调度中的可变维度问题,通过复杂神经网络架构实现高效调度优化,实验表明其在不同规模问题中表现优于传统调度规则。

Comments 24th IEEE International Conference on Machine Learning and Applications (ICMLA 2025) in Boca Raton, USA. Project page: https://github.com/DiegoHitzges/Deep-Learning-for-Unrelated-Machines-Scheduling . 8 pages, 4 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.03903 2025-12-23 cs.LG

Averaging $n$-step Returns Reduces Variance in Reinforcement Learning

平均 $n$-步回报降低强化学习中的方差

Brett Daley, Martha White, Marlos C. Machado

机构 * Department of Computing Science, University of Alberta, Edmonton, AB, Canada(阿尔伯塔大学计算机科学系) Alberta Machine Intelligence Institute(阿尔伯塔机器智能研究所) Canada CIFAR AI Chair(加拿大CIFAR人工智能主席)

AI总结 本文提出复合回报方法,通过加权平均 $n$-步回报降低强化学习中的方差,提升样本效率和算法性能。

Comments ICML 2024. 27 pages, 7 figures, 3 tables. Fixed minor equation typos

详情

展开后加载摘要…

URL PDF HTML 收藏
2301.11321 2025-12-23 cs.LG

Trajectory-Aware Eligibility Traces for Off-Policy Reinforcement Learning

轨迹感知的eligibility traces用于非策略强化学习

Brett Daley, Martha White, Christopher Amato, Marlos C. Machado

机构 * Department of Computing Science, University of Alberta, Edmonton, AB, Canada(阿尔伯塔大学计算机科学系) Alberta Machine Intelligence Institute(阿尔伯塔机器智能研究所) Canada CIFAR AI Chair(加拿大CIFAR人工智能主席) Khoury College of Computer Sciences, Northeastern University, Boston, MA, USA(东北大学计算机科学学院)

AI总结 本文提出了一种多步算子,用于表达轨迹感知和每决策方法,并通过理论分析为非策略强化学习提供了收敛保证,同时引入RBIS方法在不同λ值下实现稳健性能。

Comments ICML 2023. 18 pages, 4 figures, 1 table. Fixed off-by-1 error in Tightrope Problem

详情

展开后加载摘要…

URL PDF HTML 收藏