QSIM: Mitigating Overestimation in Multi-Agent Reinforcement Learning via Action Similarity Weighted Q-Learning
QSIM:通过动作相似性加权Q学习缓解多智能体强化学习中的过估计
专题命中 多智能体 :agent(title,abstract);multi-agent(title,abstract);分类 cs.AI、cs.LG;planning(comments)
AI总结 QSIM通过动作相似性加权Q学习缓解多智能体强化学习中的过估计问题,提升学习稳定性与性能。
Comments 19 pages, 15 figures, 7tables. Accepted to the 36th International Conference on Automated Planning and Scheduling (ICAPS 2026)