A Primer on SO(3) Action Representations in Deep Reinforcement Learning
深度强化学习中SO(3)动作表示入门
AI总结 针对SO(3)动作表示在强化学习中的选择问题,系统评估了多种表示在PPO、SAC、TD3算法下的效果,发现切向量表示最可靠。
Comments Published at The Fourteenth International Conference on Learning Representations (ICLR 2026)