arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

作者

Chelsea Finn

Robotics / Machine Learning

至 收录 3
2608.09138 2026-08-12 cs.RO cs.AI 版本更新

SpeedTuning: Speeding Up Policy Execution with Lightweight Reinforcement Learning

SpeedTuning:用轻量强化学习加速策略执行

David D. Yuan, Tony Z. Zhao, Kaylee Burns, Chelsea Finn

机构 * Stanford University(斯坦福大学)

AI总结 SpeedTuning是一种轻量强化学习框架,可预测动作最优执行速度,在无需额外数据采集的情况下,将机器人操作策略加速超2.4倍且保持足够成功率,适用于多种动态精确任务。

Comments 10 pages, 12 figures. This arXiv version includes an appendix with qualitative simulation rollouts and additional ablations. Published at ICRA 2025

Journal ref 2025 IEEE International Conference on Robotics and Automation (ICRA), pp. 1184-1192, 2025

URL PDF HTML 收藏
2607.27203 2026-08-05 cs.LG 版本更新

Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning?

在线RL微调真的需要预训练Q函数吗?

Perry Dong, Ron Polonsky, Dorsa Sadigh, Chelsea Finn

机构 * Stanford University(斯坦福大学)

AI总结 本文针对在线RL微调是否需预训练Q函数的问题,发现预训练Q函数增益有限,提出IPE方法,在连续控制基准中使微调性能平均提升1.26倍。

Comments Fixed typo in author name

URL PDF HTML 收藏
2606.32027 2026-08-04 cs.RO cs.AI cs.LG 版本更新

Freeform Preference Learning for Robotic Manipulation

自由形式偏好学习用于机器人操作

Marcel Torne, Anubha Mahajan, Abhijnya Bhat, Chelsea Finn

机构 * Stanford University(斯坦福大学)

AI总结 提出自由形式偏好学习(FPL),通过自然语言定义偏好轴并学习语言条件奖励模型,在长时域操作任务中比稀疏奖励和二元偏好方法提升38个百分点,支持行为组合与测试时策略引导。

URL PDF HTML 收藏