arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

International Conference on Intelligent Robots and Systems · 会议 · Robotics

2026-06-24 至 2026-06-24 共收录 3
2505.21916 2026-06-24 cs.RO 版本更新

Prior Reinforce: Goal-Conditioned Dynamic Manipulation with Limited Trials

Prior Reinforce: 有限试验下的目标条件动态操控

Yihang Hu, Pingyue Sheng, Yuyang Liu, Shengjie Wang, Yang Gao

机构 * IIIS, Tsinghua University(清华大学智能学院) Shanghai Qi Zhi Institute(上海启智研究院) Spirit AI

AI总结 提出Prior Reinforce框架,利用条件扩散模型从少量演示学习运动流形,并在低维条件空间通过反馈驱动优化适应新目标,实现少至十次试验内的动态操控。

Comments Accepted to the 2026 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.18371 2026-06-24 eess.SY cs.MA cs.RO cs.SY 版本更新

Policy Gradient with Self-Attention for Model-Free Distributed Nonlinear Multi-Agent Games

基于自注意力的策略梯度用于无模型分布式非线性多智能体博弈

Eduardo Sebastián, Maitrayee Keskar, Eeman Iqbal, Eduardo Montijano, Carlos Sagüés, Nikolay Atanasov

机构 * Department of Computer Science and Technology, University of Cambridge(计算机科学与技术系,剑桥大学) Department of Electrical and Computer Engineering, University of California San Diego(电气与计算机工程系,加州大学圣地亚哥分校) RoPeRt group, at DIIS - I3A, Universidad de Zaragoza(RoPeRt组,DIIS - I3A,阿拉贡大学)

AI总结 提出一种分布式策略结构,通过策略梯度学习,利用自注意力层处理时变通信拓扑,解决无模型非线性多智能体博弈问题,在多种场景中表现优异。

Comments The paper has been accepted and will be presented at IEEE/RSJ IROS 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.17070 2026-06-24 cs.RO cs.AI 版本更新

MuTRAP: Multi-trigger Trojans Attacking Robot Task Planning Systems

MuTRAP: 攻击机器人任务规划系统的多触发器木马

Mohaiminul Al Nahian, Zainab Altaweel, David Reitano, Sabbir Ahmed, Shiqi Zhang, Adnan Siraj Rakin

机构 * Binghamton University (SUNY)(宾夕法尼亚州立大学布林顿分校)

AI总结 提出首个针对LLM辅助机器人任务规划器的多触发器木马攻击MuTRAP,通过少量任务特定参数注入后门,并优化触发器词以激活特定恶意行为,揭示当前基于LLM的规划器的安全漏洞。

Comments Accepted for publication at the 2026 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS)

详情

展开后加载摘要…

URL PDF HTML 收藏