arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

期刊&会议

IEEE T-RO

IEEE Transactions on Robotics · 期刊 · Robotics

共收录 512
2608.31048 2026-09-10 eess.IV 版本更新

OmniRAS: Standardizing Foundation Model Training and Evaluation in Robot-Assisted Surgery

OmniRAS:标准化机器人辅助手术领域的基础模型训练与评估

Leonardo Borgioli, Neil Getty, Wenli Xiu, Jessica Cassiani, Alvaro Ducas, Carlos Agustin Orda, Hira Waris, Fangfang Xia, Rick Stevens, Pier Cristoforo Giulianotti, Milos Zefran

AI总结 本文针对机器人辅助手术领域基础模型稀缺的问题,提出OmniRAS系列V-JEPA-2.1编码器,发布专属数据集并完成大规模预训练与多任务评估,取得最优适配结果。

Comments 18 pages, 11 figures, 4 tables. Under review at IEEE Transactions on Robotics (T-RO)

详情

展开后加载摘要…

URL PDF HTML 收藏
2609.07440 2026-09-09 cs.RO 新提交

Open-Set Ego-Noise Separation for Legged-Robot Audition via Annotation-Free Adaptation and Pretrained-Model Transfer

基于无标注适应与预训练模型迁移的四足机器人听觉开集自身噪声分离

Koki Shoda, Jun Younes Louhi Kasahara, Aoba Koyanagi, Qi An, Atsushi Yamashita

机构 * The University of Tokyo(东京大学)

AI总结 本文提出一种无标注适应的开集自身噪声分离框架,利用RecurGraph选择噪声片段、Transfer-DiT迁移预训练模型,在双足和四足机器人上验证了分离性能提升。

Comments 14 pages, 9 figures, 4 tables. Submitted to IEEE Transactions on Robotics

详情

展开后加载摘要…

URL PDF HTML 收藏
2609.06896 2026-09-09 cs.RO 新提交

Distributed Secure Learning Control for Large-scale Multirobots under Stealthy Actuator Attacks

分布式安全学习控制:面向隐蔽执行器攻击下的大规模多机器人系统

Xinglong Zhang, Qingwen Ma, Cong Li, Hui Yin, Changxin Zhang, Yueying Wang, Wei Pan, Xin Xu

机构 * College of Intelligence Science and Technology, National University of Defense Technology(国防科技大学智能科学学院) College of Mechanical and Vehicle Engineering, Hunan University(湖南大学机械与运载工程学院) Department of Computer Science, The University of Manchester(曼彻斯特大学计算机科学系)

AI总结 针对大规模多机器人在隐蔽执行器攻击下的安全控制问题,提出分布式安全学习控制框架,结合强化学习与分布式模型预测控制,采用博弈论架构在线学习攻防策略,经仿真与实验验证有效。

Comments 23 pages, 23 figures. A revised version of this manuscript has been accepted to IEEE Transactions on Robotics

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.04094 2026-09-09 cs.RO math.OC 交叉投稿

Object-Reconstruction-Aware Whole-body Control of Mobile Manipulators

面向物体重建的移动机械臂整体控制

Fatih Dursun, Bruno Vilhena Adorno, Simon Watson, Wei Pan

机构 * University of Manchester(曼彻斯特大学)

AI总结 本文提出一种高效方法,通过计算最信息区域的焦点点,使移动机械臂在路径上保持该点在摄像头视野中,从而在不增加路径规划的情况下提升物体重建效率。

Comments 19 pages, 17 figures, 5 tables. Under Review for the IEEE Transactions on Robotics (T-RO)

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08485 2026-09-09 cs.RO 版本更新

Agile and Generalized Legged Locomotion via Attention-Based Neural Map Encoding

AME-2: 通过基于注意力的神经地图编码实现敏捷且通用的腿式运动

Chong Zhang, Victor Klemm, Fan Yang, Marco Hutter

机构 * Robotic Systems Lab, ETH Zurich, Switzerland(机器人系统实验室,苏黎世联邦理工学院,瑞士) Secure, Reliable, and Intelligent Systems Lab, ETH Zurich, Switzerland(安全、可靠和智能系统实验室,苏黎世联邦理工学院,瑞士) ETH AI Center, Switzerland(苏黎世联邦理工学院人工智能中心,瑞士)

AI总结 本文提出AME-2框架,结合基于注意力的神经地图编码,实现敏捷且通用的腿式运动,通过学习的映射管道提升地形表示的鲁棒性,验证了在四足和双足机器人上的有效性。

Comments Conditionally accepted by IEEE Transactions on Robotics (T-RO). Previously known as AME-2

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.12427 2026-09-09 cs.RO cs.SY eess.SY 版本更新

Temporal Cascading of Planning and Control for Quadrotor MPC

四旋翼MPC的时间级联规划与控制

Rudolf Reiter, Chao Qin, Leonard Bauersfeld, Davide Scaramuzza

AI总结 研究四旋翼空中任务规划与控制问题,提出UNIQUE架构,以时间级联取代分层堆叠,将规划问题作为多阶段MPC的第二尾视野求解,对齐成本、推导约束并引入过渡约束,提升了闭环跟踪性能。

Journal ref IEEE Transactions on Robotics 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20901 2026-09-09 cs.CV 版本更新

Event-Based De-Snowing for Autonomous Driving

自动驾驶中的基于事件相机的去雪方法

Manasi Muglikar, Nico Messikommer, Marco Cannici, Davide Scaramuzza

机构 * University of Zurich(苏黎世大学)

AI总结 本文提出利用事件相机和基于注意力的模块,通过识别雪花条纹遮挡来恢复背景强度,在DSEC-Snow数据集上实现PSNR提升3dB,并改善下游视觉任务性能。

Comments Accepted at IEEE Transactions on Robotics(T-RO), 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2609.03175 2026-09-04 cs.RO cs.SY eess.SY 新提交

Real-Time Shape Control of Multi-Segment Soft Robotic Arms Using Koopman Operators with Global and Local Observables

基于含全局与局部可观测量的Koopman算子的多段柔性机械臂实时形状控制

Jiahe Wang, Eron Ristich, Sultan Haidar Ali, Eric weissman, Lei Zhang, Wanxin Jin, Yi Ren, Jiefeng Sun

AI总结 本文针对多段柔性机械臂形状控制难题,提出结合全局与局部可观测量的Koopman算子模型预测控制框架,经数值与物理实验验证,该框架可实现多段柔性机械臂的实时、可扩展且精准的形状控制。

Comments 18 pages, 14 figures, submitted to TRO

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.09869 2026-09-04 cs.RO cs.SY eess.SY

Redundancy Resolution at Position Level

Alin Albu-Schäffer, Arne Sachtler

Journal ref IEEE Transactions on Robotics, vol. 39, no. 6, pp. 4240-4261, Dec. 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.22006 2026-09-03 cs.RO 版本更新

Parallel Reference-Centric Continuous-Time Relative Localization with Augmented Clamped Non-Uniform B-Splines

并行连续时间相对定位与增强钳制非均匀B样条

Jiadong Lu, Zhehan Li, Tao Han, Miao Xu, Jiajun Lv, Chao Xu, Yanjun Cao

机构 * State Key Laboratory of Industrial Control Technology, Institute of Cyber-Systems and Control, Zhejiang University(工业控制技术国家重点实验室,系统与控制研究所,浙江大学) Huzhou Institute of Zhejiang University(浙江大学湖州研究院) Department of Automatic Control, Faculty of Engineering, Lund University(自动控制系,工程学院, Lund大学)

AI总结 本文提出CT-RIO框架,利用钳制非均匀B样条实现高精度、低延迟的多机器人相对定位,通过灵活的节点管理提升计算效率。

Comments 21 pages, 23 figures, submitted to IEEE Transactions on Robotics

详情

展开后加载摘要…

URL PDF HTML 收藏
2609.01061 2026-09-02 cs.RO cs.LG cs.SY eess.SY 新提交

Accelerating Reinforcement Learning via MPC Solver-Gradient Guidance for Weights-varying MPC

通过针对权重变化的模型预测控制(MPC)的MPC求解器梯度引导加速强化学习

Baha Zarrouki, Arslan Thobani, Jasper Hoffmann, Mattia Piccinini, Rudolf Reiter, Felix Jahncke, Sébastien Gros, Davide Scaramuzza, Johannes Betz

AI总结 该研究提出SG-RL算法,将MPC求解器梯度作为辅助引导,在PPO中实现MPC代价权重自适应,在自主赛车平台上样本效率提升70.6%,性能优于GB-PL且可零样本泛化。

Comments Submitted to IEEE Transactions on Robotics (T-RO). 18 pages, 13 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.19610 2026-08-26 cs.RO 版本更新

Look as You Leap: Planning Simultaneous Motion and Perception for High-DOF Robots

跳跃时观察:为高自由度机器人同时规划运动与感知

Qingxi Meng, Emiliano Flores, Carlos Quintero-Peña, Peizhu Qian, Zachary Kingston, Shannan K. Hamlin, Vaibhav Unhelkar, Lydia E. Kavraki

机构 * Department of Computer Science, Rice University(计算机科学系,里士大学) Ken Kennedy Institute, Rice University(肯尼迪研究所,里士大学) Houston Methodist Academic Institute(休斯顿 Methodist 学术研究院)

AI总结 本文提出了一种基于GPU并行的感知评分引导的概率道路图规划器,用于高自由度机器人在动态和静态环境中的运动与感知同时规划。

Comments 20 pages, 14 figures, Accepted to T-RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.17819 2026-08-19 cs.RO 新提交

Effector-Centric NMPC of Tiltable-Multirotors for Offset-Free Omnidirectional Aerial Manipulation

面向可倾转多旋翼飞行器的以 effector 为中心的非线性模型预测控制(NMPC),用于无偏移全向空中操作

Jinjie Li, Yicheng Chen, Johannes Kübel, Haokun Liu, Junichiro Sugihara, Moju Zhao

AI总结 本研究针对可倾转多旋翼飞行器,设计并提出以 effector 为中心的 NMPC 框架,整合构型优化、奇异处理与扰动补偿,通过真实实验验证其可实现无偏移全向空中操作的可行性。

Comments 22 pages, 26 figures. Accepted to IEEE Transactions on Robotics (T-RO). This arXiv version includes a two-page appendix with additional implementation details

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.16221 2026-08-18 cs.RO 新提交

Deep Probabilistic Indoor Gas Source Localization via Physical Dependency-Guided Sequential Inference

基于物理依赖引导的序列推理的深度概率室内气体源定位

Seunghwan Kim, Hyungjin Kim, Junhee Lee, Hyondong Oh

机构 * Ulsan National Institute of Science and Technology (UNIST)(蔚山科学技术院) Korea Advanced Institute of Science and Technology (KAIST)(韩国科学技术院)

AI总结 针对室内气体源定位难题,本文提出深度概率框架,结合物理依赖的序列推理,在模拟与真实机器人实验中均优于基线方法,可实现高效准确的在线气体源定位。

Comments 18 pages, 22 figures, 5 tables. Submitted to IEEE Transactions on Robotics

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.15700 2026-08-18 cs.LG cs.AI 版本更新

Contraction-Aware Reinforcement Learning for Nonlinear Control with Statistical Robustness

面向非线性控制的收缩感知强化学习:具备统计鲁棒性

Minjae Cho, Hiroyasu Tsukamoto, Huy T. Tran

机构 * The Grainger College of Engineering University of Illinois Urbana-Champaign, US(格拉inger工程学院伊利诺伊大学厄巴纳-香槟分校)

AI总结 本研究针对非线性路径跟踪问题中控制收缩度量(CCM)策略的长期最优性与鲁棒性缺陷,提出收缩感知强化学习(CARL)算法,将CCM与RL结合,经模拟和真实机器人实验验证其性能与鲁棒性。

Comments Accepted at IEEE Transactions on Robotics

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.17982 2026-08-10 cs.RO cs.CV

HI-SLAM2: Geometry-Aware Gaussian SLAM for Fast Monocular Scene Reconstruction

HI-SLAM2:基于几何的高斯SLAM用于快速单目场景重建

Wei Zhang, Qing Cheng, David Skuddis, Niclas Zeller, Daniel Cremers, Norbert Haala

机构 * Institute for Photogrammetry and Geoinformatics, University of Stuttgart(摄影测量与地理信息学研究所,斯图加特大学) Technical University of Munich(慕尼黑技术大学) Karlsruhe University of Applied Sciences(卡尔斯鲁厄应用科学大学) Munich Center for Machine Learning(慕尼黑机器学习中心)

AI总结 HI-SLAM2通过结合单目先验与学习密集SLAM,利用3D高斯散射实现快速准确的单目场景重建,超越现有神经SLAM和RGB-D方法。

Journal ref IEEE Transactions on Robotics, Volume 41, Pages 6478-6493, Year 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.02834 2026-08-05 cs.RO cs.SY eess.SY 新提交

Biconvex Optimization for Smooth Minimum-Time Trajectories around Convex Obstacles

凸障碍物周围平滑最小时间轨迹的双凸优化

Peter Werner, Tobia Marcucci, Daniela Rus

机构 * Massachusetts Institute of Technology(麻省理工学院) University of California Santa Barbara(加州大学圣巴巴拉分校)

AI总结 本研究提出一种双凸优化方法,用于凸障碍物周围的最小时间运动规划,可保证收敛、支持任意阶导数约束,实验表明其轨迹质量高、鲁棒性强且计算效率与现有方法相当。

Comments 18 pages, 9 figures, 4 tables. Submitted to IEEE Transactions on Robotics. Project page: https://wernerpe.github.io/bmtp-website/ Code: https://github.com/wernerpe/pybmtp

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.21666 2026-08-04 cs.RO cs.CV 版本更新

Uncertainty Quantification for Visual Object Pose Estimation: S-Lemma Ellipsoidal Bounds

视觉目标姿态估计的不确定性量化

Lorenzo Shaikewitz, Charis Georgiou, Luca Carlone

机构 * Massachusetts Institute of Technology(麻省理工学院)

AI总结 SLUE通过凸优化方法为视觉目标姿态估计提供高概率的椭球不确定性界,有效缩小平移界限并保持方向精度。

Comments 18 pages, 9 figures. Code available: https://github.com/MIT-SPARK/PoseUncertaintySets. Published in IEEE Transactions on Robotics

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.17565 2026-07-30 cs.RO cs.SY eess.SY

Aerial Robots Carrying Flexible Cables: Dynamic Shape Optimal Control via Spectral Method Model

Yaolei Shen, Antonio Franchi, Chiara Gabellieri

机构 * University of Twente(特温特大学) Sapienza University of Rome(罗马第一大学)

Journal ref IEEE Transactions on Robotics (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.06636 2026-07-23 cs.RO

Design, Control, and Motion Strategy for DELTA: Transformable Multilink Multirotor for Air-Ground Hybrid Locomotion and Manipulation

DELTA:可变形多连杆多旋翼飞行器的设计、控制与运动策略——用于空地混合运动与操作

Kazuki Sugihara, Moju Zhao, Takuzumi Nishio, Kei Okada, Masayuki Inaba

AI总结 本文提出一种新型多连杆多旋翼机器人DELTA,通过在每个连杆上安装推进器并利用关节驱动,实现了地面滚动、空中飞行及多种环境下的操作能力,并设计了基于非线性优化的实时控制方法和考虑接触约束的运动策略。

Comments 20 pages, 31 figures

Journal ref IEEE Transactions on Robotics, vol. 42, pp. 2765-2785, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.12190 2026-07-21 cs.RO cs.CV

Bundle Adjustment in the Eager Mode

急切模式下的捆绑调整

Zitong Zhan, Huan Xu, Zihang Fang, Xinpeng Wei, Yaoyu Hu, Chen Wang

机构 * Spatial AI & Robotics (SAIR) Lab, University at Buffalo(空间人工智能与机器人实验室,布法罗大学) Georgia Institute of Technology(佐治亚理工学院) Purdue University(普渡大学) Carnegie Mellon University(卡内基梅隆大学)

AI总结 本文提出了一种与PyTorch无缝集成的高效急切模式捆绑调整库,通过稀疏感知的自动微分设计和GPU加速的稀疏运算,提升了在机器人应用中捆绑调整的运行效率和性能。

Journal ref IEEE Transactions on Robotics (T-RO), 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.30318 2026-06-30 cs.RO

Chronos: A Physics-Informed Full-History Framework for Non-Markovian Long-Horizon Manipulation

Chronos: 一种基于物理信息的全历史框架用于非马尔可夫长时域操作

Yulin Zhou, Yimeng Wang, Nengyu Wang, Shaojia Xing, Shiyun Tu, Xiang Li, Jingkai Zhang, Ningbo Jiang, Yuankai Lin, Hua Yang, Xiangrui Zeng, Zhouping Yin

机构 * School of Mechanical Science and Engineering, Huazhong University of Science and Technology(华中科技大学机械科学与工程学院)

AI总结 针对现有策略依赖马尔可夫假设导致记忆依赖任务失败的问题,提出Chronos框架,将观测历史作为策略潜状态,通过选择性状态空间模型传播因果历史状态,并利用二阶薛定谔桥预测加速度场,在16个模拟和4个真实任务中显著超越基线。

Comments 20 pages, 10 figures. Submitted to IEEE Transactions on Robotics

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.20962 2026-06-23 cs.RO cs.MA 新提交

Heterogeneous Policy Networks for Composite Robot Team Communication and Coordination

异构策略网络用于复合机器人团队的通信与协调

Esmaeil Seraj, Rohan Paleja, Luis Pimentel, Kin Man Lee, Zheyuan Wang, Daniel Martin, Matthew Sklar, John Zhang, Zahi Kakish, Matthew Gombolay

机构 * Georgia Institute of Technology(佐治亚理工学院) Carnegie Mellon University(卡内基梅隆大学) Sandia National Laboratories(桑迪亚国家实验室)

AI总结 提出异构策略网络(HetNet),利用异构图注意力网络学习异构机器人团队的高效协作与通信策略,实现端到端训练的二值化消息传递,在多个领域性能提升5.84%至707.65%,通信带宽降低200倍。

Comments IEEE Transactions on Robotics (T-RO)

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.06828 2026-06-23 cs.RO cs.AI 版本更新

NeuPAN: Direct Point Robot Navigation with End-to-End Model-based Learning

NeuPAN: 基于端到端模型学习的直接点机器人导航

Ruihua Han, Shuai Wang, Shuaijun Wang, Zeqing Zhang, Jianjun Chen, Shijie Lin, Chengyang Li, Chengzhong Xu, Yonina C. Eldar, Qi Hao, Jia Pan

AI总结 提出NeuPAN,一种实时、高精度、无地图、易部署的机器人运动规划器,通过紧耦合感知-控制框架和端到端模型学习,直接处理原始点云生成无碰撞运动,避免感知到控制的误差传播。

Comments Accepted by TRO 2025; project website: https://hanruihua.github.io/neupan_project/

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.12681 2026-06-12 cs.RO cs.LG cs.SY eess.SY 版本更新

Adaptive Model-Predictive Control of a Soft Continuum Robot Using a Physics-Informed Neural Network Based on Cosserat Rod Theory

基于Cosserat杆理论物理信息神经网络的软体连续机器人自适应模型预测控制

Johann Licher, Max Bartholdt, Henrik Krauss, Tim-Lukas Habich, Thomas Seel, Moritz Schappler

机构 * Institute of Mechatronic Systems, Leibniz University Hannover(机械系统研究所,汉诺威莱布尼茨大学) Department of Advanced Interdisciplinary Studies, The University of Tokyo(先进跨学科研究部,东京大学) Institute of Assembly Technology and Robotics, Leibniz University of Hannover(组装技术与机器人研究所,汉诺威莱布尼茨大学)

AI总结 提出一种基于域解耦物理信息神经网络(DD-PINN)的实时非线性模型预测控制框架,实现软体连续机器人的高精度动态控制,位置误差低于3 mm。

Comments Submitted to IEEE Transactions on Robotics, 20 pages, 14 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
1902.10607 2026-06-04 cs.RO cs.SY eess.SY

Necessary and Sufficient Conditions for Passivity of Velocity-Sourced Impedance Control of Series Elastic Actuators

速度源阻抗控制系列弹性驱动器被动性的必要和充分条件

Fatih Emre Tosun, Volkan Patoglu

机构 * Faculty of Engineering and Natural Sciences, Sabancı University(工程与自然科学学院,萨班奇大学)

AI总结 本文研究了系列弹性驱动器速度源阻抗控制架构的被动性条件,提出了非保守的设计指南以实现null阻抗和纯弹簧的触觉显示,并强调了在积分控制器中包含物理阻尼的重要性。

Comments Submitted to IEEE T-RO, 12 pages, 10 figures, 7 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
1710.05465 2026-06-04 cs.AI cs.RO cs.SY eess.SY

Flow: A Modular Learning Framework for Mixed Autonomy Traffic

Flow: 一种用于混合自主性的模块化学习框架

Cathy Wu, Aboudy Kreidieh, Kanaad Parvate, Eugene Vinitsky, Alexandre M Bayen

机构 * Laboratory for Information and Decision Systems, Massachusetts Institute of Technology(信息与决策实验室,麻省理工学院) Institute of Data, Systems, and Society, Massachusetts Institute of Technology(数据、系统与社会研究所,麻省理工学院) Department of Mechanical Engineering, University of California, Berkeley(机械工程系,加州大学伯克利分校)

AI总结 本文提出了一种模块化学习框架,利用深度强化学习解决复杂交通动态问题,通过提高系统层面的速度,使学习到的控制法则在仅有4-7%的自动驾驶汽车参与度下,相比人类驾驶性能提升高达57%。此外,在单车道交通中,一个仅使用局部观测的小型神经网络控制法则能够消除拥堵现象,达到近最优性能。

Comments 17 pages, 8 figures, 5 tables. 2021 IEEE Transactions on Robotics (T-RO)

详情

展开后加载摘要…

URL PDF HTML 收藏
1807.08229 2026-06-04 cs.AI cs.RO cs.SY eess.SY

Optimal Continuous State POMDP Planning with Semantic Observations: A Variational Approach

基于语义观测的最优连续状态POMDP规划:一种变分方法

Luke Burks, Ian Loefgren, Nisar Ahmed

AI总结 本文提出了一种基于变分方法的最优规划策略,针对语义观测下的连续状态部分可观测马尔可夫决策过程(CPOMDP)进行改进,通过变分贝叶斯方法解决混合连续-离散概率模型的表示和推理问题,提升了动态决策任务的效率和鲁棒性。

Comments Final version accepted to IEEE Transactions on Robotics (in press as of August 2019)

详情

展开后加载摘要…

URL PDF HTML 收藏
1801.09627 2026-06-04 cs.LG cs.RO cs.SY eess.SY

Barrier-Certified Adaptive Reinforcement Learning with Applications to Brushbot Navigation

具有应用的障碍证书自适应强化学习:Brushbot导航

Motoya Ohnishi, Li Wang, Gennaro Notomista, Magnus Egerstedt

机构 * School of Electrical Engineering, Royal Institute of Technology(皇家理工学院电气工程学院) Georgia Institute of Technology(佐治亚理工学院) RIKEN Center for Advanced Intelligence Project(日本理化学研究所高级智能研究中心) School of Mechanical Engineering(机械工程学院)

AI总结 本文提出了一种安全学习框架,结合自适应模型学习算法和障碍证书,用于具有可能非平稳智能体动态的系统。通过稀疏优化技术提取模型的动态结构,并利用学习的模型结合控制障碍证书来约束策略(反馈控制器),以保持安全性,即避免特定的不利状态空间区域。在某些条件下,保证了在安全被非平稳性破坏后,以李雅普诺夫稳定性的方式恢复安全。此外,将动作-价值函数近似重新公式化,使任何基于内核的非线性函数估计方法都能应用于我们的自适应学习框架。最后,保证了障碍证书策略优化的解是全局最优的,确保在温和条件下进行贪心策略改进。所得到的框架通过四旋翼无人机的模拟进行验证,该无人机此前在安全学习文献中被假设为平稳性,然后在动态未知、高度复杂且非平稳的Brushbot机器人上进行测试。

Comments ©2019 IEEE. Personal use of this material is permitted. Permission from IEEE must be obtained for all other uses, in any current or future media, including reprinting/republishing this material for advertising or promotional purposes, creating new collective works, for resale or redistribution to servers or lists, or reuse of any copyrighted component of this work in other works

Journal ref Published in IEEE Transactions on Robotics, 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1709.03426 2026-06-04 cs.RO cs.SY eess.SY

Trajectory Synthesis for Fisher Information Maximization

信息最大化的轨迹合成

Andrew D. Wilson, Jarvis A. Schultz, Todd D. Murphey

AI总结 本文提出一种连续时间优化方法,通过改进Fisher信息矩阵的范数来生成局部最优轨迹,用于动态系统参数估计,实验验证显示轨迹优化显著提升了参数估计精度。

Comments 12 pages

Journal ref IEEE Transactions on Robotics, vol. 30, no. 6, pp. 1358-1370, 2014

详情

展开后加载摘要…

URL PDF HTML 收藏