arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

机器人 / 具身智能

机器人、具身智能、机器人学习、操作、导航和具身世界模型。

共收录 4134 信号源:cs.RO, cs.AI, cs.CV, cs.LG

1. 模仿学习与强化学习 4134 篇

2504.16417 2025-11-18 eess.SY cs.SY 50%

Anytime Safe Reinforcement Learning

Pol Mestres, Arnau Marzabal, Jorge Cortés

专题命中 模仿学习与强化学习 :navigation(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.16436 2025-11-04 quant-ph 50%

Taming quantum systems: A tutorial for using shortcuts-to-adiabaticity, quantum optimal control, and reinforcement learning

Callum W. Duncan, Pablo M. Poggi, Marin Bukov, Nikolaj Thomas Zinner, Steve Campbell

专题命中 模仿学习与强化学习 :manipulation(abstract)

Comments 73 pages, 15 figures. Data associated with this manuscript version are openly available on Zenodo, https://doi.org/10.5281/zenodo.17169846 ; Jupyter notebooks are available on GitHub, https://github.com/nqd-lab/quctrl-tutorial

Journal ref PRX Quantum 6, 040201 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.20907 2025-10-30 eess.SY cs.SY 50%

Combining Deep Reinforcement Learning with a Jerk-Bounded Trajectory Generator for Kinematically Constrained Motion Planning

Seyed Adel Alizadeh Kolagar, Mehdi Heydari Shahna, Jouni Mattila

专题命中 模仿学习与强化学习 :robotic(abstract)

Comments This paper has been submitted to the IEEE for potential publication

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01492 2025-10-03 eess.SY cs.SY 50%

Off-Policy Reinforcement Learning with Anytime Safety Guarantees via Robust Safe Gradient Flow

Pol Mestres, Arnau Marzabal, Jorge Cortés

专题命中 模仿学习与强化学习 :navigation(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.23425 2025-09-30 cs.MA 50%

Situational Awareness for Safe and Robust Multi-Agent Interactions Under Uncertainty

Benjamin Alcorn, Eman Hammad

专题命中 模仿学习与强化学习 :robotics(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.03479 2025-09-04 cs.CL 50%

Design and Optimization of Reinforcement Learning-Based Agents in Text-Based Games

Haonan Wang, Mingjia Zhao, Junfeng Sun, Wei Liu

专题命中 模仿学习与强化学习 :world model(abstract)

Comments 6 papges

Journal ref Copyright (c) 2025 International Journal of Computer Science and Information Technology International Journal of Computer Science and Information Technology International Journal of Computer Science and Information Technology

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.17185 2025-08-26 math.OC cs.SY eess.SY 50%

Linear Dynamics meets Linear MDPs: Closed-Form Optimal Policies via Reinforcement Learning

Abed AlRahman Al Makdah, Oliver Kosut, Lalitha Sankar, Shaofeng Zou

专题命中 模仿学习与强化学习 :robotics(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.05236 2025-08-08 cs.MA 50%

Towards Language-Augmented Multi-Agent Deep Reinforcement Learning

Maxime Toquebiau, Jae-Yun Jun, Faïz Benamar, Nicolas Bredeche

专题命中 模仿学习与强化学习 :embodied agent(abstract)

Comments Accespted at the European Conference on Artificial Intelligence 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.00645 2025-08-04 physics.flu-dyn physics.comp-ph 50%

SmartFlow: A CFD-solver-agnostic deep reinforcement learning framework for computational fluid dynamics on HPC platforms

Maochao Xiao, Yuning Wang, Felix Rodach, Bernat Font, Marius Kurz, Pol Suárez, Di Zhou, Francisco Alcántara-Ávila, Ting Zhu, Junle Liu, Ricard Montalà, Jiawei Chen, Jean Rabault, Oriol Lehmkuhl, Andrea Beck, Johan Larsson, Ricardo Vinuesa, Sergio Pirozzoli

专题命中 模仿学习与强化学习 :navigation(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23080 2025-08-01 cs.MA 50%

Causal-Inspired Multi-Agent Decision-Making via Graph Reinforcement Learning

Jing Wang, Yan Jin, Fei Ding, Chongfeng Wei

专题命中 模仿学习与强化学习 :navigation(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.04053 2025-07-10 quant-ph 50%

Enhanced Qubit Readout via Reinforcement Learning

Aniket Chatterjee, Jonathan Schwinger, Yvonne Y. Gao

专题命中 模仿学习与强化学习 :manipulation(abstract)

Comments 11 pages, 4 figures

Journal ref Phys. Rev. Applied 23, 054057 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.13543 2025-07-09 cs.MA cs.CY 50%

Origin-Destination Pattern Effects on Large-Scale Mixed Traffic Control via Multi-Agent Reinforcement Learning

Muyang Fan, Songyang Liu, Shuai Li, Weizi Li

专题命中 模仿学习与强化学习 :robotic(abstract)

Comments Accepted to IEEE International Conference on Intelligent Transportation Systems (ITSC), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04998 2025-07-08 q-bio.NC 50%

Linking Homeostasis to Reinforcement Learning: Internal State Control of Motivated Behavior

Naoto Yoshida, Henning Sprekeler, Boris Gutkin

专题命中 模仿学习与强化学习 :robotic(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04160 2025-07-08 cs.HC 50%

HyperSumm-RL: A Dialogue Summarization Framework for Modeling Leadership Perception in Social Robots

Subasish Das

专题命中 模仿学习与强化学习 :navigation(abstract)

Comments 6 pages with references

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.03948 2025-07-08 cond-mat.stat-mech cond-mat.soft 50%

Kinetic theory of decentralized learning for smart active matter

Gerhard Jung, Misaki Ozawa, Eric Bertin

专题命中 模仿学习与强化学习 :robotics(abstract)

Comments Phys. Rev. Lett. 134, 248302 (Editors' Suggestion)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.22112 2025-07-01 cs.IR 50%

Reward Balancing Revisited: Enhancing Offline Reinforcement Learning for Recommender Systems

Wenzheng Shu, Yanxiang Zeng, Yongxiang Tang, Teng Sha, Ning Luo, Yanhua Cheng, Xialong Liu, Fan Zhou, Peng Jiang

专题命中 模仿学习与强化学习 :world model(abstract)

Comments Accepted in Companion Proceedings of the ACM Web Conference 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.12497 2025-06-19 eess.SY cs.SY 50%

Wasserstein-Barycenter Consensus for Cooperative Multi-Agent Reinforcement Learning

Ali Baheri

专题命中 模仿学习与强化学习 :navigation(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.01710 2025-06-03 cs.CL 50%

Reasoning-Table: Exploring Reinforcement Learning for Table Reasoning

Fangyu Lei, Jinxiang Meng, Yiming Huang, Tinghong Chen, Yun Zhang, Shizhu He, Jun Zhao, Kang Liu

机构 * Institute of Automation, CAS(中国科学院自动化研究所) University of Chinese Academy of Sciences(中国科学院大学)

专题命中 模仿学习与强化学习 :manipulation(abstract)

Comments Work in progress

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.12486 2025-05-29 cs.CL 50%

EPO: Explicit Policy Optimization for Strategic Reasoning in LLMs via Reinforcement Learning

Xiaoqian Liu, Ke Wang, Yongbin Li, Yuchuan Wu, Wentao Ma, Aobo Kong, Fei Huang, Jianbin Jiao, Junge Zhang

机构 * University of Chinese Academy of Sciences(中国科学院大学) Tongyi Lab(通义实验室) The Key Laboratory of Cognition and Decision Intelligence for Complex Systems, Institute of Automation, Chinese Academy of Sciences(复杂系统认知与决策智能重点实验室,中国科学院自动化研究所)

专题命中 模仿学习与强化学习 :navigation(abstract)

Comments ACL2025 main

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.07257 2025-05-13 cs.IR 50%

DARLR: Dual-Agent Offline Reinforcement Learning for Recommender Systems with Dynamic Reward

Yi Zhang, Ruihong Qiu, Xuwei Xu, Jiajun Liu, Sen Wang

专题命中 模仿学习与强化学习 :world model(abstract)

Comments SIGIR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.23290 2025-04-01 cs.NI 50%

Efficient Twin Migration in Vehicular Metaverses: Multi-Agent Split Deep Reinforcement Learning with Spatio-Temporal Trajectory Generation

Junlong Chen, Jiawen Kang, Minrui Xu, Fan Wu, Hongliang Zhang, Huawei Huang, Dusit Niyato, Shiwen Mao

专题命中 模仿学习与强化学习 :navigation(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.10395 2025-03-14 physics.ins-det 50%

Optical stabilization for laser communication satellite systems through proportional-integral-derivative (PID) control and reinforcement learning approach

A. Reutov, S. Vorobey, A. Katanskiy, V. Balakirev, R. Bakhshaliev, K. Barbyshev, V. Merzlinkin, V. Tekaev

专题命中 模仿学习与强化学习 :navigation(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.00385 2025-03-04 eess.SY cs.SY 50%

Model-Agnostic Meta-Policy Optimization via Zeroth-Order Estimation: A Linear Quadratic Regulator Perspective

Yunian Pan, Tao Li, Quanyan Zhu

专题命中 模仿学习与强化学习 :robotics(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.14995 2025-02-24 eess.IV 50%

Reinforcement Learning for Ultrasound Image Analysis A Comprehensive Review of Advances and Applications

Maha Ezzelarab, Midhila Madhusoodanan, Shrimanti Ghosh, Geetika Vadali, Jacob Jaremko, Abhilash Hareendranathan

专题命中 模仿学习与强化学习 :navigation(abstract)

Comments 36 pages, 4 figures, 1 table

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.07112 2025-02-12 eess.SP 50%

Smell of Source: Learning-Based Odor Source Localization with Molecular Communication

Ayse Sila Okcu, Ozgur B. Akan

专题命中 模仿学习与强化学习 :robotics(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.10990 2025-01-22 cs.DL cs.SI physics.soc-ph 50%

Societal citations undermine the function of the science reward system

Xiaokai Li, An Zeng, Ying Fan

专题命中 模仿学习与强化学习 :manipulation(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.19639 2025-01-19 cs.MA 50%

RMIO: A Model-Based MARL Framework for Scenarios with Observation Loss in Some Agents

Zifeng Shi, Meiqin Liu, Senlin Zhang, Ronghao Zheng, Shanling Dong

专题命中 模仿学习与强化学习 :world model(abstract)

Comments 17 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.16985 2024-12-11 quant-ph 50%

Model-aware reinforcement learning for high-performance Bayesian experimental design in quantum metrology

Federico Belliardo, Fabio Zoratti, Florian Marquardt, Vittorio Giovannetti

专题命中 模仿学习与强化学习 :manipulation(abstract)

Comments 45 pages, 10 figures

Journal ref Quantum 8, 1555 (2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.03501 2024-11-07 cs.MS cs.NA cs.SY eess.SY math.NA 50%

The Python LevelSet Toolbox (LevelSetPy)

Lekan Molu

专题命中 模仿学习与强化学习 :robotics(abstract)

Journal ref The 63rd IEEE Conference on Decision and Control, Milan, 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.01060 2024-09-04 cs.CE 50%

Multiagent Reinforcement Learning Enhanced Decision-making of Crew Agents During Floor Construction Process

Bin Yang, Boda Liu, Yilong Han, Xin Meng, Yifan Wang, Hansi Yang, Jianzhuang Xia

专题命中 模仿学习与强化学习 :robotics(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏