arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

机器人 / 具身智能

机器人、具身智能、机器人学习、操作、导航和具身世界模型。

共收录 4134 信号源:cs.RO, cs.AI, cs.CV, cs.LG

1. 模仿学习与强化学习 4134 篇

2203.10519 2022-07-20 cs.RO cs.LG cs.NE 62%

Reinforcement learning reward function in unmanned aerial vehicle control tasks

Mikhail S. Tovarnov, Nikita V. Bykov

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.RO、cs.LG

Journal ref J. Phys.: Conf. Ser. 2308 012004 (2022)

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.02007 2022-07-08 cs.LG cs.AI 62%

The StarCraft Multi-Agent Challenges+ : Learning of Multi-Stage Tasks and Environmental Factors without Precise Reward Functions

Mingyu Kim, Jihwan Oh, Yongsik Lee, Joonkee Kim, Seonghwan Kim, Song Chong, Se-Young Yun

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.AI、cs.LG

Comments ICML Workshop: AI for Agent Based Modeling 2022 Spotlight

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.00978 2022-07-05 cs.LG cs.RO 62%

Renaissance Robot: Optimal Transport Policy Fusion for Learning Diverse Skills

Julia Tan, Ransalu Senanayake, Fabio Ramos

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.RO、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.15233 2022-06-29 cs.RO cs.LG 62%

Solving the Real Robot Challenge using Deep Reinforcement Learning

Robert McCarthy, Francisco Roldan Sanchez, Qiang Wang, David Cordova Bulens, Kevin McGuinness, Noel O'Connor, Stephen J. Redmond

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO、cs.LG

Comments Published in AICS 2021 (http://ceur-ws.org/Vol-3105/paper41.pdf). Paper updated to clarify procedure used to train the policy

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.12030 2022-06-27 cs.LG cs.AI 62%

Phasic Self-Imitative Reduction for Sparse-Reward Goal-Conditioned Reinforcement Learning

Yunfei Li, Tian Gao, Jiaqi Yang, Huazhe Xu, Yi Wu

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.AI、cs.LG

Comments Accepted at ICML 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.11795 2022-06-24 cs.LG cs.AI 62%

Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos

Bowen Baker, Ilge Akkaya, Peter Zhokhov, Joost Huizinga, Jie Tang, Adrien Ecoffet, Brandon Houghton, Raul Sampedro, Jeff Clune

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.11299 2022-06-24 cs.LG cs.RO 62%

Latent Policies for Adversarial Imitation Learning

Tianyu Wang, Nikhil Karnwal, Nikolay Atanasov

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.RO、cs.LG

Comments 8 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.03654 2022-06-09 cs.NE cs.AI cs.LG 62%

Solving the Spike Feature Information Vanishing Problem in Spiking Deep Q Network with Potential Based Normalization

Yinqian Sun, Yi Zeng, Yang Li

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.03312 2022-06-08 cs.NE cs.AI cs.LG 62%

Neuro-Nav: A Library for Neurally-Plausible Reinforcement Learning

Arthur Juliani, Samuel Barnett, Brandon Davis, Margaret Sereno, Ida Momennejad

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.04156 2022-06-07 cs.LG cs.AI 62%

Showing Your Offline Reinforcement Learning Work: Online Evaluation Budget Matters

Vladislav Kurenkov, Sergey Kolesnikov

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI、cs.LG

Comments ICML 2022, Spotlight; https://tinkoff-ai.github.io/eop/

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.00142 2022-06-02 cs.LG cs.AI cs.CL 62%

IGLU Gridworld: Simple and Fast Environment for Embodied Dialog Agents

Artem Zholus, Alexey Skrynnik, Shrestha Mohanty, Zoya Volovikova, Julia Kiseleva, Artur Szlam, Marc-Alexandre Coté, Aleksandr I. Panov

专题命中 模仿学习与强化学习 :embodied agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.15023 2022-05-31 cs.LG cs.AI 62%

Scalable Multi-Agent Model-Based Reinforcement Learning

Vladimir Egorov, Aleksei Shpilman

专题命中 模仿学习与强化学习 :world model(abstract);分类 cs.AI、cs.LG

Comments AAMAS'2022, cite https://dl.acm.org/doi/abs/10.5555/3535850.3535894

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.12401 2022-05-26 cs.LG cs.AI 62%

Reward Uncertainty for Exploration in Preference-based Reinforcement Learning

Xinran Liang, Katherine Shu, Kimin Lee, Pieter Abbeel

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.AI、cs.LG

Comments ICLR 2022. Last two authors advised equally

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.10787 2022-05-24 cs.LG cs.AI 62%

A Dirichlet Process Mixture of Robust Task Models for Scalable Lifelong Reinforcement Learning

Zhi Wang, Chunlin Chen, Daoyi Dong

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.AI、cs.LG

Comments Manuscript accepted by IEEE Transactions on Cybernetics, 2022, DOI: DOI: 10.1109/TCYB.2022.3170485

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.10464 2022-05-24 cs.AI cs.LO cs.RO 62%

Synthesis from Satisficing and Temporal Goals

Suguman Bansal, Lydia Kavraki, Moshe Y. Vardi, Andrew Wells

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.08013 2022-05-23 cs.LG cs.CV 62%

Continual learning on 3D point clouds with random compressed rehearsal

Maciej Zamorski, Michał Stypułkowski, Konrad Karanowski, Tomasz Trzciński, Maciej Zięba

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.CV、cs.LG

Comments 10 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.09448 2022-05-20 cs.AI cs.CV 62%

Image Augmentation Based Momentum Memory Intrinsic Reward for Sparse Reward Visual Scenes

Zheng Fang, Biao Zhao, Guizhong Liu

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.AI、cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2111.08575 2022-05-19 cs.RO cs.CV 62%

GRI: General Reinforced Imitation and its Application to Vision-Based Autonomous Driving

Raphael Chekroun, Marin Toromanoff, Sascha Hornauer, Fabien Moutarde

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.RO、cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.13970 2022-05-16 cs.LG cs.AI 62%

Intrinsically Motivated Self-supervised Learning in Reinforcement Learning

Yue Zhao, Chenzhuang Du, Hang Zhao, Tiejun Li

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2112.03482 2022-05-13 cs.LG cs.AI cs.HC 62%

Combining Learning from Human Feedback and Knowledge Engineering to Solve Hierarchical Tasks in Minecraft

Vinicius G. Goecks, Nicholas Waytowich, David Watkins-Valls, Bharat Prakash

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.AI、cs.LG

Comments Submitted to the AAAI 2022 Spring Symposium on Machine Learning and Knowledge Engineering for Hybrid Intelligence (AAAI-MAKE 2022)

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.05230 2022-05-12 cs.LG cs.AI 62%

Developing cooperative policies for multi-stage reinforcement learning tasks

Jordan Erskine, Chris Lehnert

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.AI、cs.LG

Comments This paper supersedes the rejected paper "Developing cooperative policies for multi-stage tasks". arXiv admin note: substantial text overlap with arXiv:2007.00203

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.11296 2022-04-25 cs.LG cs.AI 62%

Reinforcement Learning in Practice: Opportunities and Challenges

Yuxi Li

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.08573 2022-04-20 cs.LG cs.RO 62%

Training and Evaluation of Deep Policies using Reinforcement Learning and Generative Models

Ali Ghadirzadeh, Petra Poklukar, Karol Arndt, Chelsea Finn, Ville Kyrki, Danica Kragic, Mårten Björkman

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.RO、cs.LG

Comments arXiv admin note: substantial text overlap with arXiv:2007.13134

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.13733 2022-04-13 cs.RO cs.LG 62%

Blocks Assemble! Learning to Assemble with Large-Scale Structured Reinforcement Learning

Seyed Kamyar Seyed Ghasemipour, Daniel Freeman, Byron David, Shixiang Shane Gu, Satoshi Kataoka, Igor Mordatch

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.RO、cs.LG

Comments Accompanying project webpage can be found at: https://sites.google.com/view/learning-direct-assembly

详情

展开后加载摘要…

URL PDF HTML 收藏
2112.05923 2022-04-05 cs.LG cs.AI cs.DC 62%

ElegantRL-Podracer: Scalable and Elastic Library for Cloud-Native Deep Reinforcement Learning

Xiao-Yang Liu, Zechu Li, Zhuoran Yang, Jiahao Zheng, Zhaoran Wang, Anwar Walid, Jian Guo, Michael I. Jordan

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.AI、cs.LG

Comments 9 pages, 7 figures

Journal ref Deep Reinforcement Learning Workshop, NeurIPS 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.04605 2022-04-01 cs.LG cs.AI cs.NE 62%

Instance Weighted Incremental Evolution Strategies for Reinforcement Learning in Dynamic Environments

Zhi Wang, Chunlin Chen, Daoyi Dong

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.AI、cs.LG

Comments Accepted by IEEE Transactions on Neural Networks and Learning Systems, 2022, DOI: 10.1109/TNNLS.2022.3160173

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.07454 2022-03-16 cs.LG cs.AI 62%

L2Explorer: A Lifelong Reinforcement Learning Assessment Environment

Erik C. Johnson, Eric Q. Nguyen, Blake Schreurs, Chigozie S. Ewulum, Chace Ashcraft, Neil M. Fendley, Megan M. Baker, Alexander New, Gautam K. Vallabha

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI、cs.LG

Comments 10 Pages submitted to AAAI AI for Open Worlds Symposium 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.07723 2022-03-10 cs.LG cs.AI 62%

Targeted Attack on Deep RL-based Autonomous Driving with Learned Visual Patterns

Prasanth Buddareddygari, Travis Zhang, Yezhou Yang, Yi Ren

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.AI、cs.LG

Comments 7 pages, 4 figures; Accepted at ICRA 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.04272 2022-03-09 cs.LG cs.AI stat.ME 62%

Policy-Based Bayesian Experimental Design for Non-Differentiable Implicit Models

Vincent Lim, Ellen Novoseller, Jeffrey Ichnowski, Huang Huang, Ken Goldberg

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI、cs.LG

Comments 15 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.02857 2022-03-08 cs.LG cs.RO cs.SY eess.SY 62%

Leveraging Reward Gradients For Reinforcement Learning in Differentiable Physics Simulations

Sean Gillen, Katie Byl

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏