arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

共收录 93793 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. 规划决策 35686 篇

2005.09611 2021-07-29 cs.RO cs.AI cs.IT math.IT 83%

Information-Theoretic Abstractions for Planning in Agents with Computational Constraints

Daniel T. Larsson, Dipankar Maity, Panagiotis Tsiotras

专题命中 规划决策 :planning(title,abstract);agent(abstract);分类 cs.AI

Journal ref 2021 IEEE Robotics and Automation Letters (RA-L)

详情

展开后加载摘要…

URL PDF HTML 收藏
2107.12544 2021-07-28 cs.AI 83%

Human-Level Reinforcement Learning through Theory-Based Modeling, Exploration, and Planning

Pedro A. Tsividis, Joao Loula, Jake Burga, Nathan Foss, Andres Campero, Thomas Pouncy, Samuel J. Gershman, Joshua B. Tenenbaum

专题命中 规划决策 :planning(title,abstract);agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2011.06450 2021-07-23 cs.AI cs.CV 83%

A deep Q-Learning based Path Planning and Navigation System for Firefighting Environments

Manish Bhattarai, Manel Martinez-Ramon

专题命中 规划决策 :planning(title,abstract);agent(abstract);分类 cs.AI

Comments Accepted to ICAART2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2105.04088 2021-05-11 cs.AI 83%

PEARL: Parallelized Expert-Assisted Reinforcement Learning for Scene Rearrangement Planning

Hanqing Wang, Zan Wang, Wei Liang, Lap-Fai Yu

专题命中 规划决策 :planning(title,abstract);agent(abstract);分类 cs.AI

Comments 7 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2104.07276 2021-04-16 cs.AI 83%

Adaptive Belief Discretization for POMDP Planning

Divya Grover, Christos Dimitrakakis

专题命中 规划决策 :planning(title,abstract);agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.10642 2021-04-12 cs.AI cs.RO 83%

Knowledge-Based Hierarchical POMDPs for Task Planning

Sergio A. Serrano, Elizabeth Santiago, Jose Martinez-Carranza, Eduardo Morales, L. Enrique Sucar

专题命中 规划决策 :planning(title,abstract);agent(abstract);分类 cs.AI

Journal ref Journal of Intelligent & Robotic Systems 101 (2021) 1-30

详情

展开后加载摘要…

URL PDF HTML 收藏
2011.08345 2021-04-07 cs.LG cs.RO 83%

Distilling a Hierarchical Policy for Planning and Control via Representation and Reinforcement Learning

Jung-Su Ha, Young-Jin Park, Hyeok-Joo Chae, Soon-Seo Park, Han-Lim Choi

专题命中 规划决策 :planning(title,abstract);agent(abstract);分类 cs.LG

Comments ICRA 2021, the first two authors contributed equally

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.14489 2021-03-29 cs.AI cs.FL cs.SY eess.SY 83%

Probabilistic Planning with Preferences over Temporal Goals

Jie Fu

专题命中 规划决策 :planning(title,abstract);agent(abstract);分类 cs.AI

Comments 6 pages, 8 figures, Accepted by American Control Conference (ACC) 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.01171 2021-03-26 cs.AI 83%

Expected Value of Communication for Planning in Ad Hoc Teamwork

William Macke, Reuth Mirsky, Peter Stone

专题命中 规划决策 :planning(title,abstract);autonomous agent(abstract);分类 cs.AI

Comments 10 pages, 6 figure, Published at AAAI 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2102.07185 2021-02-16 cs.MA cs.AI 83%

Partial Disclosure of Private Dependencies in Privacy Preserving Planning

Rotem Lev Lehman, Guy Shani, Roni Stern

专题命中 规划决策 :planning(title,abstract);agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2102.06284 2021-02-15 cs.RO cs.LG 83%

Large Scale Distributed Collaborative Unlabeled Motion Planning with Graph Policy Gradients

Arbaaz Khan, Vijay Kumar, Alejandro Ribeiro

专题命中 规划决策 :planning(title);agent(abstract);multi-agent(abstract);分类 cs.LG

Comments arXiv admin note: substantial text overlap with arXiv:1909.10704

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.12976 2021-02-09 cs.LG cs.SY eess.SY math.OC stat.ML 83%

Learning and Planning for Time-Varying MDPs Using Maximum Likelihood Estimation

Melkior Ornik, Ufuk Topcu

专题命中 规划决策 :planning(title,abstract);agent(abstract);分类 cs.LG

Comments To be published in Journal of Machine Learning Research

详情

展开后加载摘要…

URL PDF HTML 收藏
2102.00834 2021-02-02 cs.AI 83%

Counterfactual Planning in AGI Systems

Koen Holtman

专题命中 规划决策 :planning(title,abstract);agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.00544 2021-01-28 cs.LG cs.IT cs.RO eess.SP math.IT stat.ML 83%

UAV Path Planning for Wireless Data Harvesting: A Deep Reinforcement Learning Approach

Harald Bayerlein, Mirco Theile, Marco Caccamo, David Gesbert

专题命中 规划决策 :planning(title,abstract);agent(abstract);分类 cs.LG

Comments Code available under https://github.com/hbayerlein/uav_data_harvesting, IEEE Global Communications Conference (GLOBECOM) 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.07517 2021-01-15 cs.RO cs.HC cs.LG cs.SY eess.SY 83%

MATS: An Interpretable Trajectory Forecasting Representation for Planning and Control

Boris Ivanovic, Amine Elhafsi, Guy Rosman, Adrien Gaidon, Marco Pavone

专题命中 规划决策 :planning(title,abstract);agent(abstract);分类 cs.LG

Comments 14 pages, 6 figures, 1 table. All code, models, and data can be found at https://github.com/StanfordASL/MATS . Conference on Robot Learning (CoRL) 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2012.14464 2021-01-01 cs.RO cs.AI 83%

Disentangled Planning and Control in Vision Based Robotics via Reward Machines

Alberto Camacho, Jacob Varley, Deepali Jain, Atil Iscen, Dmitry Kalashnikov

专题命中 规划决策 :planning(title,abstract);agent(abstract);分类 cs.AI

Comments Accepted to the Deep Reinforcement Learning Workshop at Neural Information Processing Systems (2020)

详情

展开后加载摘要…

URL PDF HTML 收藏
2012.13037 2020-12-25 cs.AI 83%

SPOTTER: Extending Symbolic Planning Operators through Targeted Reinforcement Learning

Vasanth Sarathy, Daniel Kasenberg, Shivam Goel, Jivko Sinapov, Matthias Scheutz

专题命中 规划决策 :planning(title,abstract);agent(abstract);分类 cs.AI

Comments Accepted to AAMAS 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2011.13098 2020-11-30 cs.RO cs.AI cs.SY eess.SY 83%

An End-to-end Deep Reinforcement Learning Approach for the Long-term Short-term Planning on the Frenet Space

Majid Moghadam, Ali Alizadeh, Engin Tekin, Gabriel Hugh Elkaim

专题命中 规划决策 :planning(title,abstract);agent(abstract);分类 cs.AI

Comments submitted to International Conference on Robotics and Automation (ICRA 2021)

详情

展开后加载摘要…

URL PDF HTML 收藏
2011.03590 2020-11-30 cs.RO cs.LG cs.SY eess.SY 83%

Reactive motion planning with probabilistic safety guarantees

Yuxiao Chen, Ugo Rosolia, Chuchu Fan, Aaron D. Ames, Richard Murray

专题命中 规划决策 :planning(title,abstract);agent(abstract);分类 cs.LG

Comments In the Conference on Robotic Learning 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2011.09034 2020-11-19 cs.AI cs.RO 83%

Domain Concretization from Examples: Addressing Missing Domain Knowledge via Robust Planning

Akshay Sharma, Piyush Rajesh Medikeri, Yu Zhang

专题命中 规划决策 :planning(title,abstract);agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.14259 2020-10-28 cs.CL cs.CV 83%

Visually-Grounded Planning without Vision: Language Models Infer Detailed Plans from High-level Instructions

Peter A. Jansen

专题命中 规划决策 :planning(title,abstract);agent(abstract);分类 cs.CL

Comments Accepted to Findings of EMNLP. V2: corrected typo Table 1; margins Table 3

详情

展开后加载摘要…

URL PDF HTML 收藏
2006.07532 2020-10-27 cs.AI 83%

Online Bayesian Goal Inference for Boundedly-Rational Planning Agents

Tan Zhi-Xuan, Jordyn L. Mann, Tom Silver, Joshua B. Tenenbaum, Vikash K. Mansinghka

专题命中 规划决策 :planning(title,abstract);agent(abstract);分类 cs.AI

Comments Accepted to NeurIPS 2020. 10 pages (excl. references), 6 figures/tables. (Supplement: 8 pages, 11 figures/tables). Code available at: https://github.com/ztangent/Plinf.jl

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.12639 2020-08-31 cs.RO cs.AI 83%

Path Planning for Shepherding a Swarm in a Cluttered Environment using Differential Evolution

Saber Elsayed, Hemant Singh, Essam Debie, Anthony Perry, Benjamin Campbell, Robert Hunjet, Hussein Abbass

专题命中 规划决策 :planning(title,abstract);agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2004.12960 2020-08-04 cs.HC cs.AI 83%

Tradeoff-Focused Contrastive Explanation for MDP Planning

Roykrong Sukkerd, Reid Simmons, David Garlan

专题命中 规划决策 :planning(title,abstract);agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
1611.00230 2020-07-16 cs.RO cs.AI 83%

Towards Blended Reactive Planning and Acting using Behavior Trees

Michele Colledanchise, Diogo Almeida, Petter Ögren

专题命中 规划决策 :planning(title,abstract);agent(abstract);分类 cs.AI

Journal ref 2019 International Conference on Robotics and Automation (ICRA)

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.05655 2020-07-14 cs.CV cs.AI cs.RO 83%

Evolving Graphical Planner: Contextual Global Planning for Vision-and-Language Navigation

Zhiwei Deng, Karthik Narasimhan, Olga Russakovsky

专题命中 规划决策 :planning(title,abstract);agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
1912.06787 2020-07-08 cs.RO cs.AI 83%

PODDP: Partially Observable Differential Dynamic Programming for Latent Belief Space Planning

Dicong Qiu, Yibiao Zhao, Chris L. Baker

专题命中 规划决策 :planning(title,abstract);autonomous agent(abstract);分类 cs.AI

Comments 16 pages, 6 figures, preprint

Journal ref Robotics: Science and Systems, 2020. 69.1-69.10

详情

展开后加载摘要…

URL PDF HTML 收藏
2002.03327 2020-06-24 cs.LG stat.ML 83%

Reward Tweaking: Maximizing the Total Reward While Planning for Short Horizons

Chen Tessler, Shie Mannor

专题命中 规划决策 :planning(title,abstract);agent(abstract);分类 cs.LG

Comments Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2005.03770 2020-05-11 cs.LG cs.CV stat.ML 83%

Planning from Images with Deep Latent Gaussian Process Dynamics

Nathanael Bosch, Jan Achterhold, Laura Leal-Taixé, Jörg Stückler

专题命中 规划决策 :planning(title,abstract);agent(abstract);分类 cs.LG

Comments Accepted for publication at the 2nd Annual Conference on Learning for Dynamics and Control (L4DC) 2020, with supplementary material. First two authors contributed equally

详情

展开后加载摘要…

URL PDF HTML 收藏
1912.02945 2020-04-24 cs.LG cs.MA cs.RO stat.ML 83%

A pedestrian path-planning model in accordance with obstacle's danger with reinforcement learning

Thanh-Trung Trinh, Dinh-Minh Vu, Masaomi Kimura

专题命中 规划决策 :planning(title,abstract);agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏