arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

共收录 15819 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. Agent评测 15819 篇

1911.08111 2020-02-06 cs.IT cs.AI cs.LG math.IT 62%

Placement Optimization of Aerial Base Stations with Deep Reinforcement Learning

Jin Qiu, Jiangbin Lyu, Liqun Fu

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments 6 pages, 4 figures, accepted for publication in 2020 IEEE International Conference on Communications (ICC 2020)

详情

展开后加载摘要…

URL PDF HTML 收藏
2002.00149 2020-02-04 cs.LG cs.AI 62%

Periodic Intra-Ensemble Knowledge Distillation for Reinforcement Learning

Zhang-Wei Hong, Prabhat Nagarajan, Guilherme Maeda

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments 8 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.02690 2020-02-03 cs.CL cs.AI 62%

SIMMC: Situated Interactive Multi-Modal Conversational Data Collection And Evaluation Platform

Paul A. Crook, Shivani Poddar, Ankita De, Semir Shafi, David Whitney, Alborz Geramifard, Rajen Subba

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.CL

Comments ASRU 2019 (demonstration)

详情

展开后加载摘要…

URL PDF HTML 收藏
2001.08034 2020-01-23 cs.CL cs.AI cs.CV 62%

ManyModalQA: Modality Disambiguation and QA over Diverse Inputs

Darryl Hannan, Akshay Jain, Mohit Bansal

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.CL

Comments AAAI 2020 (10 pages)

详情

展开后加载摘要…

URL PDF HTML 收藏
1909.12892 2020-01-23 cs.LG cs.AI stat.ML 62%

Automated curricula through setter-solver interactions

Sebastien Racaniere, Andrew K. Lampinen, Adam Santoro, David P. Reichert, Vlad Firoiu, Timothy P. Lillicrap

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Journal ref International Conference on Learning Representations, 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
1807.06736 2020-01-14 cs.CL cs.LG cs.SD eess.AS 62%

Forward Attention in Sequence-to-sequence Acoustic Modelling for Speech Synthesis

Jing-Xuan Zhang, Zhen-Hua Ling, Li-Rong Dai

专题命中 Agent评测 :agent(abstract);分类 cs.CL、cs.LG

Comments 5 pages, 3 figures, 2 tables. Published in IEEE International Conference on Acoustics, Speech and Signal Processing 2018 (ICASSP2018)

Journal ref IEEE International Conference on Acoustics, Speech and Signal Processing (2018) 4789-4793

详情

展开后加载摘要…

URL PDF HTML 收藏
1909.09849 2020-01-13 cs.MA cs.AI cs.LG 62%

Multiagent Evaluation under Incomplete Information

Mark Rowland, Shayegan Omidshafiei, Karl Tuyls, Julien Perolat, Michal Valko, Georgios Piliouras, Remi Munos

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2001.00234 2020-01-03 quant-ph cs.AI cs.LG 62%

Reinforcement Quantum Annealing: A Quantum-Assisted Learning Automata Approach

Ramin Ayanzadeh, Milton Halem, Tim Finin

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1912.12294 2019-12-30 cs.RO cs.AI cs.CV cs.LG 62%

Learning by Cheating

Dian Chen, Brady Zhou, Vladlen Koltun, Philipp Krähenbühl

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Paper published in CoRL2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1912.10113 2019-12-24 eess.SY cs.AI cs.LG cs.SY 62%

Teaching robots to perceive time -- A reinforcement learning approach (Extended version)

Inês Lourenço, Bo Wahlberg, Rodrigo Ventura

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments 13 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
1806.01946 2019-12-24 cs.AI cs.LG 62%

Learning to Understand Goal Specifications by Modelling Reward

Dzmitry Bahdanau, Felix Hill, Jan Leike, Edward Hughes, Arian Hosseini, Pushmeet Kohli, Edward Grefenstette

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments 19 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
1810.08272 2019-12-20 cs.AI cs.CL 62%

BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning

Maxime Chevalier-Boisvert, Dzmitry Bahdanau, Salem Lahlou, Lucas Willems, Chitwan Saharia, Thien Huu Nguyen, Yoshua Bengio

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.CL

Comments Accepted at ICLR 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1912.06101 2019-12-13 cs.LG cs.AI 62%

The PlayStation Reinforcement Learning Environment (PSXLE)

Carlos Purves, Cătălina Cangea, Petar Veličković

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1912.04472 2019-12-11 cs.LG cs.AI stat.ML 62%

Deep Bayesian Reward Learning from Preferences

Daniel S. Brown, Scott Niekum

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Workshop on Safety and Robustness in Decision Making at the 33rd Conference on Neural Information Processing Systems (NeurIPS) 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1905.13469 2019-12-10 cs.LG cs.AI cs.NE 62%

Interval timing in deep reinforcement learning agents

Ben Deverett, Ryan Faulkner, Meire Fortunato, Greg Wayne, Joel Z. Leibo

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments 11 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
1912.00074 2019-12-03 cs.LG cs.AI stat.ML 62%

Quadratic Q-network for Learning Continuous Control for Autonomous Vehicles

Pin Wang, Hanhan Li, Ching-Yao Chan

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Machine Learning for Autonomous Driving Workshop on NeurIPS, 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1905.07579 2019-12-03 cs.LG cs.AI stat.ML 62%

Combining Experience Replay with Exploration by Random Network Distillation

Francesco Sovrano

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments 8 pages, 6 figures, accepted as full-paper at IEEE Conference on Games (CoG) 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.11260 2019-11-27 cs.LG cs.AI stat.ML 62%

Deep Reinforcement Learning for Multi-Driver Vehicle Dispatching and Repositioning Problem

John Holler, Risto Vuorio, Zhiwei Qin, Xiaocheng Tang, Yan Jiao, Tiancheng Jin, Satinder Singh, Chenxi Wang, Jieping Ye

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments ICDM 2019 Short Paper

详情

展开后加载摘要…

URL PDF HTML 收藏
1811.12470 2019-11-26 cs.LG cs.AI cs.CR stat.ML 62%

Analyzing Federated Learning through an Adversarial Lens

Arjun Nitin Bhagoji, Supriyo Chakraborty, Prateek Mittal, Seraphin Calo

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Extended version of paper accepted to ICML 2019, code available at https://github.com/inspire-group/ModelPoisoning; 19 pages, 14 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.10134 2019-11-25 cs.LG cs.AI 62%

A Transfer Learning Method for Goal Recognition Exploiting Cross-Domain Spatial Features

Thibault Duhamel, Mariane Maynard, Froduald Kabanza

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.09724 2019-11-25 stat.ML cs.AI cs.LG 62%

Information-Theoretic Confidence Bounds for Reinforcement Learning

Xiuyuan Lu, Benjamin Van Roy

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.04021 2019-11-14 cs.AI cs.LG cs.NE 62%

DRiLLS: Deep Reinforcement Learning for Logic Synthesis

Abdelrahman Hosny, Soheil Hashemi, Mohamed Shalan, Sherief Reda

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments ASPDAC'2020

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.04870 2019-11-13 cs.AI cs.LG cs.MA stat.ML 62%

Network Classifiers With Output Smoothing

Elsa Rizk, Roula Nassif, Ali H. Sayed

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1807.06333 2019-11-12 cs.LG cs.AI stat.ML 62%

Foundations for Restraining Bolts: Reinforcement Learning with LTLf/LDLf restraining specifications

Giuseppe De Giacomo, Luca Iocchi, Marco Favorito, Fabio Patrizi

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Journal ref ICAPS 2019: 128-136

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.02166 2019-11-07 cs.LG cs.AI 62%

Distributional Reward Decomposition for Reinforcement Learning

Zichuan Lin, Li Zhao, Derek Yang, Tao Qin, Guangwen Yang, Tie-Yan Liu

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments NeurlPS 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.00954 2019-11-05 cs.LG cs.AI stat.ML 62%

Problem Dependent Reinforcement Learning Bounds Which Can Identify Bandit Structure in MDPs

Andrea Zanette, Emma Brunskill

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Journal ref International Conference on Machine Learning, 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.00584 2019-11-05 cs.RO cs.AI cs.LG cs.MA 62%

A Perceived Environment Design using a Multi-Modal Variational Autoencoder for learning Active-Sensing

Timo Korthals, Malte Schilling, Jürgen Leitner

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Extended Abstract for the IROS 2019 Workshop on Deep Probabilistic Generative Models for Cognitive Architecture in Robotics

详情

展开后加载摘要…

URL PDF HTML 收藏
1905.12588 2019-11-01 cs.LG cs.AI stat.ML 62%

Meta-Learning Representations for Continual Learning

Khurram Javed, Martha White

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Accepted at NeurIPS19, 15 pages, 10 figures, open-source, representation learning, continual learning, online learning

详情

展开后加载摘要…

URL PDF HTML 收藏
1910.08719 2019-10-31 cs.LG cs.AI cs.GT stat.ML 62%

Explainable AI: Deep Reinforcement Learning Agents for Residential Demand Side Cost Savings in Smart Grids

Hareesh Kumar, Priyanka Mary Mammen, Krithi Ramamritham

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1910.12911 2019-10-30 cs.LG cs.AI stat.ML 62%

Generalization in Reinforcement Learning with Selective Noise Injection and Information Bottleneck

Maximilian Igl, Kamil Ciosek, Yingzhen Li, Sebastian Tschiatschek, Cheng Zhang, Sam Devlin, Katja Hofmann

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Published at Neurips 2019

详情

展开后加载摘要…

URL PDF HTML 收藏