arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

共收录 4786 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. 记忆与上下文管理 4786 篇

1804.07881 2018-04-24 cs.CL 57%

Event Extraction with Generative Adversarial Imitation Learning

Tongtao Zhang, Heng Ji

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
1804.06512 2018-04-19 cs.CL 57%

Dialogue Learning with Human Teaching and Feedback in End-to-End Trainable Task-Oriented Dialogue Systems

Bing Liu, Gokhan Tur, Dilek Hakkani-Tur, Pararth Shah, Larry Heck

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.CL

Comments To appear in NAACL 2018 as a long paper

详情

展开后加载摘要…

URL PDF HTML 收藏
1804.02813 2018-04-10 cs.HC cs.LG cs.NE 57%

An Adaptive Learning Method of Personality Trait Based Mood in Mental State Transition Network by Recurrent Neural Network

Takumi Ichimura, Kosuke Tanabe, Toshiyuki Yamashita

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments 6 pages, 9 figures, Proc. of IEEE 7th International Workshop on Computational Intelligence and Applications (IWCIA2014)

详情

展开后加载摘要…

URL PDF HTML 收藏
1803.03916 2018-03-13 cs.LG stat.ML 57%

Deep reinforcement learning for time series: playing idealized trading games

Xiang Gao

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1709.05706 2018-02-15 cs.AI 57%

Memory Augmented Control Networks

Arbaaz Khan, Clark Zhang, Nikolay Atanasov, Konstantinos Karydis, Vijay Kumar, Daniel D. Lee

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
1709.01308 2018-02-13 cs.AI 57%

BOOK: Storing Algorithm-Invariant Episodes for Deep Reinforcement Learning

Simyung Chang, YoungJoon Yoo, Jaeseok Choi, Nojun Kwak

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments 8 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
1801.07747 2018-01-25 cs.MA cs.AI 57%

Quantified Degrees of Group Responsibility (Extended Abstract)

Vahid Yazdanpanah, Mehdi Dastani

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments Presented in the 27th Belgian-Netherlands Conference on Artificial Intelligence (BNAIC 2015), Hasselt, Belgium

详情

展开后加载摘要…

URL PDF HTML 收藏
1801.03138 2018-01-11 cs.AI 57%

Deep In-GPU Experience Replay

Ben Parr

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments Source code (uses TensorFlow): https://github.com/bparr/gpu-experience-replay

详情

展开后加载摘要…

URL PDF HTML 收藏
1712.01169 2017-12-05 cs.LG stat.ML 57%

Episodic memory for continual model learning

David G. Nagy, Gergő Orbán

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments CLDL at NIPS 2016

详情

展开后加载摘要…

URL PDF HTML 收藏
1709.01490 2017-11-02 cs.AI 57%

Active Exploration for Learning Symbolic Representations

Garrett Andersen, George Konidaris

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
1710.07328 2017-10-23 cs.GT cs.LG math.OC 57%

Online Monotone Games

Ian Gemp, Sridhar Mahadevan

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1710.02210 2017-10-09 cs.AI 57%

Exploration in Feature Space for Reinforcement Learning

Suraj Narayanan Sasikumar

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments Masters thesis. Australian National University, May 2017. 65 pp

详情

展开后加载摘要…

URL PDF HTML 收藏
1612.03653 2017-09-22 cs.AI cs.RO 57%

Learning to Drive using Inverse Reinforcement Learning and Deep Q-Networks

Sahand Sharifzadeh, Ioannis Chiotellis, Rudolph Triebel, Daniel Cremers

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments NIPS workshop on Deep Learning for Action and Interaction, 2016

详情

展开后加载摘要…

URL PDF HTML 收藏
1707.07767 2017-07-26 cs.LG cs.RO 57%

Bellman Gradient Iteration for Inverse Reinforcement Learning

Kun Li, Yanan Sui, Joel W. Burdick

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1706.08090 2017-06-27 cs.AI 57%

Count-Based Exploration in Feature Space for Reinforcement Learning

Jarryd Martin, Suraj Narayanan Sasikumar, Tom Everitt, Marcus Hutter

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments Conference: Twenty-sixth International Joint Conference on Artificial Intelligence (IJCAI-17), 8 pages, 1 figure

详情

展开后加载摘要…

URL PDF HTML 收藏
1703.01310 2017-06-15 cs.AI 57%

Count-Based Exploration with Neural Density Models

Georg Ostrovski, Marc G. Bellemare, Aaron van den Oord, Remi Munos

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
1706.04081 2017-06-14 math.OC cs.LG math.ST stat.TH 57%

Interaction-Based Distributed Learning in Cyber-Physical and Social Networks

Francesco Sasso, Angelo Coluccia, Giuseppe Notarstefano

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1705.05690 2017-06-12 cs.NI cs.LG 57%

A Long Short-Term Memory Recurrent Neural Network Framework for Network Traffic Matrix Prediction

Abdelhadi Azzouni, Guy Pujolle

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.LG

Comments Submitted for peer review. arXiv admin note: text overlap with arXiv:1402.1128 by other authors

详情

展开后加载摘要…

URL PDF HTML 收藏
1705.09724 2017-05-30 cs.CL 57%

Semi-Supervised Model Training for Unbounded Conversational Speech Recognition

Shane Walker, Morten Pedersen, Iroro Orife, Jason Flaks

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
1705.08498 2017-05-25 cs.LG 57%

Clinical Intervention Prediction and Understanding using Deep Networks

Harini Suresh, Nathan Hunt, Alistair Johnson, Leo Anthony Celi, Peter Szolovits, Marzyeh Ghassemi

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1703.02239 2017-05-17 cs.AI 57%

Functions that Emerge through End-to-End Reinforcement Learning - The Direction for Artificial General Intelligence -

Katsunari Shibata

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.AI

Comments The Multi-disciplinary Conference on Reinforcement Learning and Decision Making (RLDM) 2017, 5 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
1603.06805 2017-05-05 q-fin.TR cs.LG q-fin.CP 57%

Using real-time cluster configurations of streaming asynchronous features as online state descriptors in financial markets

Dieter Hendricks

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments 19 pages, 6 figures, 3 tables, under review at Pattern Recognition Letters

详情

展开后加载摘要…

URL PDF HTML 收藏
1611.02167 2017-03-24 cs.LG 57%

Designing Neural Network Architectures using Reinforcement Learning

Bowen Baker, Otkrist Gupta, Nikhil Naik, Ramesh Raskar

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1703.01988 2017-03-07 cs.LG stat.ML 57%

Neural Episodic Control

Alexander Pritzel, Benigno Uria, Sriram Srinivasan, Adrià Puigdomènech, Oriol Vinyals, Demis Hassabis, Daan Wierstra, Charles Blundell

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1702.00858 2017-02-06 cs.AI 57%

The Value of Inferring the Internal State of Traffic Participants for Autonomous Freeway Driving

Zachary Sunberg, Christopher Ho, Mykel Kochenderfer

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
1701.00632 2017-01-04 cs.PL cs.SE 57%

A Simulation Tool for tccp Programs

María-del-Mar Gallardo, Leticia Lavado, Laura Panizo

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.SE

Comments In Proceedings WLP'15/'16/WFLP'16, arXiv:1701.00148

Journal ref EPTCS 234, 2017, pp. 120-134

详情

展开后加载摘要…

URL PDF HTML 收藏
1611.05397 2016-11-17 cs.LG cs.NE 57%

Reinforcement Learning with Unsupervised Auxiliary Tasks

Max Jaderberg, Volodymyr Mnih, Wojciech Marian Czarnecki, Tom Schaul, Joel Z Leibo, David Silver, Koray Kavukcuoglu

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1608.07440 2016-08-29 q-bio.QM cs.AI 57%

Activity Networks with Delays An application to toxicity analysis

Franck Delaplace, Cinzia Di Giusto, Jean-Louis Giavitto, Hanna Klaudel

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
1604.05129 2016-05-31 q-bio.NC cs.AI stat.ML 57%

Memory shapes time perception and intertemporal choices

Pedro A. Ortega, Naftali Tishby

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments 24 pages, 4 figures, 2 tables. Submitted

详情

展开后加载摘要…

URL PDF HTML 收藏
1603.08789 2016-03-30 cs.AI 57%

Using Enthymemes to Fill the Gap between Logical Argumentation and Revision of Abstract Argumentation Frameworks

Jean-Guy Mailly

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏