arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

共收录 15867 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. Agent评测 15867 篇

2202.02886 2022-06-22 cs.AI 57%

Leveraging Approximate Symbolic Models for Reinforcement Learning via Skill Diversity

Lin Guan, Sarath Sreedharan, Subbarao Kambhampati

专题命中 Agent评测 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.08723 2022-06-20 cs.CL 57%

CookDial: A dataset for task-oriented dialogs grounded in procedural documents

Yiwei Jiang, Klim Zaporojets, Johannes Deleu, Thomas Demeester, Chris Develder

专题命中 Agent评测 :agent(abstract);分类 cs.CL

Comments The dataset and codes are available at https://github.com/YiweiJiang2015/CookDial

Journal ref Applied Intelligence, 1-19 (2022)

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.11474 2022-06-20 stat.ML cs.LG 57%

Residual Bootstrap Exploration for Stochastic Linear Bandit

Shuang Wu, Chi-Hua Wang, Yuantong Li, Guang Cheng

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments Accepted by UAI 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.05378 2022-06-20 cs.LG 57%

Feature and Parameter Selection in Stochastic Linear Bandits

Ahmadreza Moradipari, Berkay Turan, Yasin Abbasi-Yadkori, Mahnoosh Alizadeh, Mohammad Ghavamzadeh

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Journal ref International Conference on Machine Learning, 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2201.12018 2022-06-17 cs.LG cs.CR cs.DS 57%

Transfer Learning In Differential Privacy's Hybrid-Model

Refael Kohen, Or Sheffet

专题命中 Agent评测 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.07680 2022-06-16 cs.LG physics.geo-ph 57%

Learning Large-scale Subsurface Simulations with a Hybrid Graph Network Simulator

Tailin Wu, Qinchen Wang, Yinan Zhang, Rex Ying, Kaidi Cao, Rok Sosič, Ridwan Jalali, Hassan Hamam, Marko Maucec, Jure Leskovec

专题命中 Agent评测 :planning(abstract);分类 cs.LG

Comments SIGKDD 2022; 11 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.04574 2022-06-16 cs.LG quant-ph stat.ML 57%

Reinforcement-Learning-Based Variational Quantum Circuits Optimization for Combinatorial Problems

Sami Khairy, Ruslan Shaydulin, Lukasz Cincio, Yuri Alexeev, Prasanna Balaprakash

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Journal ref Proceedings of the Machine Learning and the Physical Sciences workshop at Conference on Neural Information Processing Systems (NeurIPS 2019)

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.14798 2022-06-15 cs.GT cs.AI cs.MA econ.TH 57%

Random Rank: The One and Only Strategyproof and Proportionally Fair Randomized Facility Location Mechanism

Haris Aziz, Alexander Lam, Mashbat Suzuki, Toby Walsh

专题命中 Agent评测 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.07498 2022-06-15 quant-ph cs.LG 57%

Short Quantum Circuits in Reinforcement Learning Policies for the Vehicle Routing Problem

Fabio Sanches, Sean Weinberg, Takanori Ide, Kazumitsu Kamiya

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments 15 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.06096 2022-06-14 q-bio.NC cs.AI 57%

An Enactivist-Inspired Mathematical Model of Cognition

Vadim Weinstein, Basak Sakcak, Steven M. LaValle

专题命中 Agent评测 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2201.11921 2022-06-14 cs.LG 57%

Adaptive Best-of-Both-Worlds Algorithm for Heavy-Tailed Multi-Armed Bandits

Jiatai Huang, Yan Dai, Longbo Huang

专题命中 Agent评测 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.05967 2022-06-14 cs.CV cs.LG 57%

GoToNet: Fast Monocular Scene Exposure and Exploration

Tom Avrech, Evgenii Zheltonozhskii, Chaim Baskin, Ehud Rivlin

专题命中 Agent评测 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.05835 2022-06-14 q-fin.PM cs.LG cs.MA econ.GN q-fin.EC 57%

Deep Reinforcement Learning for Optimal Investment and Saving Strategy Selection in Heterogeneous Profiles: Intelligent Agents working towards retirement

Fatih Ozhamaratli, Paolo Barucca

专题命中 Agent评测 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.05373 2022-06-14 math.GT cs.LG 57%

An application of neural networks to a problem in knot theory and group theory (untangling braids)

Alexei Lisitsa, Mateo Salles, Alexei Vernitski

专题命中 Agent评测 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.05355 2022-06-14 cs.AI cs.HC 57%

Social Practices for Social Driven Conversations in Serious Games

Agnese Augello, Manuel Gentile, Frank Dignum

专题命中 Agent评测 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.04751 2022-06-13 cs.CL 57%

Defending Compositionality in Emergent Languages

Michal Auersperger, Pavel Pecina

专题命中 Agent评测 :agent(abstract);分类 cs.CL

Comments Accepted to NAACL SRW 22

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.10895 2022-06-10 cs.LG math.ST stat.ML stat.TH 57%

Contextual Information-Directed Sampling

Botao Hao, Tor Lattimore, Chao Qin

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments Accepted at ICML 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.04091 2022-06-10 stat.ML cs.LG 57%

Uplifting Bandits

Yu-Guan Hsieh, Shiva Prasad Kasiviswanathan, Branislav Kveton

专题命中 Agent评测 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2107.05762 2022-06-10 cs.LG 57%

Strategic Instrumental Variable Regression: Recovering Causal Relationships From Strategic Responses

Keegan Harris, Daniel Ngo, Logan Stapleton, Hoda Heidari, Zhiwei Steven Wu

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments In the 39th International Conference on Machine Learning (ICML 2022)

详情

展开后加载摘要…

URL PDF HTML 收藏
1812.06003 2022-06-08 cs.LG stat.ML 57%

Nonparametric inference of interaction laws in systems of agents from trajectory data

Fei Lu, Mauro Maggioni, Sui Tang, Ming Zhong

专题命中 Agent评测 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.03021 2022-06-08 cs.CL 57%

Plot Writing From Pre-Trained Language Models

Yiping Jin, Vishakha Kadam, Dittaya Wanvarie

专题命中 Agent评测 :planning(abstract);分类 cs.CL

Comments Accepted to INLG 2022; 16 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.01896 2022-06-07 cs.LG 57%

Adaptive Tree Backup Algorithms for Temporal-Difference Reinforcement Learning

Brett Daley, Isaac Chan

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments RLDM 2022. 4 pages, 1 figure

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.13252 2022-06-07 cs.AI 57%

The Quest for a Common Model of the Intelligent Decision Maker

Richard S. Sutton

专题命中 Agent评测 :agent(abstract);分类 cs.AI

Comments Will appear as an extended abstract at the fifth Multi-disciplinary Conference on Reinforcement Learning and Decision Making, held in Providence, Rhode Island, June 8-11, 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.01587 2022-06-06 cs.HC cs.AI cs.RO 57%

Employing Socially Interactive Agents for Robotic Neurorehabilitation Training

Rhythm Arora, Matteo Lavit Nicora, Pooja Prajod, Daniele Panzeri, Elisabeth André, Patrick Gebhard, Matteo Malosio

专题命中 Agent评测 :agent(abstract);分类 cs.AI

Comments The 5th Workshop on Behavior Adaptation Interaction and Learning for Assistive Robotics (BAILAR)

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.00304 2022-06-02 cs.RO cs.AI cs.HC 57%

Perception-Intention-Action Cycle in Human-Robot Collaborative Tasks

J. E. Dominguez-Vidal, Nicolas Rodriguez, Rene Alquezar, Alberto Sanfeliu

专题命中 Agent评测 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.07217 2022-06-02 cs.LG 57%

Online Nonsubmodular Minimization with Delayed Costs: From Full Information to Bandit Feedback

Tianyi Lin, Aldo Pacchiano, Yaodong Yu, Michael I. Jordan

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments Accepted by ICML 2022; The first three authors contributed equally to this work; 36 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.13079 2022-05-27 cs.LG 57%

Learning to Query Internet Text for Informing Reinforcement Learning Agents

Kolby Nottingham, Alekhya Pyla, Sameer Singh, Roy Fox

专题命中 Agent评测 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2104.08570 2022-05-27 cs.CL 57%

Crossing the Conversational Chasm: A Primer on Natural Language Processing for Multilingual Task-Oriented Dialogue Systems

Evgeniia Razumovskaia, Goran Glavaš, Olga Majewska, Edoardo M. Ponti, Anna Korhonen, Ivan Vulić

专题命中 Agent评测 :agent(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.12240 2022-05-25 cs.CL 57%

VIRATrustData: A Trust-Annotated Corpus of Human-Chatbot Conversations About COVID-19 Vaccines

Roni Friedman, João Sedoc, Shai Gretz, Assaf Toledo, Rose Weeks, Naor Bar-Zeev, Yoav Katz, Noam Slonim

专题命中 Agent评测 :agent(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.05061 2022-05-25 cs.LG 57%

On the Verge of Solving Rocket League using Deep Reinforcement Learning and Sim-to-sim Transfer

Marco Pleines, Konstantin Ramthun, Yannik Wegener, Hendrik Meyer, Matthias Pallasch, Sebastian Prior, Jannik Drögemüller, Leon Büttinghaus, Thilo Röthemeyer, Alexander Kaschwig, Oliver Chmurzynski, Frederik Rohkrähmer, Roman Kalkreuth, Frank Zimmer, Mike Preuss

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments Accepted at IEEE Conference on Games 2022, 8 pages, 6 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏