arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

共收录 15867 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. Agent评测 15867 篇

2205.10802 2022-05-24 cs.LG eess.SP 57%

Inverse-Inverse Reinforcement Learning. How to Hide Strategy from an Adversarial Inverse Reinforcement Learner

Kunal Pattanayak, Vikram Krishnamurthy, Christopher Berry

专题命中 Agent评测 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.10473 2022-05-24 cs.LG 57%

De novo design of protein target specific scaffold-based Inhibitors via Reinforcement Learning

Andrew D. McNaughton, Mridula S. Bontha, Carter R. Knutson, Jenna A. Pope, Neeraj Kumar

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments Published at the MLDD workshop, ICLR 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.04745 2022-05-24 q-fin.TR cs.LG 57%

Reinforcement Learning for Systematic FX Trading

Gabriel Borrageiro, Nick Firoozye, Paolo Barucca

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Journal ref IEEE Access, vol. 10, pp. 5024-5036, 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.09729 2022-05-20 cs.AI 57%

Reinforcement Learning with Brain-Inspired Modulation can Improve Adaptation to Environmental Changes

Eric Chalmers, Artur Luczak

专题命中 Agent评测 :agent(abstract);分类 cs.AI

Comments 9 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.09289 2022-05-20 cs.LG 57%

Routing and Placement of Macros using Deep Reinforcement Learning

Mrinal Mathur

专题命中 Agent评测 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.09473 2022-05-20 eess.IV cs.CV cs.LG q-bio.NC 57%

DBSegment: Fast and robust segmentation of deep brain structures -- Evaluation of transportability across acquisition domains

Mehri Baniasadi, Mikkel V. Petersen, Jorge Goncalves, Andreas Horn, Vanja Vlasov, Frank Hertel, Andreas Husch

专题命中 Agent评测 :planning(abstract);分类 cs.LG

Comments The data used have mistakes. No one has time to correct the data and add a new version, that is why we would like to retract it. Once we have the correct version we will resubmit to arxiv

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.12494 2022-05-20 cs.LG cs.RO stat.ML 57%

SEMI: Self-supervised Exploration via Multisensory Incongruity

Jianren Wang, Ziwen Zhuang, Hang Zhao

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments Accepted at ICRA 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.08716 2022-05-19 cs.LG 57%

No More Pesky Hyperparameters: Offline Hyperparameter Tuning for RL

Han Wang, Archit Sakhadeo, Adam White, James Bell, Vincent Liu, Xutong Zhao, Puer Liu, Tadashi Kozuno, Alona Fyshe, Martha White

专题命中 Agent评测 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.07519 2022-05-17 econ.TH cs.AI 57%

Fair Shares: Feasibility, Domination and Incentives

Moshe Babaioff, Uriel Feige

专题命中 Agent评测 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.06198 2022-05-13 cs.LG cs.CV 57%

Embodied vision for learning object representations

Arthur Aubret, Céline Teulière, Jochen Triesch

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments 6 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.00784 2022-05-13 cs.LO cs.AI cs.CC 57%

On verifying expectations and observations of intelligent agents

Sourav Chakraborty, Avijeet Ghosh, Sujata Ghosh, François Schwarzentruber

专题命中 Agent评测 :agent(abstract);分类 cs.AI

Comments Accepted in IJCAI-ECAI 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.08542 2022-05-11 cs.CL 57%

Hey AI, Can You Solve Complex Tasks by Talking to Agents?

Tushar Khot, Kyle Richardson, Daniel Khashabi, Ashish Sabharwal

专题命中 Agent评测 :agent(abstract);分类 cs.CL

Comments Accepted to Findings of ACL 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.03884 2022-05-10 cs.LG cs.SY eess.SY math.OC 57%

Decentralized Stochastic Optimization with Inherent Privacy Protection

Yongqiang Wang, H. Vincent Poor

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments Accepted as a full paper to IEEE Transactions on Automatic Control

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.06257 2022-05-06 cs.LG cs.RO 57%

Maximum Entropy RL (Provably) Solves Some Robust RL Problems

Benjamin Eysenbach, Sergey Levine

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments Published at ICLR 2022. Blog post and videos: https://bair.berkeley.edu/blog/2021/03/10/maxent-robust-rl/. arXiv admin note: text overlap with arXiv:1910.01913

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.00824 2022-05-03 cs.LG 57%

Exploration in Deep Reinforcement Learning: A Survey

Pawel Ladosz, Lilian Weng, Minwoo Kim, Hyondong Oh

专题命中 Agent评测 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.00467 2022-05-03 cs.RO cs.AI 57%

Shape Change and Control of Pressure-based Soft Agents

Federico Pigozzi

专题命中 Agent评测 :agent(abstract);分类 cs.AI

Comments Accepted at ALife'22 conference as full paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.12951 2022-04-29 cs.CL cs.HC 57%

An End-to-End Dialogue Summarization System for Sales Calls

Abedelkadir Asi, Song Wang, Roy Eisenstadt, Dean Geckt, Yarin Kuper, Yi Mao, Royi Ronen

专题命中 Agent评测 :agent(abstract);分类 cs.CL

Comments To be published in NAACL 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2112.08048 2022-04-29 cs.CL 57%

Dynamic Human Evaluation for Relative Model Comparisons

Thórhildur Thorleiksdóttir, Cedric Renggli, Nora Hollenstein, Ce Zhang

专题命中 Agent评测 :agent(abstract);分类 cs.CL

Comments accepted at LREC 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.07827 2022-04-28 cs.AI 57%

Enabling risk-aware Reinforcement Learning for medical interventions through uncertainty decomposition

Paul Festor, Giulia Luise, Matthieu Komorowski, A. Aldo Faisal

专题命中 Agent评测 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.11370 2022-04-26 cs.CV cs.AI cs.RO 57%

Deep Reinforcement Learning Using a Low-Dimensional Observation Filter for Visual Complex Video Game Playing

Victor Augusto Kich, Junior Costa de Jesus, Ricardo Bedin Grando, Alisson Henrique Kolling, Gabriel Vinícius Heisler, Rodrigo da Silva Guerra

专题命中 Agent评测 :agent(abstract);分类 cs.AI

Comments Paper accepted at the SB Games conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.10509 2022-04-25 cs.CL 57%

Towards Multi-Turn Empathetic Dialogs with Positive Emotion Elicitation

Shihang Wang, Xinchao Xu, Wenquan Wu, Zheng-Yu Niu, Hua Wu, Haifeng Wang

专题命中 Agent评测 :agent(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2112.07774 2022-04-25 cs.AI cs.HC cs.MA 57%

Assessing Human Interaction in Virtual Reality With Continually Learning Prediction Agents Based on Reinforcement Learning Algorithms: A Pilot Study

Dylan J. A. Brenneis, Adam S. Parker, Michael Bradley Johanson, Andrew Butcher, Elnaz Davoodi, Leslie Acker, Matthew M. Botvinick, Joseph Modayil, Adam White, Patrick M. Pilarski

专题命中 Agent评测 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.09500 2022-04-21 eess.SY cs.LG cs.SY 57%

A Reinforcement Learning-based Volt-VAR Control Dataset and Testing Environment

Yuanqi Gao, Nanpeng Yu

专题命中 Agent评测 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2201.04805 2022-04-19 cs.LG cs.SY eess.SY math.OC 57%

Non-Stationary Representation Learning in Sequential Linear Bandits

Yuzhen Qin, Tommaso Menara, Samet Oymak, ShiNung Ching, Fabio Pasqualetti

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments 24 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.07299 2022-04-18 cs.CL 57%

Where to Go for the Holidays: Towards Mixed-Type Dialogs for Clarification of User Goals

Zeming Liu, Jun Xu, Zeyang Lei, Haifeng Wang, Zheng-Yu Niu, Hua Wu

专题命中 Agent评测 :agent(abstract);分类 cs.CL

Comments ACL2022 Main conference. First two authors contributed equally to this work

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.06835 2022-04-15 cs.RO cs.AI 57%

GloCAL: Glocalized Curriculum-Aided Learning of Multiple Tasks with Application to Robotic Grasping

Anil Kurkcu, Cihan Acar, Domenico Campolo, Keng Peng Tee

专题命中 Agent评测 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.05245 2022-04-12 cs.LG cs.IT math.IT 57%

Approximate Top-$m$ Arm Identification with Heterogeneous Reward Variances

Ruida Zhou, Chao Tian

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments AISTATS 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.04350 2022-04-12 cs.LG cs.AR cs.CR 57%

Hardware Trojan Insertion Using Reinforcement Learning

Amin Sarihi, Ahmad Patooghy, Peter Jamieson, Abdel-Hameed A. Badawy

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments This paper was accepted for publication in GLSVLSI'22

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.07968 2022-04-11 cs.CL 57%

A Few-Shot Semantic Parser for Wizard-of-Oz Dialogues with the Precise ThingTalk Representation

Giovanni Campagna, Sina J. Semnani, Ryan Kearns, Lucas Jun Koba Sato, Silei Xu, Monica S. Lam

专题命中 Agent评测 :agent(abstract);分类 cs.CL

Comments Published in Findings of ACL 2022, 9 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.03487 2022-04-08 cs.LG cs.CV cs.RO 57%

Optimizing the Long-Term Behaviour of Deep Reinforcement Learning for Pushing and Grasping

Rodrigo Chau

专题命中 Agent评测 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏