arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

共收录 15867 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. Agent评测 15867 篇

2207.08258 2022-07-26 cs.LG 57%

Minimum Description Length Control

Ted Moskovitz, Ta-Chu Kao, Maneesh Sahani, Matthew M. Botvinick

专题命中 Agent评测 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.09004 2022-07-26 physics.comp-ph cs.LG 57%

Learning effective stochastic differential equations from microscopic simulations: linking stochastic numerics to deep learning

Felix Dietrich, Alexei Makeev, George Kevrekidis, Nikolaos Evangelou, Tom Bertalan, Sebastian Reich, Ioannis G. Kevrekidis

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments 38 pages, includes supplemental material

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.11431 2022-07-26 cs.RO cs.AI 57%

Epersist: A Self Balancing Robot Using PID Controller And Deep Reinforcement Learning

Ghanta Sai Krishna, Dyavat Sumith, Garika Akshay

专题命中 Agent评测 :agent(abstract);分类 cs.AI

Comments 4 Pages, 6 Figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.11152 2022-07-25 q-fin.TR cs.LG 57%

Learn Continuously, Act Discretely: Hybrid Action-Space Reinforcement Learning For Optimal Execution

Feiyang Pan, Tongzhe Zhang, Ling Luo, Jia He, Shuoling Liu

专题命中 Agent评测 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.10573 2022-07-22 cs.CL 57%

AI Based Chatbot: An Approach of Utilizing On Customer Service Assistance

Rejwan Bin Sulaiman

专题命中 Agent评测 :agent(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.08735 2022-07-19 cs.LG stat.ML 57%

An Information-Theoretic Analysis of Bayesian Reinforcement Learning

Amaury Gouverneur, Borja Rodríguez-Gálvez, Tobias J. Oechtering, Mikael Skoglund

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments 10 pages: 6 of the main text, 1 of references, and 3 of appendices

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.08426 2022-07-19 cs.GT cs.LG cs.MA 57%

Fast Convergence of Optimistic Gradient Ascent in Network Zero-Sum Extensive Form Games

Georgios Piliouras, Lillian Ratliff, Ryann Sim, Stratis Skoulakis

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments To appear in SAGT 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.08379 2022-07-19 cs.AI 57%

Inspector: Pixel-Based Automated Game Testing via Exploration, Detection, and Investigation

Guoqing Liu, Mengzhang Cai, Li Zhao, Tao Qin, Adrian Brown, Jimmy Bischoff, Tie-Yan Liu

专题命中 Agent评测 :agent(abstract);分类 cs.AI

Comments Accepted as IEEE CoG2022 proceedings paper (Oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2201.11824 2022-07-19 cs.RO cs.AI 57%

Empirical Estimates on Hand Manipulation are Recoverable: A Step Towards Individualized and Explainable Robotic Support in Everyday Activities

Alexander Wich, Holger Schultheis, Michael Beetz

专题命中 Agent评测 :agent(abstract);分类 cs.AI

Journal ref https://www.ifaamas.org/Proceedings/aamas2022/pdfs/p1382.pdf

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.07570 2022-07-18 cs.LG 57%

The Nature of Temporal Difference Errors in Multi-step Distributional Reinforcement Learning

Yunhao Tang, Mark Rowland, Rémi Munos, Bernardo Ávila Pires, Will Dabney, Marc G. Bellemare

专题命中 Agent评测 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.07105 2022-07-15 stat.ML cs.LG math.OC 57%

Continuous-time Analysis for Variational Inequalities: An Overview and Desiderata

Tatjana Chavdarova, Ya-Ping Hsieh, Michael I. Jordan

专题命中 Agent评测 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.09397 2022-07-14 cs.HC cs.AI 57%

Using Psychological Characteristics of Situations for Social Situation Comprehension in Support Agents

Ilir Kola, Catholijn M. Jonker, M. Birna van Riemsdijk

专题命中 Agent评测 :agent(abstract);分类 cs.AI

Comments 21 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
1811.06736 2022-07-14 cs.GT cs.LG econ.TH stat.ML 57%

Learning Approximately Optimal Contracts

Alon Cohen, Moran Koren, Argyrios Deligkas

专题命中 Agent评测 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.04557 2022-07-12 cs.GT cs.CY cs.DC cs.LG econ.TH 57%

Mechanisms that Incentivize Data Sharing in Federated Learning

Sai Praneeth Karimireddy, Wenshuo Guo, Michael I. Jordan

专题命中 Agent评测 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.12089 2022-07-12 cs.CV cs.AI 57%

Sim-To-Real Transfer of Visual Grounding for Human-Aided Ambiguity Resolution

Georgios Tziafas, Hamidreza Kasaei

专题命中 Agent评测 :agent(abstract);分类 cs.AI

Comments Accepted CoLLAs 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2011.13465 2022-07-12 cs.LG cs.SY eess.SY 57%

Exploring grid topology reconfiguration using a simple deep reinforcement learning approach

Medha Subramanian, Jan Viebahn, Simon H. Tindemans, Benjamin Donnot, Antoine Marot

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Journal ref 2021 IEEE Madrid PowerTech, 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.09575 2022-07-08 cs.AI cs.HC cs.MA 57%

Human Engagement Providing Evaluative and Informative Advice for Interactive Reinforcement Learning

Adam Bignold, Francisco Cruz, Richard Dazeley, Peter Vamplew, Cameron Foale

专题命中 Agent评测 :agent(abstract);分类 cs.AI

Comments 23 pages, 15 figures

Journal ref Neural Computing and Applications, 1-16 (2022)

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.02641 2022-07-07 cs.GT cs.AI cs.DS econ.TH math.CO 57%

Reforming an Envy-Free Matching

Takehiro Ito, Yuni Iwamasa, Naonori Kakimura, Naoyuki Kamiyama, Yusuke Kobayashi, Yuta Nozaki, Yoshio Okamoto, Kenta Ozeki

专题命中 Agent评测 :agent(abstract);分类 cs.AI

Comments AAAI 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.02100 2022-07-06 cs.AI 57%

Generating Game Levels of Diverse Behaviour Engagement

Keyuan Zhang, Jiayu Bai, Jialin Liu

专题命中 Agent评测 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.05183 2022-07-04 cs.CL cs.SD eess.AS 57%

Building an ASR Error Robust Spoken Virtual Patient System in a Highly Class-Imbalanced Scenario Without Speech Data

Vishal Sunder, Prashant Serai, Eric Fosler-Lussier

专题命中 Agent评测 :agent(abstract);分类 cs.CL

Comments 5 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.14267 2022-06-30 cs.LG q-fin.TR 57%

Applications of Reinforcement Learning in Finance -- Trading with a Double Deep Q-Network

Frensi Zejnullahu, Maurice Moser, Joerg Osterrieder

专题命中 Agent评测 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.09592 2022-06-29 math.OC cs.LG 57%

A Reinforcement Learning Approach to the Stochastic Cutting Stock Problem

Anselmo R. Pitombeira-Neto, Arthur H. Fonseca Murta

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments 22 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2112.11701 2022-06-28 cs.AI 57%

Maximum Entropy Population-Based Training for Zero-Shot Human-AI Coordination

Rui Zhao, Jinming Song, Yufeng Yuan, Hu Haifeng, Yang Gao, Yi Wu, Zhongqian Sun, Yang Wei

专题命中 Agent评测 :agent(abstract);分类 cs.AI

Comments Accepted by NeurIPS Cooperative AI Workshop, 2021, link: https://www.cooperativeai.com/workshop/neurips-2021#Workshop-Papers. Under review at a conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2011.01445 2022-06-28 cs.LG 57%

Towards Fundamental Limits of Multi-armed Bandits with Random Walk Feedback

Tianyu Wang, Lin F. Yang, Zizhuo Wang

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments major revision

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.11791 2022-06-24 cs.LG cs.AR 57%

Open-source FPGA-ML codesign for the MLPerf Tiny Benchmark

Hendrik Borras, Giuseppe Di Guglielmo, Javier Duarte, Nicolò Ghielmetti, Ben Hawks, Scott Hauck, Shih-Chieh Hsu, Ryan Kastner, Jason Liang, Andres Meza, Jules Muhizi, Tai Nguyen, Rushil Roy, Nhan Tran, Yaman Umuroglu, Olivia Weng, Aidan Yokuda, Michaela Blott

专题命中 Agent评测 :workflow(abstract);分类 cs.LG

Comments 15 pages, 7 figures, Contribution to 3rd Workshop on Benchmarking Machine Learning Workloads on Emerging Hardware (MLBench) at 5th Conference on Machine Learning and Systems (MLSys)

详情

展开后加载摘要…

URL PDF HTML 收藏
2012.05874 2022-06-24 cs.GT cs.AI 57%

Hindsight and Sequential Rationality of Correlated Play

Dustin Morrill, Ryan D'Orazio, Reca Sarfati, Marc Lanctot, James R. Wright, Amy Greenwald, Michael Bowling

专题命中 Agent评测 :agent(abstract);分类 cs.AI

Comments Corrected technical report for the paper with the same title in the proceedings of the thirty-fifth AAAI Conference on Artificial Intelligence (AAAI-21), February 2-9, 2021, Virtual. Compared to v5, this version fixes the realized terminal history indicators in the diagram describing MacQueen's counterexample. 27 pages and 16 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.10924 2022-06-23 cs.CL cs.CR cs.NI 57%

Enhancing Networking Cipher Algorithms with Natural Language

John E. Ortega

专题命中 Agent评测 :agent(abstract);分类 cs.CL

Comments 12 pages, David C. Wyld et al. (Eds): CONEDU, CSITA, MLCL, ISPR, NATAP, ARIN - 2022 pp. 43-54, 2022. CS & IT - CSCP 2022 DOI: 10.5121/csit.2022.121013

Journal ref David C. Wyld et al. (Eds): CONEDU, CSITA, MLCL, ISPR, NATAP, ARIN - 2022, pp. 43-54, 2022. CS & IT

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.10524 2022-06-22 cs.LG cs.SY eess.SY 57%

Lyapunov Density Models: Constraining Distribution Shift in Learning-Based Control

Katie Kang, Paula Gradu, Jason Choi, Michael Janner, Claire Tomlin, Sergey Levine

专题命中 Agent评测 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.10313 2022-06-22 cs.RO cs.LG 57%

Active Inference for Robotic Manipulation

Tim Schneider, Boris Belousov, Hany Abdulsamad, Jan Peters

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments Published at "The Multi-disciplinary Conference on Reinforcement Learning and Decision Making (RLDM)" 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.09348 2022-06-22 cs.LG cs.GT math.OC 57%

Nested bandits

Matthieu Martin, Panayotis Mertikopoulos, Thibaud Rahier, Houssam Zenati

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments 35 pages, 14 figures; to appear in ICML 2022

详情

展开后加载摘要…

URL PDF HTML 收藏