arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

共收录 15866 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. Agent评测 15866 篇

2304.03321 2023-04-10 cs.LG cs.SY eess.SY 57%

Adaptive Decision-Making with Constraints and Dependent Losses: Performance Guarantees and Applications to Online and Nonlinear Identification

Michael Muehlebach

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments 8 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
1812.11276 2023-04-07 cs.LG stat.ML 57%

Learn to Interpret Atari Agents

Zhao Yang, Song Bai, Li Zhang, Philip H. S. Torr

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments An old report. Uploaded for archival purposes only

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.00364 2023-04-04 q-fin.CP cs.AI 57%

Mastering Pair Trading with Risk-Aware Recurrent Reinforcement Learning

Weiguang Han, Jimin Huang, Qianqian Xie, Boyi Zhang, Yanzhao Lai, Min Peng

专题命中 Agent评测 :agent(abstract);分类 cs.AI

Comments 8 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.14362 2023-04-04 cs.LG 57%

Federated Learning Using Variance Reduced Stochastic Gradient for Probabilistically Activated Agents

M. R. Rostami, S. S. Kia

专题命中 Agent评测 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.17617 2023-04-03 cs.LG 57%

An evaluation of time series forecasting models on water consumption data: A case study of Greece

Ioannis Kontopoulos, Antonios Makris, Konstantinos Tserpes, Theodora Varvarigou

专题命中 Agent评测 :planning(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.07154 2023-03-28 stat.ML cs.LG 57%

Risk-aware linear bandits with convex loss

Patrick Saux, Odalric-Ambrym Maillard

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Journal ref International Conference on Artificial Intelligence and Statistics (AISTATS), Apr 2023, Valencia, Spain

详情

展开后加载摘要…

URL PDF HTML 收藏
2112.15287 2023-03-24 math.OC cs.DC cs.LG cs.MA 57%

Distributed Random Reshuffling over Networks

Kun Huang, Xiao Li, Andre Milzarek, Shi Pu, Junwen Qiu

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments 20 pages, 13 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2011.13203 2023-03-22 cs.AI 57%

Everyone Knows that Everyone Knows: Gossip Protocols for Super Experts

Hans van Ditmarsch, Malvin Gattinger, Rahim Ramezanian

专题命中 Agent评测 :agent(abstract);分类 cs.AI

Journal ref Studia Logica 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.06601 2023-03-21 cs.LG cs.RO 57%

Causal Confusion and Reward Misidentification in Preference-Based Reward Learning

Jeremy Tien, Jerry Zhi-Yang He, Zackory Erickson, Anca D. Dragan, Daniel S. Brown

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments In the proceedings of the Eleventh International Conference on Learning Representations (ICLR 2023). https://iclr.cc/virtual/2023/poster/10822

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.09946 2023-03-20 eess.SY cs.LG cs.MA cs.RO cs.SY 57%

An Adaptive Fuzzy Reinforcement Learning Cooperative Approach for the Autonomous Control of Flock Systems

Shuzheng Qu, Mohammed Abouheaf, Wail Gueaieb, Davide Spinello

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments 7 pages, 2 figures

Journal ref IEEE International Conference on Robotics and Automation (ICRA) 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.02181 2023-03-17 cs.LG stat.AP 57%

DISCOVER: Deep identification of symbolically concise open-form PDEs via enhanced reinforcement-learning

Mengge Du, Yuntian Chen, Dongxiao Zhang

专题命中 Agent评测 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.06484 2023-03-16 cs.RO cs.LG cs.SY eess.SY 57%

Robust High-speed Running for Quadruped Robots via Deep Reinforcement Learning

Guillaume Bellegarda, Yiyu Chen, Zhuochen Liu, Quan Nguyen

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Journal ref 2022 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS), Kyoto, Japan, 2022, pp. 10364-10370

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.10691 2023-03-15 cs.RO cs.LG cs.SY eess.SY 57%

Provably Safe Reinforcement Learning via Action Projection using Reachability Analysis and Polynomial Zonotopes

Niklas Kochdumper, Hanna Krasowski, Xiao Wang, Stanley Bak, Matthias Althoff

专题命中 Agent评测 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.05733 2023-03-13 cs.LG 57%

Provably Efficient Model-Free Algorithms for Non-stationary CMDPs

Honghao Wei, Arnob Ghosh, Ness Shroff, Lei Ying, Xingyu Zhou

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Journal ref AISTATS 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.03768 2023-03-13 cs.AI cs.MA 57%

Catch Me If You Can: Improving Adversaries in Cyber-Security With Q-Learning Algorithms

Arti Bandhana, Ondřej Lukáš, Sebastian Garcia, Tomáš Kroupa

专题命中 Agent评测 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.04911 2023-03-10 eess.IV cs.CV cs.LG 57%

Reverse Engineering Breast MRIs: Predicting Acquisition Parameters Directly from Images

Nicholas Konz, Maciej A. Mazurowski

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments Paper accepted at MIDL 2023. Code available at https://github.com/mazurowski-lab/MRI-IAP-prediction

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.00081 2023-03-09 cs.LG cs.CR 57%

Sampling Attacks on Meta Reinforcement Learning: A Minimax Formulation and Complexity Analysis

Tao Li, Haozhe Lei, Quanyan Zhu

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments updates: github repo posted

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.02735 2023-03-08 stat.ML cs.LG cs.NA math.NA math.ST stat.TH 57%

Learning particle swarming models from data with Gaussian processes

Jinchao Feng, Charles Kulick, Yunxiang Ren, Sui Tang

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments 44 pages; Appendix 5 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.01275 2023-03-07 cs.LG 57%

ReLOAD: Reinforcement Learning with Optimistic Ascent-Descent for Last-Iterate Convergence in Constrained MDPs

Ted Moskovitz, Brendan O'Donoghue, Vivek Veeriah, Sebastian Flennerhag, Satinder Singh, Tom Zahavy

专题命中 Agent评测 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.09322 2023-03-06 cs.LG 57%

Entropy Augmented Reinforcement Learning

Jianfei Ma

专题命中 Agent评测 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.00672 2023-03-02 cs.AI 57%

Forward-PECVaR Algorithm: Exact Evaluation for CVaR SSPs

Willy Arthur Silva Reis, Denis Benevolo Pais, Valdinei Freire, Karina Valdivia Delgado

专题命中 Agent评测 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.01379 2023-03-02 cs.RO cs.AI 57%

Extraneousness-Aware Imitation Learning

Ray Chen Zheng, Kaizhe Hu, Zhecheng Yuan, Boyuan Chen, Huazhe Xu

专题命中 Agent评测 :agent(abstract);分类 cs.AI

Comments 7 pages, 6 figures; accepted to ICRA 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2012.08950 2023-03-02 cs.CV cs.LG 57%

Revocable Deep Reinforcement Learning with Affinity Regularization for Outlier-Robust Graph Matching

Chang Liu, Zetian Jiang, Runzhong Wang, Junchi Yan, Lingxiao Huang, Pinyan Lu

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments Proceedings of The Eleventh International Conference on Learning Representations (ICLR 2023)

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.14176 2023-03-01 cs.AI cs.CE math.OC 57%

Reinforcement Learning with Depreciating Assets

Taylor Dohmen, Ashutosh Trivedi

专题命中 Agent评测 :agent(abstract);分类 cs.AI

Comments Full version of extended abstract appearing in the proceedings of AAMAS 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.04179 2023-02-28 cs.LG 57%

A Scale-Independent Multi-Objective Reinforcement Learning with Convergence Analysis

Mohsen Amidzadeh

专题命中 Agent评测 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2112.00818 2023-02-28 cs.CY cs.DC cs.GT cs.LG 57%

Models of fairness in federated learning

Kate Donahue, Jon Kleinberg

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments Preliminary version accepted for oral presentation at Neurips workshop on Learning in Presence of Strategic Behavior (2021) and EAAMO (2022). Final version accepted at The Web Conference (2023)

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.12520 2023-02-27 cs.LG cs.SY eess.SY 57%

A Novel Demand Response Model and Method for Peak Reduction in Smart Grids -- PowerTAC

Sanjay Chandlekar, Arthik Boroju, Shweta Jain, Sujit Gujar

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments 11 pages, 5 figures, 2 tables, Accepted as an Extended Abstract in AAMAS'23

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.12320 2023-02-27 math.OC cs.LG cs.SY eess.SY 57%

Dynamic Regret Analysis of Safe Distributed Online Optimization for Convex and Non-convex Problems

Ting-Jui Chang, Sapana Chaudhary, Dileep Kalathil, Shahin Shahrampour

专题命中 Agent评测 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.01935 2023-02-23 cs.CL cs.HC 57%

CAB: Empathetic Dialogue Generation with Cognition, Affection and Behavior

Pan Gao, Donghong Han, Rui Zhou, Xuejiao Zhang, Zikun Wang

专题命中 Agent评测 :agent(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.05455 2023-02-21 cs.RO cs.LG 57%

Benchmark for Models Predicting Human Behavior in Gap Acceptance Scenarios

Julian Frederik Schumann, Jens Kober, Arkady Zgonnikov

专题命中 Agent评测 :planning(abstract);分类 cs.LG

Comments 11 pages, 5 figures, accepted by in IEEE Transactions on Intelligent Vehicles, 2023

详情

展开后加载摘要…

URL PDF HTML 收藏