arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

共收录 15801 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. Agent评测 15801 篇

2305.00365 2023-08-03 cs.LG cs.AI cs.SY eess.SY 62%

A Transfer Learning Approach to Minimize Reinforcement Learning Risks in Energy Optimization for Smart Buildings

Mikhail Genkin, J. J. McArthur

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments 31 pages, 9 figures, submitted to the journal Energy and Buildings

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.13447 2023-07-26 cs.RO cs.AI cs.LG 62%

A behavioural transformer for effective collaboration between a robot and a non-stationary human

Ruaridh Mon-Williams, Theodoros Stouraitis, Sethu Vijayakumar

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments 8 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.02119 2023-07-26 cs.AI cs.LG 62%

Diversity Induced Environment Design via Self-Play

Dexun Li, Wenjun Li, Pradeep Varakantham

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.12926 2023-07-25 cs.LG cs.AI cs.HC 62%

Contextual Bandits and Imitation Learning via Preference-Based Active Queries

Ayush Sekhari, Karthik Sridharan, Wen Sun, Runzhe Wu

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.12158 2023-07-25 cs.LG cs.AI cs.HC 62%

DIP-RL: Demonstration-Inferred Preference Learning in Minecraft

Ellen Novoseller, Vinicius G. Goecks, David Watkins, Josh Miller, Nicholas Waytowich

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Paper accepted at The Many Facets of Preference Learning Workshop at the International Conference on Machine Learning (ICML), Honolulu, Hawaii, USA, 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.12143 2023-07-25 cs.AI cs.LG q-bio.NC 62%

Emergence of Adaptive Circadian Rhythms in Deep Reinforcement Learning

Aqeel Labash, Florian Fletzer, Daniel Majoral, Raul Vicente

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments ICML 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.14823 2023-07-25 cs.LG cs.AI 62%

Co-Imitation Learning without Expert Demonstration

Kun-Peng Ning, Hu Xu, Kun Zhu, Sheng-Jun Huang

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.11288 2023-07-24 cs.LG cs.AI stat.ML 62%

Kernelized Offline Contextual Dueling Bandits

Viraj Mehta, Ojash Neopane, Vikramjeet Das, Sen Lin, Jeff Schneider, Willie Neiswanger

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.00127 2023-07-07 cs.LG cs.AI cs.SY eess.SY 62%

Optimal Scheduling in IoT-Driven Smart Isolated Microgrids Based on Deep Reinforcement Learning

Jiaju Qi, Lei Lei, Kan Zheng, Simon X. Yang, Xuemin, Shen

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.12888 2023-07-06 cs.LG cs.AI stat.ML 62%

Meta-Learning for Simple Regret Minimization

Mohammadjavad Azizi, Branislav Kveton, Mohammad Ghavamzadeh, Sumeet Katariya

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.00923 2023-07-04 cs.LG cs.AI 62%

Achieving Stable Training of Reinforcement Learning Agents in Bimodal Environments through Batch Learning

E. Hurwitz, N. Peace, G. Cevora

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.15591 2023-06-28 cs.LG cs.NI cs.SE 62%

Learning to Sail Dynamic Networks: The MARLIN Reinforcement Learning Framework for Congestion Control in Tactical Environments

Raffaele Galliera, Mattia Zaccarini, Alessandro Morelli, Roberto Fronteddu, Filippo Poltronieri, Niranjan Suri, Mauro Tortonesi

专题命中 Agent评测 :agent(abstract);分类 cs.LG、cs.SE

Comments 6 pages, 4 figures, IEEE conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.15559 2023-06-28 cs.CR cs.AI cs.LG 62%

RansomAI: AI-powered Ransomware for Stealthy Encryption

Jan von der Assen, Alberto Huertas Celdrán, Janik Luechinger, Pedro Miguel Sánchez Sánchez, Gérôme Bovet, Gregorio Martínez Pérez, Burkhard Stiller

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.14911 2023-06-28 cs.CL cs.AI 62%

"You might think about slightly revising the title": identifying hedges in peer-tutoring interactions

Yann Raphalen, Chloé Clavel, Justine Cassell

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.CL

Comments Published in Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (ACL), 2022

Journal ref Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (ACL), Volume 1: long papers (2022)

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.14178 2023-06-27 eess.SY cs.AI cs.LG cs.SY 62%

A Framework for dynamically meeting performance objectives on a service mesh

Forough Shahab Samani, Rolf Stadler

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.09643 2023-06-19 cs.LG cs.AI stat.ME 62%

BISCUIT: Causal Representation Learning from Binary Interactions

Phillip Lippe, Sara Magliacane, Sindy Löwe, Yuki M. Asano, Taco Cohen, Efstratios Gavves

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Published in: Uncertainty in Artificial Intelligence (UAI 2023). Project page: https://phlippe.github.io/BISCUIT/

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.08958 2023-06-16 cs.CV cs.AI cs.LG 62%

Temporally-Extended Prompts Optimization for SAM in Interactive Medical Image Segmentation

Chuyun Shen, Wenhao Li, Ya Zhang, Xiangfeng Wang

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments 17 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.07402 2023-06-14 cs.CL cs.AI 62%

The economic trade-offs of large language models: A case study

Kristen Howell, Gwen Christian, Pavel Fomitchov, Gitit Kehat, Julianne Marzulla, Leanne Rolston, Jadin Tredup, Ilana Zimmerman, Ethan Selfridge, Joseph Bradley

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.CL

Comments Paper to be published at the Association for Computational Linguistics in the Industry Track 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.07730 2023-06-14 cs.LG cs.AI 62%

How to Reuse and Compose Knowledge for a Lifetime of Tasks: A Survey on Continual Learning and Functional Composition

Jorge A. Mendez, Eric Eaton

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Published in Transactions on Machine Learning Research (TMLR), June 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.02622 2023-06-13 cs.CV cs.AI cs.LG 62%

Integrating Distributed Architectures in Highly Modular RL Libraries

Albert Bou, Sebastian Dittert, Gianni De Fabritiis

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments 9 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.05747 2023-06-12 cs.AI cs.LG 62%

An End-to-End Reinforcement Learning Approach for Job-Shop Scheduling Problems Based on Constraint Programming

Pierre Tassel, Martin Gebser, Konstantin Schekotihin

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments To be published at ICAPS 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.04765 2023-06-09 cs.AI cs.CL 62%

The HCI Aspects of Public Deployment of Research Chatbots: A User Study, Design Recommendations, and Open Challenges

Morteza Behrooz, William Ngan, Joshua Lane, Giuliano Morse, Benjamin Babcock, Kurt Shuster, Mojtaba Komeili, Moya Chen, Melanie Kambadur, Y-Lan Boureau, Jason Weston

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.04595 2023-06-08 cs.LG cs.AI cs.RO 62%

Generalization Across Observation Shifts in Reinforcement Learning

Anuj Mahajan, Amy Zhang

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.02918 2023-06-08 cs.AI cs.LG 62%

A Filtering-based General Approach to Learning Rational Constraints of Epistemic Graphs

Xiao Chi

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments 18 pages, 6 figures, submitted to CLAR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.05551 2023-06-06 cs.LG cs.AI cs.RO 62%

Causal Counterfactuals for Improving the Robustness of Reinforcement Learning

Tom He, Jasmina Gajcin, Ivana Dusparic

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Accepted to ARMS-2023 (ARMS-2023: AAMAS 2023 Workshop on Autonomous Robots and Multirobot Systems)

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.01157 2023-06-05 cs.LG cs.AI 62%

Delphic Offline Reinforcement Learning under Nonidentifiable Hidden Confounding

Alizée Pace, Hugo Yèche, Bernhard Schölkopf, Gunnar Rätsch, Guy Tennenholtz

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2301.13393 2023-06-05 cs.LG cs.AI cs.IT math.IT stat.ML 62%

Probably Anytime-Safe Stochastic Combinatorial Semi-Bandits

Yunlong Hou, Vincent Y. F. Tan, Zixin Zhong

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments To be presented at ICML 2023. 57 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.03094 2023-05-30 cs.RO cs.AI cs.LG 62%

VIMA: General Robot Manipulation with Multimodal Prompts

Yunfan Jiang, Agrim Gupta, Zichen Zhang, Guanzhi Wang, Yongqiang Dou, Yanjun Chen, Li Fei-Fei, Anima Anandkumar, Yuke Zhu, Linxi Fan

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments ICML 2023 Camera-ready version. Project website: https://vimalabs.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.15455 2023-05-29 cs.LG cs.AI 62%

A Simulation Environment and Reinforcement Learning Method for Waste Reduction

Sami Jullien, Mozhdeh Ariannezhad, Paul Groth, Maarten de Rijke

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments 20 pages, 4 figures, 4 tables, 3 listings, 1 algorithm

Journal ref TMLR, May 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.13869 2023-05-26 physics.acc-ph cs.AI cs.LG cs.SY eess.SY 62%

Trend-Based SAC Beam Control Method with Zero-Shot in Superconducting Linear Accelerator

Xiaolong Chen, Xin Qi, Chunguang Su, Yuan He, Zhijun Wang, Kunxiang Sun, Chao Jin, Weilong Chen, Shuhui Liu, Xiaoying Zhao, Duanyang Jia, Man Yi

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏