arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

共收录 15801 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. Agent评测 15801 篇

2210.13417 2022-10-25 cs.AI cs.LG 62%

Avalon: A Benchmark for RL Generalization Using Procedurally Generated Worlds

Joshua Albrecht, Abraham J. Fetterman, Bryden Fogelman, Ellie Kitanidis, Bartosz Wróblewski, Nicole Seo, Michael Rosenthal, Maksis Knutins, Zachary Polizzi, James B. Simon, Kanjun Qiu

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Accepted to NeurIPS Datasets and Benchmarks 2022. Video and links to all code, data, etc can be found at https://generallyintelligent.com/avalon/

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.12511 2022-10-25 cs.AI cs.CL cs.CV cs.RO 62%

DOROTHIE: Spoken Dialogue for Handling Unexpected Situations in Interactive Autonomous Driving Agents

Ziqiao Ma, Ben VanDerPloeg, Cristian-Paul Bara, Huang Yidong, Eui-In Kim, Felix Gervits, Matthew Marge, Joyce Chai

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.CL

Comments Findings of EMNLP, 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2105.01099 2022-10-25 cs.LG cs.AI 62%

Reinforcement Learning for Ridesharing: An Extended Survey

Zhiwei Qin, Hongtu Zhu, Jieping Ye

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Journal ref Transportation Research Part C: Emerging Technologies, Volume 144, 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.15701 2022-10-24 cs.LG cs.AI 62%

Provable General Function Class Representation Learning in Multitask Bandits and MDPs

Rui Lu, Andrew Zhao, Simon S. Du, Gao Huang

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2102.08581 2022-10-21 cs.LG cs.AI 62%

Efficient Scheduling of Data Augmentation for Deep Reinforcement Learning

Byungchan Ko, Jungseul Ok

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Journal ref Neurips 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.09579 2022-10-19 cs.LG cs.AI 62%

Unpacking Reward Shaping: Understanding the Benefits of Reward Engineering on Sample Complexity

Abhishek Gupta, Aldo Pacchiano, Yuexiang Zhai, Sham M. Kakade, Sergey Levine

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.03535 2022-10-18 cs.LG cs.AI cs.MA 62%

Influencing Long-Term Behavior in Multiagent Reinforcement Learning

Dong-Ki Kim, Matthew Riemer, Miao Liu, Jakob N. Foerster, Michael Everett, Chuangchuang Sun, Gerald Tesauro, Jonathan P. How

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Accepted to NeurIPS 2022. The earlier version was presented at the Gamification and Multiagent Solutions Workshop (ICLR 2022) with a spotlight. Code at https://github.com/dkkim93/further and videos at https://sites.google.com/view/further-marl

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.06702 2022-10-14 cs.LG cs.AI 62%

A Mixture of Surprises for Unsupervised Reinforcement Learning

Andrew Zhao, Matthieu Gaetan Lin, Yangguang Li, Yong-Jin Liu, Gao Huang

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Accepted to NeurIPS 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.15367 2022-10-11 cs.LG cs.AI 62%

Non-Markovian Reward Modelling from Trajectory Labels via Interpretable Multiple Instance Learning

Joseph Early, Tom Bewley, Christine Evers, Sarvapali Ramchurn

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments 27 pages (10 main content; 2 references; 1 checklist; 14 appendix). 14 figures (9 main content; 5 appendix). Published at NeurIPS 2022. Revisions: v2) Updated to NeurIPS camera ready version (extra experiments)

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.03586 2022-10-10 cs.LG cs.AI cs.RO 62%

CausalAgents: A Robustness Benchmark for Motion Forecasting using Causal Relationships

Rebecca Roelofs, Liting Sun, Ben Caine, Khaled S. Refaat, Ben Sapp, Scott Ettinger, Wei Chai

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Rebecca Roelofs and Liting Sun are equally contributed to the work

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.01801 2022-10-06 cs.LG cs.AI 62%

Safe Reinforcement Learning From Pixels Using a Stochastic Latent Representation

Yannick Hogewind, Thiago D. Simao, Tal Kachman, Nils Jansen

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.09597 2022-10-05 cs.LG cs.AI cs.GT 62%

Feasible Adversarial Robust Reinforcement Learning for Underspecified Environments

JB Lanier, Stephen McAleer, Pierre Baldi, Roy Fox

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Added new theory sections. Added comparison to self-play. Added adversary mixed-strategy analysis

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.01007 2022-10-04 cs.LG cs.AI 62%

Reward Learning with Trees: Methods and Evaluation

Tom Bewley, Jonathan Lawry, Arthur Richards, Rachel Craddock, Ian Henderson

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments 22 pages (9 main body). Preprint, under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.14375 2022-09-30 cs.LG cs.CL 62%

Improving alignment of dialogue agents via targeted human judgements

Amelia Glaese, Nat McAleese, Maja Trębacz, John Aslanides, Vlad Firoiu, Timo Ewalds, Maribeth Rauh, Laura Weidinger, Martin Chadwick, Phoebe Thacker, Lucy Campbell-Gillingham, Jonathan Uesato, Po-Sen Huang, Ramona Comanescu, Fan Yang, Abigail See, Sumanth Dathathri, Rory Greig, Charlie Chen, Doug Fritz, Jaume Sanchez Elias, Richard Green, Soňa Mokrá, Nicholas Fernando, Boxi Wu, Rachel Foley, Susannah Young, Iason Gabriel, William Isaac, John Mellor, Demis Hassabis, Koray Kavukcuoglu, Lisa Anne Hendricks, Geoffrey Irving

专题命中 Agent评测 :agent(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.09033 2022-09-20 cs.CR cs.AI cs.LG 62%

A Transferable and Automatic Tuning of Deep Reinforcement Learning for Cost Effective Phishing Detection

Orel Lavie, Asaf Shabtai, Gilad Katz

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.06192 2022-09-14 cs.CV cs.AI cs.CL 62%

StoryDALL-E: Adapting Pretrained Text-to-Image Transformers for Story Continuation

Adyasha Maharana, Darryl Hannan, Mohit Bansal

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.CL

Comments ECCV 2022 (33 pages; code, data, demo, model card available at https://github.com/adymaharana/storydalle)

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.04665 2022-09-14 cs.AI cs.LG 62%

Ask Before You Act: Generalising to Novel Environments by Asking Questions

Ross Murphy, Sergey Mosesov, Javier Leguina Peral, Thymo ter Doest

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2002.10066 2022-08-26 cs.LG cs.AI stat.ML 62%

Causal Strategic Linear Regression

Yonadav Shavit, Benjamin Edelman, Brian Axelrod

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments 18 pages; published at ICML 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.10592 2022-08-24 cs.RO cs.AI cs.LG 62%

DIDER: Discovering Interpretable Dynamically Evolving Relations

Enna Sachdeva, Chiho Choi

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.02256 2022-08-23 cs.HC cs.AI cs.LG 62%

Use-Case-Grounded Simulations for Explanation Evaluation

Valerie Chen, Nari Johnson, Nicholay Topin, Gregory Plumb, Ameet Talwalkar

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.07597 2022-08-17 cs.CL cs.AI 62%

Manual-Guided Dialogue for Flexible Conversational Agents

Ryuichi Takanobu, Hao Zhou, Yankai Lin, Peng Li, Jie Zhou, Minlie Huang

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.13623 2022-08-16 cs.AI cs.LG cs.NE 62%

Learning Controllable 3D Level Generators

Zehua Jiang, Sam Earle, Michael Cerny Green, Julian Togelius

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments 8 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.03374 2022-08-09 cs.LG cs.AI 62%

Learning to Generalize with Object-centric Agents in the Open World Survival Game Crafter

Aleksandar Stanić, Yujin Tang, David Ha, Jürgen Schmidhuber

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.01781 2022-08-05 cs.LG cs.AI cs.SY eess.SY 62%

Digital Twin-Assisted Efficient Reinforcement Learning for Edge Task Scheduling

Xiucheng Wang, Longfei Ma, Haocheng Li, Zhisheng Yin, Tom. Luan, Nan Cheng

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.13699 2022-07-29 cs.LG cs.AI 62%

Modelling non-reinforced preferences using selective attention

Noor Sajid, Panagiotis Tigas, Zafeirios Fountas, Qinghai Guo, Alexey Zakharov, Lancelot Da Costa

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments 4 pages, 3 figures - Workshop Track: 1st Conference on Lifelong Learning Agents, 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.06091 2022-07-26 cs.RO cs.AI cs.LG 62%

Towards Autonomous Grading In The Real World

Yakov Miron, Chana Ross, Yuval Goldfracht, Chen Tessler, Dotan Di Castro

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments 7 pages, Accepted to IEEE-IROS2022

详情

展开后加载摘要…

URL PDF HTML 收藏
1809.10283 2022-07-26 cs.RO cs.AI cs.LG cs.NE 62%

Adding Neural Network Controllers to Behavior Trees without Destroying Performance Guarantees

Christopher Iliffe Sprague, Petter Ögren

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Accepted as Regular Paper to The 61th IEEE Conference on Decision and Control (CDC 2022)

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.07710 2022-07-19 cs.AI cs.LG 62%

Outcome-Guided Counterfactuals for Reinforcement Learning Agents from a Jointly Trained Generative Latent Space

Eric Yeh, Pedro Sequeira, Jesse Hostetler, Melinda Gervasio

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.07276 2022-07-18 cs.AI cs.CL cs.HC 62%

A Flexible Schema-Guided Dialogue Management Framework: From Friendly Peer to Virtual Standardized Cancer Patient

Benjamin Kane, Catherine Giugno, Lenhart Schubert, Kurtis Haut, Caleb Wohn, Ehsan Hoque

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.13274 2022-07-15 cs.LG cs.AI 62%

Evaluating Multimodal Interactive Agents

Josh Abramson, Arun Ahuja, Federico Carnevale, Petko Georgiev, Alex Goldin, Alden Hung, Jessica Landon, Timothy Lillicrap, Alistair Muldal, Blake Richards, Adam Santoro, Tamara von Glehn, Greg Wayne, Nathaniel Wong, Chen Yan

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏