机构
*
School of Automation and Intelligent Manufacturing, Southern University of Science and Technology(南方科技大学自动化与智能制造学院)
;
Pengcheng Laboratory(鹏城实验室)
;
Carnegie Mellon University(卡内基梅隆大学)
Multi-Granular Attention-Driven Reinforcement Learning Framework for Web Intelligent Enhancement Systems
多粒度注意力驱动的强化学习框架用于Web智能增强系统
Navin Chhibber, Deepak Singh, Anokh Kishore, Nikita Chawla, K. Anguraj
机构
*
Infinity Tech Group
;
Gainwell Technologies
;
Capital One
;
Sunnyvale, CA, USA(美国硅谷)
;
Department of Electronics and Communication Engineering(电子与通信工程系)
;
Sona College of Technology(Sona科技学院)
;
Independent Researcher(独立研究员)
;
Salem, India(印度塞拉姆)
Recover, Discover, Plan: Learning Skills and Concepts from Robot Failures
恢复、发现、规划:从机器人失败中学习技能与概念
Bowen Li, Mayank Mishra, Y. Isabel Liu, Stone Tao, Nishanth Kumar, Alexander G. Gray, Ruwan Wickramarachchi, Jonathan Francis, Sebastian Scherer, Tom Silver
机构
*
CMU(卡内基梅隆大学)
;
Princeton(普林斯顿大学)
;
AI2(艾伦人工智能研究所)
;
MIT(麻省理工学院)
;
Centaur AI
;
Bosch Center for AI(博世人工智能中心)
GrowthHacker: Automated Off-Policy Evaluation Optimization Using Code-Modifying LLM Agents
GrowthHacker: 使用代码修改型LLM代理的自动离线策略评估优化
Jie JW Wu, Ayanda Patrick Herlihy, Ahmad Saleem Mirza, Ali Afoud, Fatemeh Fard
机构
*
Michigan Technological University, Houghton(密歇根技术大学)
;
Birmingham City University(伯明翰城市大学)
;
University of British Columbia, Kelowna(不列颠哥伦比亚大学, 肯洛纳)
CommentsAccepted at the Joint Workshop on Statistics and Knowledge Integration for Logic, Learning, Ethical Decisions, and LLMs (SKILLED-LLMs 2026), co-located with KR 2026 and FLoC 2026, Lisbon, Portugal
机构
*
National University of Defense Technology(国防科技大学)
;
Hefei University of Technology(合肥工业大学)
;
Nanjing University (Suzhou Campus)(南京大学(苏州校区))
;
Technical University of Munich(慕尼黑工业大学)
;
Beihang University(北京航空航天大学)
;
Newcastle University(纽卡斯尔大学)
Cooperative Long Rope Skipping via Multi-Agent Reinforcement Learning
基于多智能体强化学习的协作长绳跳绳
Zihao Wang, Shijie Peng, Kerui Wu, Yu Huang, Ruiqi Xue, Dong Liu, Tian Xu, Lei Yuan, Yang Yu
机构
*
National Key Laboratory of Novel Software Technology, Nanjing University(南京大学计算机软件新技术国家重点实验室)
;
School of Artificial Intelligence, Nanjing University(南京大学人工智能学院)
;
Beijing Academy of Artificial Intelligence, BAAI(北京智源人工智能研究院)
A Human-Sensitive Controller: Adapting to Human Musculoskeletal Disorder-Related Constraints via Reinforcement Learning
一种人类敏感控制器:通过强化学习适应人类肌肉骨骼疾病相关约束
Vitor Martins, Sara M. Cerqueira, Mercedes Balcells, Elazer R Edelman, Cristina P. Santos
机构
*
Fundação para a Ciência e Tecnologia(葡萄牙科学与技术基金会)
;
Centro de Microssistemas Eletromecânicos da Universidade do Minho(University of Minho微机电系统中心)
;
Massachusetts Institute of Technology(麻省理工学院)
;
Brigham and Women’s Hospital, Harvard Medical School(哈佛医学院布莱尔妇女医院)
;
GEVAB, IQS School of Engineering(GEVAB,IQS工程学院)
;
LABBELS-Associate Laboratory, University of Minho(University of Minho关联实验室)
Observer-based Adaptive Optimal Output Containment Control problem of Linear Heterogeneous Multi-agent Systems with Relative Output Measurements
基于观测器的自适应最优输出包容控制问题:线性异构多智能体系统中的相对输出测量
Majid Mazouchi, Mohammad Bagher Naghibi-Sistani, Seyed Kamal Hosseini Sani, Farzaneh Tatari, Hamidreza Modares
机构
*
Department of Electrical Engineering, Ferdowsi University of Mashhad, Mashhad, Iran(马什哈德法尔多西大学电气工程系)
;
Department of Electrical Engineering, University of Semnan, Semnan, Iran(塞姆南大学电气工程系)
;
Missouri University of Science(密苏里科技大学)