Decentralized End-to-End Multi-AAV Pursuit Using Predictive Spatio-Temporal Observation via Deep Reinforcement Learning
基于深度强化学习的去中心化端到端多无人艇追捕
Yude Li, Zhexuan Zhou, Huizhe Li, Yanke Sun, Yenan Wu, Yichen Lai, Yiming Wang, Youmin Gong, Jie Mei
机构
*
School of Intelligence Science and Engineering(智能科学与工程学院)
;
Guangdong Key Laboratory of Intelligent Morphing Mechanisms and Adaptive Robotics(广东省智能变形机制与自适应机器人重点实验室)
;
Shenzhen Key Lab for Advanced Motion Control and Modern Automation Equipments(深圳先进运动控制与现代自动化设备重点实验室)
;
Harbin Institute of Technology(哈尔滨工业大学)
LVLMs and Humans Ground Differently in Referential Communication
LVLMs与人类在指称交流中的基础不同
Peter Zeng, Weiling Li, Amie J. Paige, Zhengxiang Wang, Panagiotis Kaliosis, Dimitris Samaras, Gregory Zelinsky, Susan E. Brennan, Owen Rambow
机构
*
Department of Computer Science(计算机科学系)
;
Department of Psychology(心理学系)
;
Department of Linguistics(语言学系)
;
Institute for Advanced Computational Science(先进计算科学研究院)
In-Context Reinforcement Learning via Communicative World Models
通过通信世界模型进行上下文强化学习
Fernando Martinez-Lopez, Tao Li, Yingdong Lu, Juntao Chen
机构
*
Department of Computer and Information Sciences, Fordham University(福特汉姆大学计算机与信息科学系)
;
Department of Systems Engineering, City University of Hong Kong(香港城市大学系统工程系)
;
IBM Research(IBM研究院)
CommentsAccepted for publication in 2026 IEEE International Conference on Communications, Control, and Computing Technologies for Smart Grids (SmartGridComm): Workshop on Cyber-Physical Power System Resilience: Challenges and Emerging Solutions
Distributed Dynamic Associative Memory via Online Convex Optimization
通过在线凸优化实现分布式动态关联记忆
Bowen Wang, Matteo Zecchin, Osvaldo Simeone
机构
*
Department of Engineering, King’s College London(伦敦国王学院工程系)
;
Communication Systems Department, EURECOM(EURECOM通信系统部)
;
Institute for Intelligent Networked Systems (INSI) at Northeastern University London(伦敦东北大学智能网络系统研究所)