Multi-Agent Reinforcement Learning via Agent-Specific Preference
基于智能体特定偏好的多智能体强化学习
Ni Mu, Yao Luan, Yiqin Yang, Qing-Shan Jia
机构
*
Tsinghua University(清华大学)
;
Chinese Academy of Sciences(中国科学院)
;
CFINS
;
BNRist
;
Institute for Embodied Intelligence and Robotics(嵌入式智能与机器人研究所)
;
Institute of Automation(自动化研究所)
CommentsThis article has been accepted for publication in IEEE Transactions on Automation Science and Engineering. This is the author's version, which has not been fully edited, and the content may change prior to final publication. \c{opyright} 2026 IEEE. All rights reserved, including rights for text and data mining and training of artificial intelligence and similar technologies
ChannelGuard: Safe Models Do Not Compose into Safe Multi-Agent Systems
通道卫士:安全模型无法组合成安全的多智能体系统
Elias Hossain, Md Mehedi Hasan Nipu, Fatema Tuj Johora Faria, Tasfia Nuzhat Ornee, Maleeha Sheikh
机构
*
College of Engineering and Computer Science, University of Central Florida(工程与计算机科学学院,中央佛罗里达大学)
;
Department of Computer Science and Engineering, North South University(计算机科学与工程系,北南大学)
;
Computer Science and Engineering, Ahsanullah University of Science and Technology(计算机科学与工程,阿沙努拉科学与技术大学)
;
Department of Electrical and Computer Engineering, Purdue University Fort Wayne(电气与计算机工程系,普渡大学弗拉特沃恩分校)
机构
*
East China Normal University(东华大学)
;
Hefei University of Technology(合肥工业大学)
;
China University of Petroleum(中国石油大学)
;
Guangdong university of Finance & Economics(广东财经大学)
;
Alibaba Group(阿里巴巴集团)
CommentsCritical flaw in Eq.(5)/Alg.1 (Sec 3.2): scoring fails Lipschitz continuity in multi-modal spaces, causing invalid hierarchy. Thus, latency/FLOPs in Tables 2&3 are overestimated & irreproducible. Core defect unfixable by minor update. Withdraw to avoid misleading; will revise theory & experiments
机构
*
Department of Computer Science\&Engineering, University of Minnesota-Twin Cities, Minnesota, USA
;
Department of Electrical Engineering, University of Minnesota-Twin Cities, Minnesota, USA
Comments171 pages. Formalized in Lean 4 with Mathlib: 240 theorems in the elaborated environment, 141 audited headline results, cold-compiling from a clean checkout with zero custom axioms. Source, theorem-by-theorem contract, and reproducible axiom audit: https://github.com/selfreferencing/TSE_Formal. Companion to Agentic Capital
Illusion of Alignment: Detecting Hidden Disagreement in Collaborative Dialogue
对齐的错觉:检测协同对话中的隐藏分歧
Kaiming Liu, Fuwen Luo, Ziyue Wang, Jinrui Ju, Yuxuan Liu, Xuanyu Lei, Yunghwei Lai, Peng Li, Yang Liu
机构
*
College of AI, Tsinghua University(清华大学人工智能学院)
;
Institute for AI Industry Research (AIR), Tsinghua University(清华大学人工智能产业研究院)
;
Institute for AI, Tsinghua University(清华大学人工智能研究院)