机构
*
Dalian University of Technology(大连理工大学)
;
University of Oxford(牛津大学)
;
Sun Yat-sen University(中山大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
University of Science and Technology of China(中国科学技术大学)
;
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
;
Beihang University(北京航空航天大学)
机构
*
Tuojing Intelligence(拓境智能)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
Tsinghua University(清华大学)
;
University of Science and Technology of China(中国科学技术大学)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
Nanyang Technological University(南洋理工大学)
;
Beihang University(北京航空航天大学)
;
Carnegie Mellon University(卡内基梅隆大学)
;
The University of Hong Kong(香港大学)
机构
*
School of Computer Science and Engineering, Beihang University(北京航空航天大学计算机科学与工程学院)
;
School of Software, Beihang University(北京航空航天大学软件学院)
;
Shandong Inspur Intelligent Production Technology Co., Ltd(山东浪潮智能生产技术有限公司)
机构
*
Beijing Advanced Innovation Center for Future Blockchain and Privacy Computing, School of Artificial Intelligence, Beihang University(北京未来区块链与隐私计算先进创新中心,人工智能学院,北京航空航天大学)
;
Tsinghua University(清华大学)
AI总结
针对准则奖励中因忽略准则间依赖关系导致的虚假信用传播问题,提出概率图框架Graphical Event Aggregation for Rubric rewards (GEAR),通过建模潜在伯努利事件和软抑制传播实现依赖感知的奖励聚合,在多个基准上提升性能并减少信用泄漏。