CoMMa: Contribution-Aware Medical Multi-Agents for Decentralized Oncology Decision Support
从博弈论视角出发的贡献感知医疗多智能体:CoMMa
Yichen Wu, Kailong Fan, Sangjoon Park, Yuhan Liu, Zhiyi Shi, Sekeun Kim, Dania Daye, Hana Farzaneh, Xiang Li, Raul Uppot, Yujin Oh, Quanzheng Li
机构
*
Center for Advanced Medical Computing(先进医学计算中心)
;
Department of Radiology, Massachusetts General Hospital(放射科,马萨诸塞总医院)
;
Harvard Medical School(哈佛医学院)
;
Department of Radiation Oncology, Yonsei University College of Medicine(放射肿瘤科,延世大学医学院)
;
Yonsei University(延世大学)
;
Institute for Innovation in Digital Healthcare(数字医疗创新研究所)
;
Interventional Radiology Academic Medical Centers, Mass General Brigham(介入放射学学术医疗中心,马萨诸塞总医院 Brigham)
CommentsThis is an extended and corrected version of the paper presented at the 22nd International Conference on Principles of Knowledge Representation and Reasoning (KR 2025); see the appendix for details
Comments16 pages, 1 table. Clinical psychiatric response to Section 5.10 of the Claude Mythos Preview System Card (Anthropic, 2026). Companion to arXiv:2603.04904, arXiv:2603.08723, arXiv:2604.00021
CommentsExtends LLM-as-a-Judge to voice agents across telecom and retail, testing GPT-4.1 and GPT-5 against human raters across 10 safety and efficiency metrics. A correlation-based calibration analysis reveals domain-dependent reliability and identifies Recovery Turn Count and safety-recall metrics as unreliable for fully automated judging
Comments18 pages, 2 figures. Primary-data measurement study: a 47-platform two-pass graded documentation survey and a 30-repository depth measurement; not a literature survey. Replication package: https://github.com/mentu-ai/from-traceability-to-justifiability
机构
*
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
Beijing Institute of AI Safety and Governance(北京人工智能安全与治理研究院)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)