When Ethics and Payoffs Diverge: LLM Agents in Morally Charged Social Dilemmas
当伦理与收益相悖时:道德两难情境下的大语言模型智能体
Steffen Backmann, David Guzman Piedrahita, Terry Jingchen Zhang, Emanuel Tewolde, Rada Mihalcea, Bernhard Schölkopf, Zhijing Jin
机构
*
ETH Zürich(苏黎世联邦理工学院)
;
University of Zurich(苏黎世大学)
;
Carnegie Mellon University(卡内基梅隆大学)
;
University of Michigan(密歇根大学)
;
Max Planck Institute for Intelligent Systems, Tübingen(图宾根人工智能研究所)
;
University of Toronto(多伦多大学)
;
Vector Institute(向量研究所)
Comments8 pages, 7 figures, 8 tables. Accepted at the 7th Annual World AIIoT Congress (AIIoT 2026). This is the author's accepted version; the version of record will appear in IEEE Xplore
Comments41 pages, 11 tables, no figures. Preprint intended for submission to EDM 2027 / LAK 2027. Includes a reproducibility package: trained ONNX Decision Transformer, generic training script, OULAD evaluation scripts, and per-arm results CSVs
Payoff scaling shapes cooperation in LLM agents across languages
收益规模塑造跨语言LLM代理的合作行为
Trung-Kiet Huynh, Dao-Sy Duy-Minh, Thanh-Bang Cao, Phong-Hao Le, Hong-Dan Nguyen, Phu-Quy Nguyen-Lam, Minh-Luan Nguyen-Vo, Hong-Phat Pham, Phu-Hoa Pham, Thien-Kim Than, Chi-Nguyen Tran, Huy Tran, Gia-Thoai Tran-Le, Alessio Buscemi, Le Hong Trang, The Anh Han
机构
*
Faculty of Information Technology, University of Science (HCMUS), Ho Chi Minh City, Vietnam(信息技术学院,科学大学(HCMUS),胡志明市,越南)
;
Faculty of Computer Science and Engineering, Ho Chi Minh City University of Technology (HCMUT), Ho Chi Minh City, Vietnam(计算机科学与工程学院,胡志明市技术大学(HCMUT),胡志明市,越南)
;
Vietnam National University – Ho Chi Minh City (VNU-HCM), Ho Chi Minh City, Vietnam(越南国家大学——胡志明市(VNU-HCM),胡志明市,越南)
;
Luxembourg Institute of Science and Technology (LIST), Luxembourg(卢森堡科学与技术研究所(LIST),卢森堡)
;
School of Computing, Engineering and Digital Technologies, Teesside University, Middlesbrough, United Kingdom(计算、工程与数字技术学院,泰赛德大学,米德尔斯布罗,英国)
SafeMCP: Proactive Power Regulation for LLM Agent Defense via Environment-Grounded Look-Ahead Reasoning
SafeMCP:基于环境接地前瞻推理的LLM智能体防御主动功率调节
Lichao Wang, Zhaoxing Ren, Tianzhuo Yang, Jiaming Ji, Chi Harold Liu, Yaodong Yang, Juntao Dai
机构
*
Beijing Institute of Technology(北京理工大学)
;
Beijing Academy of Artificial Intelligence(北京人工智能研究院)
;
Institute for Artificial Intelligence, Peking University(北京大学人工智能研究院)
A Scoping Review of LLM-as-a-Judge in Healthcare and the MedJUDGE Framework
对LLM-as-a-Judge在医疗领域的综述及MedJUDGE框架
Chenyu Li, Zohaib Akhtar, Mingu Kwak, Yuelyu Ji, Hang Zhang, Tracey Obi, Yufan Ren, Xizhi Wu, Sonish Sivarajkumar, Harold P. Lehmann, Shyam Visweswaran, Michael J. Becich, Danielle L. Mowery, Renxuan Liu, Haoyang Sun, Yanshan Wang
机构
*
Department of Biomedical Informatics, School of Medicine, University of Pittsburgh(匹兹堡大学医学院生物医学信息学系)
;
Department of Health Information Management, School of Health and Rehabilitation Sciences, University of Pittsburgh(匹兹堡大学健康与康复科学学院健康信息管理系)
;
OpenCura, Health Innovation Consortium(OpenCura健康创新联盟)
;
Northwestern University, Kellogg School of Management(西北大学凯洛格管理学院)
;
Intelligent Systems Program, School of Computing and Information, University of Pittsburgh(匹兹堡大学计算与信息学院智能系统项目)
;
Johns Hopkins University School of Medicine Biomedical Informatics and Data Science(约翰霍普金斯大学医学院生物医学信息学与数据科学)
;
Clinical and Translational Science Institute, University of Pittsburgh(匹兹堡大学临床与转化科学研究所)
;
Institute for Biomedical Informatics, University of Pennsylvania(宾夕法尼亚大学生物医学信息学研究所)
;
Data Science, School of Computing and Information, University of Pittsburgh(匹兹堡大学计算与信息学院数据科学)
The Consensus Trap: Dissecting Subjectivity and the "Ground Truth" Illusion in Data Annotation
共识陷阱:数据标注中主观性与‘真实真相’幻觉的剖析
Sheza Munir, Benjamin Mah, Krisha Kalsi, Shivani Kapania, Julian Posada, Edith Law, Ding Wang, Syed Ishtiaque Ahmed
机构
*
University of Toronto Computer Science(多伦多大学计算机科学系)
;
University of Toronto Engineering Science(多伦多大学工程科学系)
;
Carnegie Mellon University School of Computer Science(卡内基梅隆大学计算机科学学院)
;
Yale University American Studies(耶鲁大学美国研究系)
;
University of Waterloo Computer Science(滑铁卢大学计算机科学系)
;
Google Research(谷歌研究)
;
University of Toronto(多伦多大学)
;
Carnegie Mellon University(卡内基梅隆大学)
;
Yale University(耶鲁大学)
;
University of Waterloo(滑铁卢大学)
AEGIS: An Operational Infrastructure for Post-Market Governance of Adaptive Medical AI Under US and EU Regulations
AEGIS:一种用于美国和欧盟法规下适应性医疗AI市场后治理的操作基础设施
Fardin Afdideh, Mehdi Astaraki, Fernando Seoane, Farhad Abtahi
机构
*
Department of Clinical Science, Intervention and Technology, Karolinska Institutet(临床科学、干预与技术部门,Karolinska研究院)
;
Department of Medical Radiation Physics, Stockholm University(医学辐射物理学部门,斯德哥尔摩大学)
;
Department of Oncology-Pathology, Karolinska Institutet(肿瘤学-病理学部门,Karolinska研究院)
;
Department of Clinical Physiology, Karolinska University Hospital(临床生理学部门,Karolinska大学医院)
;
Department of Textile Technology, University of Bor s(纺织技术部门,Bor s大学)
;
Department of Medical Technologies, Karolinska University Hospital(医学技术部门,Karolinska大学医院)
;
Department of Biomedical Engineering and Health System, KTH Royal Institute of Technology(生物医学工程与健康系统部门,KTH皇家理工学院)
Credibility Governance: A Social Mechanism for Collective Self-Correction under Weak Truth Signals
可信治理:在弱真相信号下的一种社会机制,用于集体自我校正
Wanying He, Yanxi Lin, Ziheng Zhou, Xue Feng, Min Peng, Qianqian Xie, Zilong Zheng, Yipeng Kang
机构
*
School of Artificial Intelligence, Wuhan University(武汉大学人工智能学院)
;
Tsinghua University(清华大学)
;
University of California, Los Angeles(加州大学洛杉矶分校)
;
State Key Laboratory of General Artificial Intelligence, BIGAI(通用人工智能国家重点实验室,BIGAI)
Buy versus Build an LLM: A Decision Framework for Governments
买还是建一个大语言模型:政府的决策框架
Jiahao Lu, Ziwei Xu, William Tjhi, Junnan Li, Antoine Bosselut, Pang Wei Koh, Mohan Kankanhalli
机构
*
National University of Singapore(新加坡国立大学)
;
AI Singapore(AI新加坡)
;
Salesforce AI Research(Salesforce AI研究)
;
EPFL(苏黎世联邦理工学院)
;
University of Washington(华盛顿大学)
;
Allen Institute for AI(人工智能研究院)