CoopEval: Benchmarking Cooperation-Sustaining Mechanisms and LLM Agents in Social Dilemmas
CoopEval:社会困境中合作维持机制与LLM代理的基准测试
机构 * Carnegie Mellon University ; Foundations of Cooperative AI Lab (FOCAL) ; Jinesis Lab, University of Toronto \& Vector Institute ; ETH Z\" u rich ; Max Planck Institute for Intelligent Systems, T\" u bingen, Germany
专题命中 推理与问题求解 :LLM(title,title_cn);分类 cs.CL、cs.AI
AI总结 研究探讨了社会困境中合作维持机制与LLM代理的交互效果,发现重复游戏和声誉系统最有效,但合作效果随对手变化而下降,且在进化压力下更有效。
Comments Published paper at the International Conference on Machine Learning (ICML) 2026. 65 pages, 38 Figures, 8 Tables, 17 Listings