Social Catalysts, Not Moral Agents: The Illusion of Alignment in LLM Societies
社交催化剂,而非道德主体:LLM社会中的对齐幻觉
Yueqing Hu, Yixuan Jiang, Zehua Jiang, Xiao Wen, Tianhong Wang
机构
*
Institute of Neuroscience, Chinese Academy of Sciences(中国科学院神经科学研究所)
;
School of Philosophy, Anhui University(安徽大学哲学学院)
;
Department of Psychology and Behavioral Sciences, Zhejiang University(浙江大学心理学与行为科学系)
;
Mental Health Education Center, North China Electric Power University(华北电力大学心理健康教育中心)
CommentsThis submission is a new version of arXiv:2509.05882v1. with a substantially revised experimental pipeline and new metrics. In particular, collaborator agents are now instantiated independently via separate API calls, rather than generated autoregressively by a single agent. All experimental results are new. Accepted as an extended abstract at AAMAS 2026
CommentsPublished in the 36th Central European Conference on Information and Intelligent Systems(CECIIS)at: Varaždin, Croatia. September 17-19/2025. ISSN 1847-2001 (Print). ISSN 1848-2295 (Online)
机构
*
Cheriton School of Computer Science, University of Waterloo(滑铁卢大学计算机科学学院)
;
School of Computing, FSE, Macquarie University(麦觉大学计算机学院)
;
School of Computing and Information System, the University of Melbourne(墨尔本大学计算机与信息系统学院)
;
Rajax Network Technology (ele.me), China(中国Rajax网络技术(饿了么))
;
Alibaba Cloud Computing, Alibaba Group, China(中国阿里云 computing,阿里集团)
When the Domain Expert Has No Time and the LLM Developer Has No Clinical Expertise: Real-World Lessons from LLM Co-Design in a Safety-Net Hospital
Avni Kothari, Patrick Vossler, Jean Digitale, Mohammad Forouzannia, Elise Rosenberg, Michele Lee, Jennee Bryant, Melanie Molina, James Marks, Lucas Zier, Jean Feng
Revisiting Replay and Gradient Alignment for Continual Pre-Training of Large Language Models
Istabrak Abbes, Gopeshh Subbaraj, Matthew Riemer, Nizar Islah, Benjamin Therien, Tsuguchika Tabaru, Hiroaki Kingetsu, Sarath Chandar, Irina Rish
机构
*
Université de Montréal(蒙特利尔大学)
;
Mila – Quebec AI Institute(魁北克人工智能研究院)
;
Chandar Research Lab(Chandar研究实验室)
;
IBM Research(IBM研究院)
;
Fujitsu Research(富士通研究院)
;
Polytechnique Montréal(蒙特利尔理工学院)
SAE-V: Interpreting Multimodal Models for Enhanced Alignment
Hantao Lou, Changye Li, Jiaming Ji, Yaodong Yang
机构
*
Institute for AI, Peking University, Beijing, China(人工智能研究院,北京大学,北京,中国)
;
State Key Laboratory of General Artificial Intelligence, Institute for AI, Peking University, Beijing, China(通用人工智能国家重点实验室,人工智能研究院,北京大学,北京,中国)
Analyzing Fine-Grained Alignment and Enhancing Vision Understanding in Multimodal Language Models
Jiachen Jiang, Jinxin Zhou, Bo Peng, Xia Ning, Zhihui Zhu
机构
*
Department of Computer Science and Engineering, The Ohio State University(计算机科学与工程系,俄亥俄州立大学)
;
Translational Data Analytics Institute, The Ohio State University(转化数据分析研究所,俄亥俄州立大学)
;
Department of Biomedical Informatics, The Ohio State University(生物医学信息学系,俄亥俄州立大学)
A Statistical Case Against Empirical Human-AI Alignment
Julian Rodemann, Esteban Garces Arias, Christoph Luther, Christoph Jansen, Thomas Augustin
机构
*
Department of Statistics, LMU Munich(统计系,慕尼黑大学)
;
Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心)
;
Research Group Neuroinformatics, Faculty of Computer Science, University of Vienna(神经信息学研究组,维也纳大学)
;
Doctoral School Computer Science, Faculty of Computer Science, University of Vienna(计算机科学博士学院,维也纳大学)
;
School of Computing & Communications, Lancaster University Leipzig(计算与通信学院,莱比锡 Lancaster 大学)