arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

2026-01-26 至 2026-01-26 共收录 1 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. AI治理与伦理 1 篇

2601.16237 2026-01-26 cs.MA cs.AI cs.CY cs.SE 62%

Computational Foundations for Strategic Coopetition: Formalizing Collective Action and Loyalty

战略竞合的计算基础:形式化集体行动与忠诚

Vik Pant, Eric Yu

专题命中 AI治理与伦理 :alignment(abstract);分类 cs.AI、cs.CY

AI总结 本文提出了一种基于忠诚度调节的效用函数,用于分析竞合环境中集体行动问题的产生与解决,通过实验验证展示了忠诚度对努力差异的显著影响。

Comments 68 pages, 22 figures. Third technical report in research program; should be read with companion arXiv:2510.18802 and arXiv:2510.24909. Adapts and extends complex actor material from Pant (2021) doctoral dissertation, University of Toronto

详情

展开后加载摘要…

URL PDF HTML 收藏