Are language models rational? The case of coherence norms and belief revision
机构 * Department of Philosophy University of North Carolina at Chapel Hill(哲学系北卡罗来纳大学教堂山分校) ; Department of Computer Science University of North Carolina at Chapel Hill(计算机科学系北卡罗来纳大学教堂山分校) ; Department of Computer Science University of Texas at Austin(计算机科学系德克萨斯大学奥斯汀分校)
专题命中 其他安全 :alignment(abstract);safety(abstract);AI safety(abstract);分类 cs.CL、cs.AI
Comments substantial expansions of sections 4 and 5, updated references, numerous smaller additions and clarifications