Mechanistic Interpretability of Socio-Political Frames in Language Models
机构 * Technische Universität Berlin, Berlin, Germany Humboldt Institute for Internet \& Society, Berlin, Germany
专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract);分类 cs.CL、cs.AI
Comments Peer-reviewed and presented at Advances in Interpretable Machine Learning and Artificial Intelligence (AIMLAI) Workshop at ECML/PKDD 2024