KLAAD: Refining Attention Mechanisms to Reduce Societal Bias in Generative Language Models
Seorin Kim, Dongyoung Lee, Jaejin Lee
机构
*
Dept. of Data Science, Seoul National University(数据科学系,首尔国立大学)
;
Dept. of Computer Science and Engineering, Seoul National University(计算机科学与工程系,首尔国立大学)
Assessing the Reliability of LLMs Annotations in the Context of Demographic Bias and Model Explanation
Hadi Mohammadi, Tina Shahedi, Pablo Mosteiro, Massimo Poesio, Ayoub Bagheri, Anastasia Giachanou
机构
*
Department of Methodology and Statistics, Utrecht University, The Netherlands(方法论与统计学系,乌特列支大学,荷兰)
;
Department of Information and Computing Sciences, Utrecht University, The Netherlands(信息与计算科学系,乌特列支大学,荷兰)
;
Queen Mary University of London, London, United Kingdom(伦敦女王玛丽大学,伦敦,英国)