MetaExplainer: A Framework to Generate Multi-Type User-Centered Explanations for AI Systems
机构 * Rensselaer Polytechnic Institute(伦斯勒理工学院) ; Amazon Science(亚马逊科学) ; IBM Research(IBM研究院)
专题命中 幻觉与事实性 :trustworthy(abstract);分类 cs.AI、cs.LG
AI 大模型
大模型对齐、安全、越狱、红队、提示注入和可信评测。
机构 * Rensselaer Polytechnic Institute(伦斯勒理工学院) ; Amazon Science(亚马逊科学) ; IBM Research(IBM研究院)
专题命中 幻觉与事实性 :trustworthy(abstract);分类 cs.AI、cs.LG
机构 * Adobe(Adobe公司) ; Oregon State University(俄勒冈州立大学)
专题命中 幻觉与事实性 :alignment(abstract);分类 cs.CL
Comments 17 pages, 8 figures, Accepted at ICCV 2025
机构 * Department of Medicine, Massachusetts General Hospital(麻省总医院内科部) ; Department of Biostatistics, Harvard T.H. Chan School of Public Health(哈佛T.H. Chan公共卫生学院生物统计学部) ; Department of Medicine, Brigham and Women’s Hospital(布里洛妇产科医院内科部)
专题命中 幻觉与事实性 :alignment(abstract);分类 cs.LG
机构 * University College London(伦敦大学学院)
专题命中 幻觉与事实性 :alignment(abstract);分类 cs.CL
Comments submitted to NLLP 2025 Workshop