What Makes a Good Generated Image? Investigating Human and Multimodal LLM Image Preference Alignment
专题命中 偏好对齐 :alignment(title)
Comments 7 pages, 9 figures, 3 tables; appendix 16 pages, 9 figures, 6 tables
AI 大模型
大模型对齐、安全、越狱、红队、提示注入和可信评测。
专题命中 偏好对齐 :alignment(title)
Comments 7 pages, 9 figures, 3 tables; appendix 16 pages, 9 figures, 6 tables
机构 * NC AI
专题命中 偏好对齐 :alignment(abstract);safety(abstract);分类 cs.CL
Comments 19 pages, 1 figure, 14 tables. Technical report for VARCO-VISION-2.0, a Korean-English bilingual VLM in 14B and 1.7B variants. Key features: multi-image understanding, OCR with text localization, improved Korean capabilities
机构 * The Pennsylvania State University(宾夕法尼亚州立大学) ; The University of North Carolina at Chapel Hill(北卡罗来纳大学教堂山分校)
专题命中 偏好对齐 :safety(abstract);分类 cs.LG
专题命中 偏好对齐 :RLHF(abstract)
Comments 5 pages, 7 figures, conference
专题命中 偏好对齐 :DPO(abstract)