Adversarial Fair Multi-View Clustering
机构 * School of Software, Dalian University of Technology(大连理工大学软件学院)
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.LG
AI 大模型
大模型对齐、安全、越狱、红队、提示注入和可信评测。
机构 * School of Software, Dalian University of Technology(大连理工大学软件学院)
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.LG
机构 * National Institute of Informatics, Japan(日本信息机构国家研究所) ; University of Wisconsin—Madison, USA(美国威斯康星大学麦迪逊分校) ; Kyoto University, Japan(日本京都大学)
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CL
Comments Accepted to ACL 2025 Findings. 14 pages
机构 * The Hong Kong University of Science and Technology(香港科学与技术大学) ; City University of Hong Kong(香港城市大学) ; University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.LG
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CY
Comments This work has been submitted to the IEEE for possible publication
专题命中 AI治理与伦理 :safety(abstract);分类 cs.CY
Comments pre-print
专题命中 AI治理与伦理 :safety(abstract);分类 cs.AI
机构 * Department of Methodology and Statistics, Utrecht University, The Netherlands(方法论与统计学系,乌特雷赫特大学,荷兰)
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CL
机构 * Dept. of Data Science, Seoul National University(数据科学系,首尔国立大学) ; Dept. of Computer Science and Engineering, Seoul National University(计算机科学与工程系,首尔国立大学)
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CL
机构 * Department of Methodology and Statistics, Utrecht University, The Netherlands(方法论与统计学系,乌特列支大学,荷兰) ; Department of Information and Computing Sciences, Utrecht University, The Netherlands(信息与计算科学系,乌特列支大学,荷兰) ; Queen Mary University of London, London, United Kingdom(伦敦女王玛丽大学,伦敦,英国)
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CL
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.AI
Comments 42 pages. Supplementary material included at end of article
机构 * Department of Computer Science, UCL Centre for Artificial Intelligence, University College London, London, UK(计算机科学系,UCL人工智能中心,伦敦大学学院,伦敦,英国) ; University of Southampton, Southampton, UK(南安普顿大学,南安普顿,英国)
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.LG
Comments 6 pages, 2 figures
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CY
Comments Accepted to ACM Conference on Fairness, Accountability, and Transparency 2025. 15 pages, 3 figures
专题命中 AI治理与伦理 :safety(abstract);分类 cs.CY
Comments 17 pages, 17 figures
机构 * Department of Nuclear Engineering, North Carolina State University(核工程系,北卡罗来纳州立大学) ; Department of Nuclear Engineering, University of Tennessee(核工程系,田纳西大学) ; Nuclear Energy and Fuel Cycle Division, Oak Ridge National Laboratory(核能与燃料循环 division,橡树岭国家实验室)
专题命中 AI治理与伦理 :safety(abstract);分类 cs.LG
Comments Accepted for inclusion in Transactions of the American Nuclear Society for the 2025 ANS Winter Conference
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CY
Comments 44 pages
机构 * University of Illinois, Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.AI
Comments 10 pages, Recsys'25 Spotlight Oral
机构 * Columbia University(哥伦比亚大学)
专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.LG
机构 * Leibniz Institute for Resilience Research(莱比锡韧性研究所) ; University Medicine Halle (Saale) of the Martin Luther University Halle-Wittenberg (MLU)(马尔堡-哈雷大学哈雷-萨勒医学院) ; German Center for Mental Health (DZPG)(德国心理健康中心) ; School of Life Sciences, Technical University of Munich(慕尼黑技术大学生命科学学院)
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.LG
机构 * Queen’s University(皇后大学)
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.LG
Comments Accepted to ICCV 2025
机构 * Faculty of Computing - Federal University of Mato Grosso do Sul(计算机学院 - 莫扎尔河大省联邦大学)
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.AI
Comments 12 pages; 2 figures; Preprint with the original submission accepted for publication at 39th Brazilian Symposium on Software Engineering (SBES)
机构 * Charité – Universitätsmedizin Berlin, Humboldt-Universität zu Berlin, Berlin Institute of Health (BIH), Berlin, Germany(柏林查理医院、洪堡-柏林大学、柏林健康研究所(BIH)、柏林)
专题命中 AI治理与伦理 :safety(abstract);分类 cs.LG
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CY
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.AI
机构 * Università della Svizzera italiana \& University of Amsterdam ; University of Amsterdam The Netherland ; Unviersity of Amsterdam The Netherland ; University of Amsterdam
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CL
Comments Accepted by ACM SIGIR Conference on Innovative Concepts and Theories in Information Retrieval (ICTIR 2025)
机构 * Department of EECS University of Arkansas Fayetteville(电子工程与计算机科学系美国阿肯色大学弗莱维尔分校) ; Department of CS Baylor University Waco(计算机科学系贝勒大学沃斯堡)
专题命中 AI治理与伦理 :safety(abstract);分类 cs.LG
专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.LG
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CL
Comments Accepted to ACL 2025
专题命中 AI治理与伦理 :safety(abstract);分类 cs.CY
Comments Accepted to ICML 2025 position paper track
机构 * Tsinghua University, Beijing, China(清华大学) ; Beijing Zhongguancun Academy, Beijing, China(北京中关村学院) ; Shanghai Qi Zhi Institute, Shanghai, China(上海启智研究所)
专题命中 AI治理与伦理 :DPO(abstract);分类 cs.AI
Comments Published in ICML 2025
机构 * Zhengzhou University(郑州大学) ; Wuhan University(武汉大学)
专题命中 AI治理与伦理 :DPO(abstract);分类 cs.CL