Responsible AI Adoption in the Public Sector: A Data-Centric Taxonomy of AI Adoption Challenges
专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.AI、cs.CY
AI 大模型
大模型对齐、安全、越狱、红队、提示注入和可信评测。
专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.AI、cs.CY
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CL、cs.AI
机构 * University of Maryland(马里兰大学) ; Capital One
专题命中 AI治理与伦理 :safety(abstract);分类 cs.CL、cs.LG
Comments 22 Pages
机构 * Department of Information Systems, W.P. Carey School of Business, Arizona State University, Tempe, AZ, USA(亚利桑那州立大学信息系统系,W.P. Carey商学院,Tempe分校) ; Department of Computer Science, Cornell University, Ithaca, NY, USA(康奈尔大学计算机科学系) ; Graduate School of Management, University of California Davis, Davis, CA, USA(加州大学戴维斯分校管理研究生院)
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CL、cs.AI
Comments Accepted at PNAS Nexus
Journal ref PNAS Nexus 2025
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.AI、cs.CY
机构 * Vrije University Amsterdam(荷兰阿姆斯特丹自由大学) ; Tri-institutional Center for Translational Research in Neuroimaging(转化神经影像研究联合中心) ; Emory University(埃默里大学) ; Key Laboratory of Genetic Evolution and Animal Models(遗传进化与动物模型重点实验室) ; Kunming Institute of Zoology(昆明动物研究所) ; Chinese Academy of Sciences Kunming(中国科学院昆明分院) ; Department of Psychiatry, Amsterdam UMC, University of Amsterdam(阿姆斯特丹大学精神病科) ; Department of Physics and Technology, UiT The Arctic University of Norway(北极大学挪威理工学院物理与技术系)
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.AI、cs.LG
Comments This manuscript has been accepted by Biomedical Signal Processing and Control and the code is available at https://github.com/TianzhengHU/BrainIB_coding/tree/main/BrainIB_GIB
机构 * Senior Research Fellow, University of Kent, UK(肯特大学高级研究员) ; Digital Child Safety Expert(数字儿童安全专家) ; Vice President of Data Science, Thorn(数据科学副总裁,Thorn)
专题命中 AI治理与伦理 :harmlessness(abstract);分类 cs.AI、cs.CY
专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.AI、cs.CY
机构 * Faculty of Computing - Federal University of Mato Grosso do Sul(计算机学院 - 短暂戈亚那联邦大学)
专题命中 AI治理与伦理 :prompt injection(abstract);分类 cs.AI、cs.LG
机构 * Department of Computer Science Virginia Tech(计算机科学系弗吉尼亚理工大学)
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.AI、cs.CY
Comments Accepted at 2025 ASEE Annual Conference & Exposition
机构 * The Ohio State University(俄亥俄州立大学)
专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.AI、cs.LG
Comments Add GPT 5 experiments
机构 * TU Delft(代尔夫特理工大学)
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.AI、cs.CY
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.AI、cs.CY
Comments Proceeding of The British Academy of Management Conference 2025, University of Kent, UK
专题命中 AI治理与伦理 :safety(abstract);分类 cs.AI、cs.CY
Comments 6 pages, no figures
Journal ref Nature, 644 (8075), 2025, 38-40
专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.AI、cs.CY
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.AI、cs.CY
Comments 10 pages, 5 figures. Accepted to the Workshop on Multimodal Continual Learning (MCL) at ICCV 2025. @2025 IEEE/CVF International Conference on Computer Vision Workshops (ICCVW), ICCV's 2025
机构 * University of Ioannina(伊奥安纳大学) ; Archimedes, Athena Research Center(阿基米德研究所) ; Boston University(波士顿大学)
专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.AI、cs.LG
Comments ECML PKDD 2025
机构 * Stanford University(斯坦福大学)
专题命中 AI治理与伦理 :safety(abstract);分类 cs.CY、cs.LG
Comments 34 pages, 13 figures
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.AI、cs.CY
Comments This work has been accepted for publication as a full paper at the AAAI/ACM Conference on AI, Ethics, and Society (AIES 2025)
专题命中 AI治理与伦理 :safety(abstract);分类 cs.AI、cs.CY
Comments 40 pages, 14 figures, 16 tables. To be published in Nature Scientific Reports
机构 * École Polytechnique Fédérale de Lausanne(瑞士联邦理工学院) ; Institute of Entrepreneurship and Management, HES-SO Valais-Wallis(创业与管理研究所) ; Institute of Informatics, HES-SO Valais-Wallis(信息研究所)
专题命中 AI治理与伦理 :safety(abstract);分类 cs.CL、cs.AI
机构 * Department of Physics University of Basel(物理系 巴塞尔大学)
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.AI、cs.LG
Comments 11+25 pages, 4+11 figures
机构 * KIIT Deemed University(KIIT大学) ; Indian Institute of Technology (IIT) Bhubaneswar(印度理工学院(Bhubaneswar分校))
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CL、cs.CY
Comments Accepted at ASI @ ICCV 2025
机构 * Namibia University of Science \& Technology 13 Jackson Kaujeua Windhoek Namibia 9000 ; Rhodes University Makhanda South Africa ; International University of Management Namibia ; Charles Darwin University Australia ; Namibia University of Science \& Technology ; Rhodes University ; International University of Management ; Charles Darwin University
专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.AI、cs.CY
专题命中 AI治理与伦理 :safety(abstract);分类 cs.AI、cs.LG
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.AI、cs.CY
Comments 2025 AAAI Conference on AI, Ethics, and Society
机构 * Microsoft Research AI for Science(微软研究院人工智能与科学研究中心) ; Novartis Biomedical Research(诺华生物医学研究) ; University of Cambridge(剑桥大学) ; Jagiellonian University(雅盖隆大学)
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.AI、cs.LG
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.AI、cs.CY
Comments Conference version: AIES 2025 (non-archival track), 12 pages
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.AI、cs.LG
机构 * Apple(苹果公司)
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CL、cs.AI