arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

2025-10-06 至 2025-10-06 共收录 3 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. AI治理与伦理 3 篇

2510.03004 2025-10-06 cs.LG cs.AI 62%

BrainIB++: Leveraging Graph Neural Networks and Information Bottleneck for Functional Brain Biomarkers in Schizophrenia

Tianzheng Hu, Qiang Li, Shu Liu, Vince D. Calhoun, Guido van Wingen, Shujian Yu

机构 * Vrije University Amsterdam(荷兰阿姆斯特丹自由大学) Tri-institutional Center for Translational Research in Neuroimaging(转化神经影像研究联合中心) Emory University(埃默里大学) Key Laboratory of Genetic Evolution and Animal Models(遗传进化与动物模型重点实验室) Kunming Institute of Zoology(昆明动物研究所) Chinese Academy of Sciences Kunming(中国科学院昆明分院) Department of Psychiatry, Amsterdam UMC, University of Amsterdam(阿姆斯特丹大学精神病科) Department of Physics and Technology, UiT The Arctic University of Norway(北极大学挪威理工学院物理与技术系)

专题命中 AI治理与伦理 :alignment(abstract);分类 cs.AI、cs.LG

Comments This manuscript has been accepted by Biomedical Signal Processing and Control and the code is available at https://github.com/TianzhengHU/BrainIB_coding/tree/main/BrainIB_GIB

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.02978 2025-10-06 cs.CY cs.AI cs.HC 62%

AI Generated Child Sexual Abuse Material -- What's the Harm?

Caoilte Ó Ciardha, John Buckley, Rebecca S. Portnoff

机构 * Senior Research Fellow, University of Kent, UK(肯特大学高级研究员) Digital Child Safety Expert(数字儿童安全专家) Vice President of Data Science, Thorn(数据科学副总裁,Thorn)

专题命中 AI治理与伦理 :harmlessness(abstract);分类 cs.AI、cs.CY

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.22774 2025-10-06 cs.AI cs.CY 62%

Bridging Ethical Principles and Algorithmic Methods: An Alternative Approach for Assessing Trustworthiness in AI Systems

Michael Papademas, Xenia Ziouvelou, Antonis Troumpoukis, Vangelis Karkaletsis

专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.AI、cs.CY

详情

展开后加载摘要…

URL PDF HTML 收藏