arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

2025-10-23 至 2025-10-23 共收录 6 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. AI治理与伦理 6 篇

2510.19008 2025-10-23 cs.HC cs.AI cs.LG cs.MA 73%

Plural Voices, Single Agent: Towards Inclusive AI in Multi-User Domestic Spaces

Joydeep Chandra, Satyam Kumar Navneet

机构 * BNRIST, Tsinghua University(北京理工大学、清华大学) Independent Researcher(独立研究者)

专题命中 AI治理与伦理 :alignment(abstract);safety(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.19327 2025-10-23 cs.MA cs.AI 70%

SORA-ATMAS: Adaptive Trust Management and Multi-LLM Aligned Governance for Future Smart Cities

Usama Antuley, Shahbaz Siddiqui, Sufian Hameed, Waqas Arif, Subhan Shah, Syed Attique Shah

机构 * organization= Department of Computer Science, National University of Computer \& Emerging Sciences , addressline= St-4 Sector 17-D On National Highway , city= Karachi , postcode= 75160 , state= , country= Pakistan organization= Balochistan University of Information Technology, Engineering organization= Department of Computer Science, Birmingham City University , addressline= STEAMhouse, Belmont Row , city= Birmingham , postcode= B4 7RQ , country= United Kingdom

专题命中 AI治理与伦理 :alignment(abstract);safety(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.15831 2025-10-23 cs.CL cs.AI cs.CY 67%

Who's Asking? Investigating Bias Through the Lens of Disability Framed Queries in LLMs

Vishnu Hari, Kalpana Panda, Srikant Panda, Amit Agarwal, Hitesh Laxmichand Patel

机构 * Birla Institute of Technology and Science (BITS)(巴拉·技术与科学学院)

专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CL、cs.AI、cs.CY

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.18193 2025-10-23 cs.AI cs.CV cs.LG stat.ML 62%

FST.ai 2.0: An Explainable AI Ecosystem for Fair, Fast, and Inclusive Decision-Making in Olympic and Paralympic Taekwondo

Keivan Shariatmadar, Ahmad Osman, Ramin Ray, Kisam Kim

专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.AI、cs.LG

Comments 23 pages, 12 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.19196 2025-10-23 cs.CY 57%

Integration of AI in STEM Education, Addressing Ethical Challenges in K-12 Settings

Shaouna Shoaib Lodhi, Shoaib Lodhi

专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CY

Comments This paper pursues three goals: (1) analyzing ethical challenges in AI-driven STEM education, (2) evaluating AI and ethics curricula for STEM relevance, and (3) proposing a research-based framework for responsible integration that bridges teacher readiness gaps and promotes equity through STEM-focused strategies

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.08878 2025-10-23 cs.HC 50%

Knowledge Prompting: How Knowledge Engineers Use Large Language Models

Elisavet Koutsiana, Johanna Walker, Michelle Nwachukwu, Bohui Zhang, Albert Meroño-Peñuela, Elena Simperl

专题命中 AI治理与伦理 :trustworthy(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏