arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

2025-09-26 至 2025-09-26 共收录 8 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 其他安全 8 篇

2509.20705 2025-09-26 cs.RO 88%

Building Information Models to Robot-Ready Site Digital Twins (BIM2RDT): An Agentic AI Safety-First Framework

Reza Akhavian, Mani Amani, Johannes Mootz, Robert Ashe, Behrad Beheshti

机构 * Department of Civil, Construction, and Environmental Engineering, San Diego State University, San Diego, CA, United States(土木、建设与环境工程系,圣地亚哥州立大学) Department of Electrical and Computer Engineering, University of California, San Diego, San Diego, CA, United States(电气与计算机工程系,加州大学圣地亚哥分校) Department of Mechanical and Aerospace Engineering, University of California, San Diego, San Diego, CA, United States(机械与航空航天工程系,加州大学圣地亚哥分校) Department of Computer Science, San Diego State University, San Diego, CA, United States(计算机科学系,圣地亚哥州立大学)

专题命中 其他安全 :safety(title,abstract);AI safety(title);alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.20751 2025-09-26 cs.CV cs.AI cs.CL 81%

Seeing Through Words, Speaking Through Pixels: Deep Representational Alignment Between Vision and Language Models

Zoe Wanying He, Sean Trott, Meenakshi Khosla

机构 * Department of Cognitive Science University of California, San Diego(认知科学系,加州大学圣地亚哥分校)

专题命中 其他安全 :alignment(title,abstract);分类 cs.CL、cs.AI

Comments Accepted at EMNLP 2025 (camera-ready)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21136 2025-09-26 cs.AI 79%

Embodied Representation Alignment with Mirror Neurons

Wentao Zhu, Zhining Zhang, Yuwei Ren, Yin Huang, Hao Xu, Yizhou Wang

机构 * Center on Frontiers of Computing Studies, School of Compter Science, Peking University(前沿计算研究中心,计算机科学学院,北京大学) Eastern Institute of Technology, Ningbo(宁波技术研究所) Qualcomm AI Research(高通人工智能研究) Inst. for Artificial Intelligence, Peking University(人工智能研究所,北京大学)

专题命中 其他安全 :alignment(title,abstract);分类 cs.AI

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.20891 2025-09-26 cs.SD 78%

AIBA: Attention-based Instrument Band Alignment for Text-to-Audio Diffusion

Junyoung Koh, Soo Yong Kim, Gyu Hyeong Choi, Yongwon Choi

机构 * Department of Artificial Intelligence, Yonsei University(人工智能系,延世大学) MAAP LAB, MODULABS(MODULABS 音频实验室) KRAFTON AI Matics Department of Media Software, Sungkyul University(媒体软件系,松谷大学)

专题命中 其他安全 :alignment(title,abstract)

Comments NeurIPS 2025 AI for Music Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21247 2025-09-26 cs.CV cs.AI 74%

Learning to Look: Cognitive Attention Alignment with Vision-Language Models

Ryan L. Yang, Dipkamal Bhusal, Nidhi Rastogi

机构 * Brown University(布朗大学) Rochester Institute of Technology(罗切斯特理工大学)

专题命中 其他安全 :alignment(title);分类 cs.AI

Comments 7 pages, neurips workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21259 2025-09-26 cs.NI cs.AI 57%

Semantic Edge-Cloud Communication for Real-Time Urban Traffic Surveillance with ViT and LLMs over Mobile Networks

Murat Arda Onsu, Poonam Lohan, Burak Kantarci, Aisha Syed, Matthew Andrews, Sean Kennedy

专题命中 其他安全 :safety(abstract);分类 cs.AI

Comments 17 pages, 12 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.20713 2025-09-26 cs.CE 50%

Difference-Guided Reasoning: A Temporal-Spatial Framework for Large Language Models

Hong Su

专题命中 其他安全 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.16080 2025-09-26 q-bio.NC cs.SD eess.AS 50%

Interpretable Embeddings of Speech Enhance and Explain Brain Encoding Performance of Audio Models

Riki Shimizu, Richard J. Antonello, Chandan Singh, Nima Mesgarani

机构 * Mortimer B. Zuckerman Mind Brain Behavior Institute(莫蒂默·B·齐克曼脑行为研究所) Columbia University(哥伦比亚大学) Microsoft Research(微软研究院)

专题命中 其他安全 :alignment(abstract)

Comments 19 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏