arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

2025-08-06 至 2025-08-06 共收录 36 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 其他安全 8 篇

2508.03393 2025-08-06 cs.SE cs.AI 57%

Agentic AI in 6G Software Businesses: A Layered Maturity Model

Muhammad Zohaib, Muhammad Azeem Akbar, Sami Hyrynsalmi, Arif Ali Khan

专题命中 其他安全 :alignment(abstract);分类 cs.AI

Comments 6 pages, 3 figures and FIT'25 Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.03053 2025-08-06 cs.RO cs.AI 57%

SkeNa: Learning to Navigate Unseen Environments Based on Abstract Hand-Drawn Maps

Haojun Xu, Jiaqi Xiang, Wu Wei, Jinyu Chen, Linqing Zhong, Linjiang Huang, Hongyu Yang, Si Liu

专题命中 其他安全 :alignment(abstract);分类 cs.AI

Comments 9 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.03042 2025-08-06 cs.LG 57%

Urban In-Context Learning: Bridging Pretraining and Inference through Masked Diffusion for Urban Profiling

Ruixing Zhang, Bo Wang, Tongyu Zhu, Leilei Sun, Weifeng Lv

专题命中 其他安全 :alignment(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02755 2025-08-06 physics.hist-ph cs.AI 57%

Beyond the Wavefunction: Qualia Abstraction Language Mechanics and the Grammar of Awareness

Mikołaj Sienicki, Krzysztof Sienicki

机构 * Polish-Japanese Academy of Information Technology(波兰-日本信息科技学院) Chair of Theoretical Physics of Naturally Intelligent Systems(自然智能系统理论物理系)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

Comments 65 pages, 49 references, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.17696 2025-08-06 cs.AI cs.SY eess.SY 57%

Enhancing AI System Resiliency: Formulation and Guarantee for LSTM Resilience Based on Control Theory

Sota Yoshihara, Ryosuke Yamamoto, Hiroyuki Kusumoto, Masanari Shimura

机构 * Graduate School of Mathematics, Nagoya University, Aichi, Japan(名古屋大学数学研究科)

专题命中 其他安全 :safety(abstract);分类 cs.AI

Comments 9 pages, 6 figures. Appendix: 16 pages. First three listed authors have equal contributions

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.08328 2025-08-06 eess.AS 50%

Multimodal Representation Loss Between Timed Text and Audio for Regularized Speech Separation

Tsun-An Hsieh, Heeyoul Choi, Minje Kim

专题命中 其他安全 :alignment(abstract)

Journal ref Interspeech 2024

详情

展开后加载摘要…

URL PDF HTML 收藏