Reasoning in Diffusion Large Language Models is Concentrated in Dynamic Confusion Zones
机构 * Institute of Computing Technology Chinese Academy of Sciences Beijing, China(计算技术研究所中国科学院北京)
专题命中 安全训练 :alignment(abstract);分类 cs.LG
AI 大模型
大模型对齐、安全、越狱、红队、提示注入和可信评测。
机构 * Institute of Computing Technology Chinese Academy of Sciences Beijing, China(计算技术研究所中国科学院北京)
专题命中 安全训练 :alignment(abstract);分类 cs.LG
机构 * Robotics and AI Group, Department of Computer Science, Electrical and Space Engineering, Luleå University of Technology(机器人与人工智能组,计算机科学、电气与空间工程系,吕勒奥技术大学)
专题命中 安全训练 :safety(abstract)
Comments 8 pages, 6 figures, submitted to the 2026 IEEE International Conference on Robotics & Automation