arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

2025-11-20 至 2025-11-20 共收录 8 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 其他安全 8 篇

2502.05934 2025-11-20 cs.AI cs.CC cs.GT cs.LG cs.MA 84%

Intrinsic Barriers and Practical Pathways for Human-AI Alignment: An Agreement-Based Complexity Analysis

Aran Nayebi

机构 * Aran Nayebi(独立研究者)

专题命中 其他安全 :alignment(title,abstract);safety(abstract);分类 cs.AI、cs.LG

Comments 21 pages, 1 figure, 1 table. To appear in AAAI 2026 Special Track on AI Alignment (oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.02505 2025-11-20 cs.CV cs.AI 57%

ESA: Energy-Based Shot Assembly Optimization for Automatic Video Editing

Yaosen Chen, Wei Wang, Tianheng Zheng, Xuming Wen, Han Yang, Yanru Zhang

机构 * Sobey Media Intelligence Laboratory(索贝媒体智能实验室) University of Electronic Science and Technology of China(电子科学与技术大学) SiChuan University(四川大学) Qinghai Normal University(青海师范大学)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.00275 2025-11-20 cs.CV cs.AI 57%

AdCare-VLM: Towards a Unified and Pre-aligned Latent Representation for Healthcare Video Understanding

Md Asaduzzaman Jabin, Hanqi Jiang, Yiwei Li, Patrick Kaggwa, Eugene Douglass, Juliet N. Sekandi, Tianming Liu

机构 * University of Georgia(佐治亚大学)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

Comments 39th Conference on Neural Information Processing Systems (NeurIPS 2025) Workshop: 7th International Workshop on Large Scale Holistic Video Understanding: Toward Video Foundation Models

Journal ref Neural Information Processing Systems (NeurIPS 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.15118 2025-11-20 cs.CV 50%

Unbiased Semantic Decoding with Vision Foundation Models for Few-shot Segmentation

Jin Wang, Bingfeng Zhang, Jian Pang, Weifeng Liu, Baodi Liu, Honglong Chen

机构 * School of Control Science and Engineering, China University of Petroleum (East China)(控制科学与工程学院,中国石油大学(华东))

专题命中 其他安全 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.15026 2025-11-20 eess.SP 50%

WiCo-MG: Wireless Channel Foundation Model for Multipath Generation via Synesthesia of Machines

Zengrui Han, Lu Bai, Xuesong Cai, Xiang Cheng

专题命中 其他安全 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.14910 2025-11-20 cs.RO 50%

Z-Merge: Multi-Agent Reinforcement Learning for On-Ramp Merging with Zone-Specific V2X Traffic Information

Yassine Ibork, Myounggyu Won, Lokesh Das

专题命中 其他安全 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.12079 2025-11-20 cs.CV 50%

Point Cloud Quantization through Multimodal Prompting for 3D Understanding

Hongxuan Li, Wencheng Zhu, Huiying Xu, Xinzhong Zhu, Pengfei Zhu

专题命中 其他安全 :alignment(abstract)

Comments Accepted by AAAI 2026. 11 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.17597 2025-11-20 cs.HC cs.CV 50%

Human-AI Collaboration and Explainability for 2D/3D Registration Quality Assurance

Sue Min Cho, Alexander Do, Russell H. Taylor, Mathias Unberath

机构 * Johns Hopkins University(约翰霍普金斯大学)

专题命中 其他安全 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏