TraceAegis: Securing LLM-Based Agents via Hierarchical and Behavioral Anomaly Detection
专题命中 其他安全 :safety(abstract)
AI 大模型
大模型对齐、安全、越狱、红队、提示注入和可信评测。
专题命中 其他安全 :safety(abstract)
专题命中 其他安全 :alignment(abstract)
Comments Published in the Proceedings of the ICML 2025 Workshop on Multi-modal Foun- dation Models and Large Language Models for Life Sciences, Vancouver, Canada. 2025
机构 * Professorship of Autonomous Vehicle Systems, TUM School of Engineering and Design, Technical University of Munich(自主车辆系统教授职位,TUM工程与设计学院,慕尼黑技术大学) ; Munich Institute of Robotics and Machine Intelligence (MIRMI)(慕尼黑机器人与机器智能研究所(MIRMI))
专题命中 其他安全 :safety(abstract)
Comments 8 pages, submitted to the IEEE ICRA 2026, Vienna, Austria
机构 * Beijing University of Posts and Telecommunications(北京邮电大学)
专题命中 其他安全 :alignment(abstract)
Comments This paper was originally submitted to ACM MM 2025 on April 12, 2025
专题命中 其他安全 :alignment(abstract)
机构 * School of Physics, Engineering and Technology, University of York(物理、工程与技术学院,约克大学) ; Department of Chemistry, University of York(化学学院,约克大学)
专题命中 其他安全 :safety(abstract)
Comments This article was originally published in the IEEE Systems, Man, and Cybernetics Society eNewsletter, September 2025 issue: https://www.ieeesmc.org/wp-content/uploads/2024/10/FeatureArticle_Sept25.pdf
Journal ref https://www.ieeesmc.org/wp-content/uploads/2024/10/FeatureArticle_Sept25.pdf
专题命中 其他安全 :alignment(abstract)
Comments 35 pages, 5 figures, 2 tables