Let the Bees Find the Weak Spots: A Path Planning Perspective on Multi-Turn Jailbreak Attacks against LLMs
专题命中 越狱攻击 :jailbreak(title);red teaming(abstract);分类 cs.CL
AI 大模型
大模型对齐、安全、越狱、红队、提示注入和可信评测。
专题命中 越狱攻击 :jailbreak(title);red teaming(abstract);分类 cs.CL
专题命中 越狱攻击 :prompt injection(abstract);分类 cs.AI
Comments Submitted to AAAI 2026 Workshop on Trust and Control in Agentic AI (TrustAgent)