The Attacker Moves Second: Stronger Adaptive Attacks Bypass Defenses Against Llm Jailbreaks and Prompt Injections
机构 * OpenAI ; Anthropic ; Google DeepMind(谷歌DeepMind) ; HackAPrompt ; Northeastern University(东北大学) ; ETH Zürich(苏黎世联邦理工学院) ; AI Sequrity Company(AI安全公司) ; MATS Main contributors(MATS主要贡献者)
专题命中 越狱攻击 :prompt injection(title,abstract);分类 cs.LG