Pattern Enhanced Multi-Turn Jailbreaking: Exploiting Structural Vulnerabilities in Large Language Models
模式增强的多轮劫持:在大语言模型中利用结构漏洞
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI
AI总结 本文提出PE-CoA框架,通过五个对话模式构建多轮劫持攻击,揭示了大语言模型在不同危害类别下的漏洞及防御局限性。