Jailbreaking Large Language Models Through Content Concretization
机构 * Networked Systems Security (NSS) Group(网络系统安全组) ; KTH Royal Institute of Technology(皇家理工学院)
专题命中 指令微调 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI
Comments Accepted for presentation in the Conference on Game Theory and AI for Security (GameSec) 2025