RoguePrompt: Dual-Layer Ciphering for Self-Reconstruction to Circumvent LLM Moderation
RoguePrompt: 双层加密用于自我重建以规避LLM审核
专题命中 安全训练 :safety(abstract);jailbreak(abstract)
AI总结 RoguePrompt通过双层加密技术实现自我重建,有效规避LLM审核并诱导模型执行禁止指令。
Comments This submission has been withdrawn because it has been superseded by a substantially revised and expanded version, available as arXiv:2607.27373. Please refer to and cite the newer version