Tree of Attacks: Jailbreaking Black-Box LLMs Automatically
机构 * Yale University(耶鲁大学) ; Robust Intelligence ; Google Research(谷歌研究院)
专题命中 效率与部署 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG
Comments Accepted for presentation at NeurIPS 2024. Code: https://github.com/RICommunity/TAP