Beyond the Black Box: Integrating Lexical and Semantic Methods in Quantitative Discourse Analysis with BERTopic
机构 * University of York(约克大学)
专题命中 其他安全 :alignment(abstract);分类 cs.CL
Comments 5 pages conference paper, 4 tables
AI 大模型
大模型对齐、安全、越狱、红队、提示注入和可信评测。
机构 * University of York(约克大学)
专题命中 其他安全 :alignment(abstract);分类 cs.CL
Comments 5 pages conference paper, 4 tables
机构 * Department of Computer Science, University of Sheffield, UK(计算机科学系,谢菲尔德大学)
专题命中 其他安全 :alignment(abstract);分类 cs.CL
Comments Accepted by EMNLP 2025
专题命中 其他安全 :safety(abstract);分类 cs.LG
Comments 18 pages, 4 figures
专题命中 其他安全 :alignment(abstract);分类 cs.AI
Comments Accepted by EMNLP 2025 Main Conference
专题命中 其他安全 :alignment(abstract);分类 cs.CL
Comments Accepted by EMNLP 2025 findings
机构 * Human Inspired Technology Research Center, Università di Padova, Padova, PD 35121 IT(人类启发技术研究中心,帕多瓦大学,帕多瓦,PD 35121 IT) ; Vrije Universiteit Amsterdam, Amsterdam, Netherlands(阿姆斯特丹自由大学,阿姆斯特丹,荷兰)
专题命中 其他安全 :safety(abstract);分类 cs.LG
专题命中 其他安全 :alignment(abstract);分类 cs.CL
机构 * University of Marburg(马尔堡大学) ; McMaster University(麦马斯特大学) ; University of Mannheim(曼海姆大学)
专题命中 其他安全 :alignment(abstract);分类 cs.CL
Comments Accepted at ICNLSP 2025, Odense, Denmark
机构 * Ume University(乌梅大学)
专题命中 其他安全 :alignment(abstract);分类 cs.AI
Comments 7 pages, accepted at VALE 2025
机构 * National Key Lab for Novel Software Technology, Nanjing University(南京大学新型软件技术国家重点实验室)
专题命中 其他安全 :safety(abstract);分类 cs.AI
专题命中 其他安全 :safety(abstract);分类 cs.CL
Comments The author information is incorrect, some contributors are not included, and the submission has not been approved by all authors
机构 * University of Illinois at Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
专题命中 其他安全 :alignment(abstract);分类 cs.CL
Comments EMNLP 2025 Findings; 25 pages, 15 figures
机构 * University of New South Wales(新南威尔士大学) ; University of Tokyo(东京大学)
专题命中 其他安全 :alignment(abstract);分类 cs.CL
Comments Accepted by EMNLP 2025 Main Conference
机构 * Case Western Reserve University(凯斯西储大学)
专题命中 其他安全 :alignment(abstract);分类 cs.LG
专题命中 其他安全 :alignment(abstract);分类 cs.AI
Comments ~7 pages; ~32 figures; compiled with pdfLaTeX. Primary category: cs.CV. (Secondary: cs.AI)
机构 * Department of Electrical and Electronics Engineering, Koç University(电子与电气工程系,科克大学) ; CEMSE Division, King Abdullah University of Science and Technology(KAUST能源科学与工程分校,国王 Abdullah 科学技术大学) ; School of Electronics and Computer Science, University of Southampton(电子与计算机科学学院,南安普顿大学)
专题命中 其他安全 :alignment(abstract);分类 cs.AI
专题命中 其他安全 :alignment(abstract);分类 cs.LG
机构 * NICE Research Group & Institute for People-Centred AI, School of Computer Science & Electronic Engineering, University of Surrey, UK(NICE研究组及以人为本的人工智能研究所、计算机科学与电子工程学院、萨里大学)
专题命中 其他安全 :safety(abstract);分类 cs.CL
Comments Accepted to RANLP 2025
机构 * The University of Sydney(悉尼大学)
专题命中 其他安全 :alignment(abstract);分类 cs.AI
Comments 10pages
专题命中 其他安全 :safety(abstract);分类 cs.LG
Comments 64 pages, 7 figures, 10 tables
Journal ref Journal of Computers in Biology and Medicine Volume 196, Part A, September 2025, 110438
专题命中 其他安全 :safety(abstract);分类 cs.LG
专题命中 其他安全 :alignment(abstract);分类 cs.AI
机构 * University of Amsterdam(阿姆斯特丹大学) ; National Institute of Informatics (NII)(日本信息处理研究所) ; Tilburg University(蒂尔堡大学) ; Southeast University(东南大学) ; Vrije Universiteit Amsterdam(阿姆斯特丹自由大学) ; Utrecht University(乌得勒支大学)
专题命中 其他安全 :alignment(abstract);分类 cs.AI
专题命中 其他安全 :alignment(abstract);分类 cs.CY
Comments 25 pages, 4 pages
专题命中 其他安全 :alignment(abstract);分类 cs.LG
机构 * University of Toronto(多伦多大学)
专题命中 其他安全 :alignment(abstract);分类 cs.AI
Journal ref PACMHCI (CSCW 2025)
机构 * Hong Kong University of Science and Technology(香港理工大学) ; Cohere
专题命中 其他安全 :alignment(abstract);分类 cs.CL
专题命中 其他安全 :safety(abstract);分类 cs.LG
Comments Presented at the 2nd Reinforcement Learning Conference (RLC2025), Edmonton, Canada. To be published in the Proceedings of the Reinforcement Learning Journal 2025
专题命中 其他安全 :alignment(abstract);分类 cs.CL
机构 * Vector Institute for AI(向量人工智能研究所) ; MIT(麻省理工学院)
专题命中 其他安全 :alignment(abstract);分类 cs.CL
Comments Accepted to COLM 2025. Code and data can be found at https://github.com/ryskina/concepts-brain-llms