HealthBranches: Synthesizing Clinically-Grounded Question Answering Datasets via Decision Pathways
专题命中 安全评测 :trustworthy(abstract);分类 cs.CL、cs.AI、cs.LG
AI 大模型
大模型对齐、安全、越狱、红队、提示注入和可信评测。
专题命中 安全评测 :trustworthy(abstract);分类 cs.CL、cs.AI、cs.LG
专题命中 安全评测 :alignment(abstract);分类 cs.AI、cs.LG
专题命中 安全评测 :trustworthy(abstract);分类 cs.AI、cs.LG
机构 * Department of Information Systems, University of Maryland, Baltimore County(信息系统系,马里兰大学巴尔的摩分校) ; Department of Mechanical and Aerospace Engineering, The George Washington University(机械与航空航天工程系,乔治华盛顿大学) ; Department of Mechanical Engineering, Baylor University(机械工程系,贝勒大学)
专题命中 安全评测 :safety(abstract);分类 cs.AI
Comments 9 pages, 4 figures, IJCAI-2025 (accepted)
机构 * School of Cyber Science and Technology, Shenzhen Campus of Sun Yat-sen University(中山大学信息科学与技术学院(深圳校区)) ; Institute of Data Science, National University of Singapore(新加坡国立大学数据科学研究所) ; College of Modern Engineering and the Engineering Research Center of Cyberspace, Yunnan University(云南大学现代工程学院及空天信息工程研究中心)
专题命中 安全评测 :alignment(abstract);分类 cs.AI
Comments Accepted by ACM MM 2025
机构 * Department of Electrical and Computer Engineering, Duke University(电子工程与计算机科学系,杜克大学) ; Department of Biostatistics and Bioinformatics, Duke University(生物统计学与生物信息学系,杜克大学) ; Departments of Biostatistics and Bioinformatics, Radiology, Electrical and Computer Engineering, and Computer Science, Duke University(生物统计学与生物信息学系、放射学、电子工程与计算机科学系,杜克大学)
专题命中 安全评测 :alignment(abstract);分类 cs.AI
Comments 3 figures, 9 pages
机构 * Texas A&M University(德克萨斯A&M大学)
专题命中 安全评测 :safety(abstract);分类 cs.AI
Comments 19 pages, 5 figures, Preprint under review. Code available at: https://github.com/taco-group/DRAMA-X
机构 * Sungkyunkwan University, S. Korea(顺天大学) ; University of Queensland, Australia(昆士兰大学)
专题命中 安全评测 :trustworthy(abstract)
Comments 11 pages, 3 tables, 5 figures, accepted for publicaiton in the 33rd ACM International Conference on Multimedia (MM '25), October 27-31, 2025, Dublin, Ireland
专题命中 安全评测 :alignment(abstract)
专题命中 安全评测 :safety(abstract)
Comments 31 pages, 5 figures, under review
专题命中 安全评测 :alignment(abstract)
机构 * Massachusetts Institute of Technology(麻省理工学院)
专题命中 安全评测 :alignment(abstract)
Comments 10 pages, accepted to MICCAI ASMUS 25
专题命中 AI治理与伦理 :alignment(abstract);safety(abstract);分类 cs.AI、cs.CY
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CL、cs.AI、cs.CY
机构 * Chalmers University of Technology, Sweden Carl von Ossietzky Universität Oldenburg, Germany Universit\'e Grenoble Alpes, France University of Warwick, United Kingdom CSX-AI, France SRI International, United States
专题命中 AI治理与伦理 :prompt injection(abstract);分类 cs.AI、cs.CY、cs.LG
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.AI、cs.LG
机构 * Apple(苹果公司)
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CL、cs.AI
机构 * Munich Center for Mathematical Philosophy, LMU Munich Munich Germany ; GATE, CNRS, Universit\'e Jean Monnet, Universit\'e Lumiere Lyon 2 Saint-Etienne France ; Munich Center for Mathematical Philosophy, LMU Munich ; GATE, CNRS, Universit\'e Jean Monnet, Universit\'e Lumiere Lyon 2
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.AI
Comments 44 pages, 21 figures, 14 tables. Updated and published version
Journal ref Journal of Artificial Intelligence Research 83, Article 25 (August 2025)
机构 * Mashang Consumer Finance Co., Ltd.(Mashang消费金融有限公司) ; The University of Sydney(悉尼大学) ; Macau University of Science and Technology(澳门科学技术大学)
专题命中 其他安全 :alignment(title,abstract);分类 cs.CL、cs.AI
Comments To be presented at ACPR 2025 Conference
专题命中 其他安全 :alignment(title,abstract);分类 cs.AI、cs.CY
机构 * University of Science and Technology of China(中国科学技术大学) ; State Key Laboratory of Cognitive Intelligence(认知智能国家重点实验室)
专题命中 其他安全 :alignment(title,abstract);分类 cs.AI
Comments Accepted to ICCV 2025
机构 * Mohamed bin Zayed University of AI(穆罕默德·本·扎耶德人工智能大学)
专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI、cs.LG
Comments 9 pages, 4 figures
机构 * Department of Artificial Intelligence and Human Health, Icahn School of Medicine at Mount Sinai(人工智能与人类健康系,伊坎医学院 Mount Sinai 分校) ; Department of Psychiatry, Icahn School of Medicine at Mount Sinai(精神病学系,伊坎医学院 Mount Sinai 分校) ; Department of Neuroscience, Icahn School of Medicine at Mount Sinai(神经科学系,伊坎医学院 Mount Sinai 分校) ; Berkman Klein Center for Internet & Society, Harvard University(互联网与社会研究中心,哈佛大学) ; IBM Research, T.J. Watson Research Center(IBM 研究,T.J. Watson 研究中心) ; Mental Illness Research, Education and Clinical Center, James J. Peters VA Medical Center(精神疾病研究、教育与临床中心,James J. Peters VA 医疗中心)
专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI、cs.LG
Comments Translational Psychiatry, in press. This work extends our research series in computational psychiatry (e.g auto annotation in arXiv:2204.05522, topic extraction in arXiv:2204.10189, and diagnosis in arXiv:2210.15603) with the introduction of LLMs to complete the full cycle of interpreting and understanding psychotherapy strategies as a comprehensive analytical framework
Journal ref Transl Psychiatry 15, 166 (2025)
机构 * Institute for People-Centred AI and Centre for Translation Studies, School of Computer Science and Electronic Engineering, University of Surrey(以人为本的人工智能研究所和翻译研究中心,计算机科学与电子工程学院,萨里大学)
专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI
Comments Accepted to COLM 2025 Conference
机构 * School of Computer Science, Faculty of Engineering, University of Sydney(悉尼大学计算机科学学院、工程学院) ; Computational Health Informatics Program, Boston Children’s Hospital(波士顿儿童医院计算健康信息学项目) ; Harvard-MIT Center for Regulatory Science and Department of Pediatrics, Harvard Medical School(哈佛-麻省理工监管科学中心和哈佛医学院儿科部门) ; Sydney School of Public Health, Faculty of Medicine and Health, University of Sydney(悉尼大学公共卫生学院、医学与健康学院)
专题命中 其他安全 :safety(abstract);分类 cs.CL、cs.AI
Comments 12 pages, 4 figures. Updated to include Table 2, Supplementary Table 1, and an additional baseline random forest model
机构 * Walmart Global Tech(沃尔玛全球科技)
专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI
机构 * Charité – Universitätsmedizin Berlin, Department of Psychiatry and Psychotherapy, Berlin, Germany(柏林查理医院医学大学精神病与心理治疗系) ; Hertie Institute for AI in Brain Health, University of Tübingen, Germany(图宾根大学健康人工智能研究所)
专题命中 其他安全 :alignment(abstract);分类 cs.AI、cs.LG
Comments Accepted to ICML 2025
机构 * Purdue University(普渡大学) ; Johns Hopkins University(约翰霍普金斯大学)
专题命中 其他安全 :safety(abstract);分类 cs.CL、cs.AI
Comments COLM 2025. The first two authors contributed equally to this work
机构 * Independent Researcher in AI and Statistics(人工智能与统计学独立研究者) ; Shahrood University of Technology(沙霍罗德大学) ; University of Pittsburgh(匹兹堡大学) ; Duquesne University(杜克森大学)
专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI
Comments 6pages
机构 * Center for Human-Compatible AI, University of California, Berkeley(人类兼容人工智能中心,加州大学伯克利分校) ; Foundations of Cooperative AI Lab, Carnegie Mellon University(协作人工智能实验室,卡内基梅隆大学)
专题命中 其他安全 :alignment(abstract);分类 cs.AI、cs.LG