LLM4Sweat: A Trustworthy Large Language Model for Hyperhidrosis Support
专题命中 安全评测 :trustworthy(title,abstract);分类 cs.CL、cs.AI
AI 大模型
大模型对齐、安全、越狱、红队、提示注入和可信评测。
专题命中 安全评测 :trustworthy(title,abstract);分类 cs.CL、cs.AI
机构 * Laboratory of Brain Atlas and Brain-inspired Intelligence, Institute of Automation Chinese Academy of Sciences (CASIA)(中国科学院自动化研究所脑图谱与类脑智能实验室) ; School of Artificial Intelligence, University of Chinese Academy of Sciences (UCAS)(中国科学院大学人工智能学院) ; School of Systems Science, Beijing Normal University(北京师范大学系统科学学院) ; School of Psychological and Cognitive Sciences & Beijing Key Laboratory of Behavior and Mental Health, Peking University(北京大学心理与认知科学学院) ; IDG/McGovern Institute for Brain Research, Peking University(北京大学IDG/ McGovern脑科学研究院) ; Institute for Artificial Intelligence & Key Laboratory of Machine Perception (Ministry of Education), Peking University(北京大学人工智能研究所) ; School of Future Technology, University of Chinese Academy of Sciences (UCAS)(中国科学院大学未来技术学院)
专题命中 安全评测 :alignment(title,abstract);分类 cs.CL、cs.AI
机构 * Engineering Department, School of Science and Technology (SST), City University of London(伦敦城市大学科学与技术学院工程系)
专题命中 安全评测 :safety(title,abstract);分类 cs.AI
机构 * Shanghai AI Lab(上海人工智能实验室) ; East China Normal University(东华大学)
专题命中 安全评测 :safety(title,abstract);分类 cs.CL
Comments Code and dataset are available at https://github.com/yangyangyang127/SafetyFlow
机构 * Department of Computer Science and Technology, College of AI, Institute for AI, Tsinghua-Bosch Joint ML Center, THBI Lab, BNRist Center, Tsinghua University(计算机科学与技术系、人工智能学院、人工智能研究所、清华-博世联合机器学习中心、THBI实验室、BNRist中心、清华大学) ; Institute of Artificial Intelligence, Beihang University(人工智能研究院、北航) ; RealAI
专题命中 安全评测 :alignment(abstract);safety(abstract);分类 cs.CL、cs.AI
Comments For Appendix, please refer to arXiv:2406.07057
专题命中 安全评测 :alignment(abstract);safety(abstract);分类 cs.LG
机构 * Ningbo Key Laboratory of Spatial Intelligence and Digital Derivative, Institute of Digital Twin, EIT(宁波空间智能与数字衍生关键实验室,数字孪生研究院,EIT) ; Logic Intelligence Technology(逻辑智能技术) ; BUPT(北京邮电大学) ; Xiamen University(厦门大学)
专题命中 安全评测 :alignment(abstract);分类 cs.CL、cs.AI、cs.LG
机构 * Department of Computing Science University of Alberta(计算科学系阿尔伯塔大学) ; ServiceNow Research(ServiceNow研究)
专题命中 安全评测 :alignment(abstract);分类 cs.CL、cs.AI、cs.LG
机构 * School of Electrical and Computer Engineering, University of Sydney(悉尼大学电气与计算机工程学院) ; School of Computer Science, University of Adelaide(阿德莱德大学计算机科学学院) ; School of Computing and Information Technology, University of Wollongong(沃林根大学计算与信息科技学院)
专题命中 安全评测 :alignment(abstract);分类 cs.CL、cs.AI
机构 * University of California at Berkeley(加州大学伯克利分校) ; Indian Institute of Technology Bombay(印度班加罗尔理工学院) ; Chalmers University of Technology and University of Gothenburg(查尔姆斯理工大学和哥德堡大学)
专题命中 安全评测 :trustworthy(abstract);分类 cs.AI、cs.LG
Comments This work has been accepted at ATVA'25
机构 * Otto von Guericke University Magdeburg(奥托·冯·格里克大学马格德堡)
专题命中 安全评测 :alignment(abstract);分类 cs.AI、cs.CY
专题命中 安全评测 :safety(abstract);分类 cs.AI、cs.CY
Comments Accepted to AAAI/ACM AIES 2025
专题命中 安全评测 :alignment(abstract);分类 cs.LG
Comments Published at COLM 2025
机构 * Dept. of Computer Engineering(计算机工程系) ; Jamia Millia Islamia(Jamia Millia Islamia大学)
专题命中 安全评测 :trustworthy(abstract);分类 cs.AI
机构 * Department of Computer Science(计算机科学系) ; Czech Technical University Prague(捷克技术大学布拉格) ; Global Priorities Institute(全球优先研究所) ; University of Oxford(牛津大学) ; Foundations of Cooperative AI Lab(合作人工智能基础实验室) ; Carnegie Mellon University(卡内基梅隆大学)
专题命中 安全评测 :safety(abstract);分类 cs.AI
机构 * Qorvex Consulting ; Kleiner Perkins ; Wentworth Institute of Higher Education
专题命中 安全评测 :trustworthy(abstract);分类 cs.AI
专题命中 安全评测 :alignment(abstract)
专题命中 安全评测 :safety(abstract)
机构 * Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) ; Mashang Consumer Finance Co, Ltd(马商消费金融有限公司)
专题命中 安全评测 :alignment(abstract)
Comments ACMMM2025
专题命中 安全评测 :alignment(abstract)