Machine Learning Systems in the IoT: Trustworthiness Trade-offs for Edge Intelligence
专题命中 安全评测 :trustworthy(abstract);分类 cs.CY、cs.LG
Comments In Proceedings of the Second International Conference on Cognitive Machine Intelligence (CogMI 2020)
AI 大模型
大模型对齐、安全、越狱、红队、提示注入和可信评测。
专题命中 安全评测 :trustworthy(abstract);分类 cs.CY、cs.LG
Comments In Proceedings of the Second International Conference on Cognitive Machine Intelligence (CogMI 2020)
专题命中 安全评测 :alignment(abstract);分类 cs.CL、cs.AI
专题命中 安全评测 :trustworthy(abstract);分类 cs.CL、cs.LG
专题命中 安全评测 :trustworthy(abstract);分类 cs.AI、cs.LG
Comments 8 pages, 5 figures, accepted at IJCNN 2021
专题命中 安全评测 :trustworthy(abstract);分类 cs.AI、cs.LG
专题命中 安全评测 :alignment(abstract);分类 cs.CL、cs.AI
Comments NAACL 2021 (16 pages)
专题命中 安全评测 :trustworthy(abstract);分类 cs.CY、cs.LG
Comments This is the submitted, pre-peer reviewed version of this paper
专题命中 安全评测 :safety(abstract);分类 cs.AI、cs.LG
专题命中 安全评测 :safety(abstract);分类 cs.AI、cs.LG
专题命中 安全评测 :trustworthy(abstract);分类 cs.AI、cs.LG
Comments 27 pages, 2 figures, submitted to "Special Issue on: Explainable AI (XAI) for Web-based Information Processing"
Journal ref Journal of Artificial Intelligence, Volume 294, May 2021, 103457
专题命中 安全评测 :trustworthy(abstract);分类 cs.AI、cs.CY
Comments Accepted to ACM FAccT 2021
专题命中 安全评测 :trustworthy(abstract);分类 cs.AI、cs.LG
Comments Accepted in AAAI2021
专题命中 安全评测 :trustworthy(abstract);分类 cs.AI、cs.LG
Comments 7 pages, 5 figures
专题命中 安全评测 :safety(abstract);分类 cs.AI、cs.LG
专题命中 安全评测 :alignment(abstract);分类 cs.CL、cs.LG
Comments Accepted for publication in Main Technical Track at AAAI 20
专题命中 安全评测 :alignment(abstract);分类 cs.CL、cs.LG
Journal ref NeurIPS 2020
专题命中 安全评测 :alignment(abstract);分类 cs.CL、cs.AI
专题命中 安全评测 :safety(abstract);分类 cs.CY、cs.LG
Comments All the authors contributed equally. This paper is an outcome of https://www.cs.ubc.ca/~kevinlb/teaching/cs532l%20-%202018-19/index.html. To be submitted to a journal in transportation or urban planning
专题命中 安全评测 :trustworthy(abstract);分类 cs.AI、cs.LG
Journal ref Proceedings of the 35th International Conference on Machine Learning, PMLR 80:4518-4527, 2018
专题命中 安全评测 :safety(abstract);分类 cs.AI、cs.LG
Comments 20 pages, 9 figures, submited to IEEE Transactions on Knowledge and Data Engineering
专题命中 安全评测 :trustworthy(abstract);分类 cs.AI、cs.LG
Comments 24 pages, 20 figures, survey paper, submitting to IEEE
专题命中 安全评测 :trustworthy(abstract);分类 cs.AI、cs.LG
Comments 17 pages, preprint
专题命中 安全评测 :safety(abstract);分类 cs.AI、cs.LG
Comments Published at Conference on Robot Learning (CoRL) 2019; (v2) contains minor updates to related works; (v3) acknowledged AWS
Journal ref Proceedings of Machine Learning Research (PMLR) Vol. 100, 2019
专题命中 安全评测 :safety(abstract);分类 cs.AI、cs.LG
Comments Conference on Fairness, Accountability, and Transparency (FAT* '20), January 27-30, 2020, Barcelona, Spain
专题命中 安全评测 :safety(abstract);分类 cs.AI、cs.LG
专题命中 安全评测 :safety(abstract);分类 cs.AI、cs.LG
专题命中 安全评测 :safety(abstract);分类 cs.CY、cs.LG
专题命中 安全评测 :trustworthy(abstract);分类 cs.AI、cs.LG
专题命中 安全评测 :trustworthy(abstract);分类 cs.AI、cs.LG
Comments 20th International Conference on Intelligent Data Engineering and Automated Learning (IDEAL), 14--16 November 2019
专题命中 安全评测 :trustworthy(abstract);分类 cs.AI、cs.CY
Comments A Computing Community Consortium (CCC) workshop report, 109 pages