Competency Assessment for Autonomous Agents using Deep Generative Models
专题命中 安全评测 :trustworthy(abstract);分类 cs.AI、cs.LG
AI 大模型
大模型对齐、安全、越狱、红队、提示注入和可信评测。
专题命中 安全评测 :trustworthy(abstract);分类 cs.AI、cs.LG
专题命中 安全评测 :safety(abstract);分类 cs.CL、cs.LG
专题命中 安全评测 :trustworthy(abstract);分类 cs.AI、cs.CY
专题命中 安全评测 :trustworthy(abstract);分类 cs.AI、cs.LG
专题命中 安全评测 :trustworthy(abstract);分类 cs.AI、cs.CY
Comments We discovered that the manuscript unintentionally contains passages that are direct quotes from previous literature, but fails to properly address them as such
专题命中 安全评测 :trustworthy(abstract);分类 cs.AI、cs.LG
专题命中 安全评测 :trustworthy(abstract);分类 cs.CL、cs.AI
Comments Undergraduate Research Thesis, The Ohio State University
专题命中 安全评测 :trustworthy(abstract);分类 cs.AI、cs.LG
专题命中 安全评测 :trustworthy(abstract);分类 cs.AI、cs.LG
Comments Paper #203
Journal ref AAAI-MLPS-2021: Association for the Advancement of Artificial Intelligence (AAAI) 2021 Spring Symposium on Combining Artificial Intelligence and Machine Learning with Physics Sciences (MLPS)
专题命中 安全评测 :safety(abstract);分类 cs.AI、cs.LG
Comments 7 pages, 4 figures, 1 table. To appear in the proceedings of 2022 IEEE International Conference on Robotics and Automation (ICRA)
专题命中 安全评测 :safety(abstract);分类 cs.AI、cs.LG
Comments Accepted to AAAI 2022. 9 pages
专题命中 安全评测 :alignment(abstract);分类 cs.AI、cs.LG
Comments Updated version with a correction. The full draft was submitted in Jan 2021. The Voice2Series project initially was launched in Sep 2020. Accepted to ICML 2021, 16 Pages
Journal ref Proceedings of the 38th International Conference on Machine Learning 2021
专题命中 安全评测 :safety(abstract);分类 cs.AI、cs.LG
专题命中 安全评测 :trustworthy(abstract);分类 cs.CL、cs.AI
Comments Code and dataset are available at https://github.com/gjdnju/MLMN
专题命中 安全评测 :safety(abstract);分类 cs.AI、cs.LG
Comments Accepted at Cooperative AI Workshop, NeurIPS 2021
专题命中 安全评测 :alignment(abstract);分类 cs.AI、cs.LG
Comments 12 pages, 5 figures
专题命中 安全评测 :safety(abstract);分类 cs.AI、cs.LG
Comments To appear in NeurIPS 2021 Workshop on Machine Learning for Autonomous Driving (ML4AD)
专题命中 安全评测 :trustworthy(abstract);分类 cs.AI、cs.LG
Comments Medical Imaging meets NeurIPS 2021 workshop
专题命中 安全评测 :trustworthy(abstract);分类 cs.AI、cs.LG
专题命中 安全评测 :alignment(abstract);分类 cs.CL、cs.AI
专题命中 安全评测 :trustworthy(abstract);分类 cs.AI、cs.LG
Comments Accepted to 35th Conference on Neural Information Processing Systems (NeurIPS), 2021
专题命中 安全评测 :safety(abstract);分类 cs.AI、cs.LG
Comments In Proceedings FMAS 2021, arXiv:2110.11527
Journal ref EPTCS 348, 2021, pp. 92-100
专题命中 安全评测 :safety(abstract);分类 cs.AI、cs.LG
Comments Published at 2021 ICML RL4RL Workshop; Submitted to 2022 PSCC
专题命中 安全评测 :safety(abstract);分类 cs.AI、cs.LG
Comments 8 pages, 6 figures
专题命中 安全评测 :safety(abstract);分类 cs.AI、cs.LG
Comments Accepted by IEEE International Conference on Computer Vision (ICCV) 2021
专题命中 安全评测 :trustworthy(abstract);分类 cs.AI、cs.LG
Comments Keywords: Admissible machine learning; InfoGram; L-Features; Information-theory; ALFA-testing, Algorithmic risk management; Fairness; Interpretability; COREml; FINEml
专题命中 安全评测 :safety(abstract);分类 cs.AI、cs.LG
Comments 12 pages, 7 figures. The final publication is available at Springer via https://doi.org/10.1007/978-3-030-79150-6_21
Journal ref In: Maglogiannis I., Macintyre J., Iliadis L. (eds) Artificial Intelligence Applications and Innovations. AIAI 2021. IFIP Advances in Information and Communication Technology, vol 627. Springer, Cham
专题命中 安全评测 :safety(abstract);分类 cs.AI、cs.LG
Comments Accepted at IROS'21. Author version with 7 pages, 5 figures, 2 tables, and 1 algorithm
专题命中 安全评测 :safety(abstract);分类 cs.AI、cs.LG
Comments Accepted by The 15th ACM International Conference on Distributed and Event-based Systems (DEBS) 2021
专题命中 安全评测 :safety(abstract);分类 cs.AI、cs.LG
Comments Accepted by the ICML 2021 workshop on "A Blessing in Disguise: The Prospects and Perils of Adversarial Machine Learning"