arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

共收录 9434 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 安全评测 9434 篇

2209.01174 2022-12-01 cs.CL 57%

Extend and Explain: Interpreting Very Long Language Models

Joel Stremmel, Brian L. Hill, Jeffrey Hertzberg, Jaime Murillo, Llewelyn Allotey, Eran Halperin

专题命中 安全评测 :trustworthy(abstract);分类 cs.CL

Comments 11 pages

Journal ref Proceedings of the 2nd Machine Learning for Health symposium, PMLR 193:218-258, 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.10354 2022-11-30 cs.LG 57%

Quantifying probabilistic robustness of tree-based classifiers against natural distortions

Christoph Schweimer, Sebastian Scher

专题命中 安全评测 :trustworthy(abstract);分类 cs.LG

Comments 9 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.14095 2022-11-28 cs.RO cs.AI cs.HC cs.MA 57%

A Hierarchical Variable Autonomy Mixed-Initiative Framework for Human-Robot Teaming in Mobile Robotics

Dimitris Panagopoulos, Giannis Petousakis, Aniketh Ramesh, Tianshu Ruan, Grigoris Nikolaou, Rustam Stolkin, Manolis Chiou

专题命中 安全评测 :safety(abstract);分类 cs.AI

Comments 6 pages, 4 figures, ICHMS 2022, First two Authors contributed equally

详情

展开后加载摘要…

URL PDF HTML 收藏
2011.09756 2022-11-28 cs.RO cs.AI 57%

Active Inference and Behavior Trees for Reactive Action Planning and Execution in Robotics

Corrado Pezzato, Carlos Hernandez Corbato, Stefan Bonhof, Martijn Wisse

专题命中 安全评测 :safety(abstract);分类 cs.AI

Comments Accepted at Transactions on Robotics (IEEE T-RO)

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.12326 2022-11-23 cs.LG 57%

PreMa: Predictive Maintenance of Solenoid Valve in Real-Time at Embedded Edge-Level

Prajwal BN, Harsha Yelchuri, Vishwanath Shastry, T. V. Prabhakar

专题命中 安全评测 :safety(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.04874 2022-11-22 cs.SI cs.CY 57%

Doctors vs. Nurses: Understanding the Great Divide in Vaccine Hesitancy among Healthcare Workers

Sajid Hussain Rafi Ahamed, Shahid Shakil, Hanjia Lyu, Xinping Zhang, Jiebo Luo

专题命中 安全评测 :trustworthy(abstract);分类 cs.CY

Comments Accepted for publication in Proceedings of the 8th Special Session on Intelligent Data Mining of 2022 IEEE International Conference on Big Data, 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.09084 2022-11-17 cs.SE cs.AI 57%

Technical Report on Neural Language Models and Few-Shot Learning for Systematic Requirements Processing in MDSE

Vincent Bertram, Miriam Boß, Evgeny Kusmenko, Imke Helene Nachmann, Bernhard Rumpe, Danilo Trotta, Louis Wachtmeister

专题命中 安全评测 :safety(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.01175 2022-11-17 eess.SY cs.LG cs.SY 57%

Robust Longitudinal Control for Vehicular Autonomous Platoons Using Deep Reinforcement Learning

Armando Alves Neto, Leonardo Amaral Mozelli

专题命中 安全评测 :safety(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.11097 2022-11-16 cs.CL 57%

A Fine-grained Interpretability Evaluation Benchmark for Neural NLP

Lijie Wang, Yaozong Shen, Shuyuan Peng, Shuai Zhang, Xinyan Xiao, Hao Liu, Hongxuan Tang, Ying Chen, Hua Wu, Haifeng Wang

专题命中 安全评测 :trustworthy(abstract);分类 cs.CL

Journal ref CoNLL 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.00973 2022-11-15 cs.LG cs.CV cs.MS eess.SP math.OC 57%

NCVX: A General-Purpose Optimization Solver for Constrained Machine and Deep Learning

Buyun Liang, Tim Mitchell, Ju Sun

专题命中 安全评测 :trustworthy(abstract);分类 cs.LG

Comments Accepted by the NeurIPS Workshop on Optimization for Machine Learning (OPT 2022). arXiv admin note: text overlap with arXiv:2111.13984

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.08566 2022-11-10 cs.AI cs.AR cs.SY eess.SY 57%

Physical Computing for Materials Acceleration Platforms

Erik Peterson, Alexander Lavin

专题命中 安全评测 :alignment(abstract);分类 cs.AI

Journal ref MATTER, VOLUME 5, ISSUE 11, P3586-3596, NOVEMBER 02, 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.03783 2022-11-08 physics.med-ph cs.AI 57%

Issues and Challenges in Applications of Artificial Intelligence to Nuclear Medicine -- The Bethesda Report (AI Summit 2022)

Arman Rahmim, Tyler J. Bradshaw, Irène Buvat, Joyita Dutta, Abhinav K. Jha, Paul E. Kinahan, Quanzheng Li, Chi Liu, Melissa D. McCradden, Babak Saboury, Eliot Siegel, John J. Sunderland, Richard L. Wahl

专题命中 安全评测 :trustworthy(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.16031 2022-11-04 cs.CV cs.CL 57%

UPainting: Unified Text-to-Image Diffusion Generation with Cross-modal Guidance

Wei Li, Xue Xu, Xinyan Xiao, Jiachen Liu, Hu Yang, Guohao Li, Zhanpeng Wang, Zhifan Feng, Qiaoqiao She, Yajuan Lyu, Hua Wu

专题命中 安全评测 :alignment(abstract);分类 cs.CL

Comments First Version, 16 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.01200 2022-11-03 cs.CL 57%

Multi-level Distillation of Semantic Knowledge for Pre-training Multilingual Language Model

Mingqi Li, Fei Ding, Dan Zhang, Long Cheng, Hongxin Hu, Feng Luo

专题命中 安全评测 :alignment(abstract);分类 cs.CL

Comments accepted at EMNLP 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.01790 2022-11-03 cs.LG 57%

Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Rohin Shah, Vikrant Varma, Ramana Kumar, Mary Phuong, Victoria Krakovna, Jonathan Uesato, Zac Kenton

专题命中 安全评测 :alignment(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1709.00084 2022-10-26 cs.RO cs.AI 57%

Behavior Trees in Robotics and AI: An Introduction

Michele Colledanchise, Petter Ögren

专题命中 安全评测 :safety(abstract);分类 cs.AI

Journal ref Chapman & Hall/CRC Artificial Intelligence and Robotics Series 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.12560 2022-10-25 cs.CL 57%

PHEE: A Dataset for Pharmacovigilance Event Extraction from Text

Zhaoyue Sun, Jiazheng Li, Gabriele Pergola, Byron C. Wallace, Bino John, Nigel Greene, Joseph Kim, Yulan He

专题命中 安全评测 :safety(abstract);分类 cs.CL

Comments 17 pages, 3 figures, EMNLP2022 accepted

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.12140 2022-10-24 cs.AI cs.MA 57%

Explainability in autonomous pedagogically structured scenarios

Minal Suresh Patil

专题命中 安全评测 :trustworthy(abstract);分类 cs.AI

Comments Explainable Agency in Artificial Intelligence Workshop - 36th AAAI Conference on Artificial Intelligence

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.10692 2022-10-21 cs.CL 57%

Separating Grains from the Chaff: Using Data Filtering to Improve Multilingual Translation for Low-Resourced African Languages

Idris Abdulmumin, Michael Beukman, Jesujoba O. Alabi, Chris Emezue, Everlyn Asiko, Tosin Adewumi, Shamsuddeen Hassan Muhammad, Mofetoluwa Adeyemi, Oreen Yousuf, Sahib Singh, Tajuddeen Rabiu Gwadabe

专题命中 安全评测 :alignment(abstract);分类 cs.CL

Comments Accepted at the Seventh Conference on Machine Translation (WMT22)

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.08742 2022-10-18 cs.CL 57%

Tencent AI Lab - Shanghai Jiao Tong University Low-Resource Translation System for the WMT22 Translation Task

Zhiwei He, Xing Wang, Zhaopeng Tu, Shuming Shi, Rui Wang

专题命中 安全评测 :alignment(abstract);分类 cs.CL

Comments WMT 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.06682 2022-10-14 cs.CV cs.AI 57%

Application-Driven AI Paradigm for Hand-Held Action Detection

Kohou Wang, Zhaoxiang Liu, Shiguo Lian

专题命中 安全评测 :safety(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.06649 2022-10-14 cs.AI cs.NI 57%

Neuro-symbolic Explainable Artificial Intelligence Twin for Zero-touch IoE in Wireless Network

Md. Shirajum Munir, Ki Tae Kim, Apurba Adhikary, Walid Saad, Sachin Shetty, Seong-Bae Park, Choong Seon Hong

专题命中 安全评测 :trustworthy(abstract);分类 cs.AI

Comments Submitted to a journal for peer review

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.06447 2022-10-13 cs.LG stat.ML 57%

Sampling in Constrained Domains with Orthogonal-Space Variational Gradient Descent

Ruqi Zhang, Qiang Liu, Xin T. Tong

专题命中 安全评测 :safety(abstract);分类 cs.LG

Comments NeurIPS 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.12227 2022-10-12 cs.CR cs.LG cs.SE 57%

Adversarial Robustness of Deep Neural Networks: A Survey from a Formal Verification Perspective

Mark Huasong Meng, Guangdong Bai, Sin Gee Teo, Zhe Hou, Yan Xiao, Yun Lin, Jin Song Dong

专题命中 安全评测 :trustworthy(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.02533 2022-10-07 cs.HC cs.CV cs.CY 57%

Towards Semi-automatic Detection and Localization of Indoor Accessibility Issues using Mobile Depth Scanning and Computer Vision

Xia Su, Kaiming Cheng, Han Zhang, Jaewook Lee, Jon E. Froehlich

专题命中 安全评测 :safety(abstract);分类 cs.CY

Comments Workshop paper presented at "The 1st ASSETS'22 Workshop on The Future or urban Accessibility (UrbanAccess'22)"

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.15438 2022-10-03 cs.SE cs.LG 57%

Empowering the trustworthiness of ML-based critical systems through engineering activities

Juliette Mattioli, Agnes Delaborde, Souhaiel Khalfaoui, Freddy Lecue, Henri Sohier, Frederic Jurie

专题命中 安全评测 :trustworthy(abstract);分类 cs.LG

Comments This work has been supported by the French government under the "France 2030" program, as part of the SystemX Technological Research Institute Research Institute

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.14749 2022-09-27 cs.RO cs.AI 57%

Adaptive Risk-Tendency: Nano Drone Navigation in Cluttered Environments with Distributional Reinforcement Learning

Cheng Liu, Erik-Jan van Kampen, Guido C. H. E. de Croon

专题命中 安全评测 :safety(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.07834 2022-09-22 cs.CL 57%

Bridging Cross-Lingual Gaps During Leveraging the Multilingual Sequence-to-Sequence Pretraining for Text Generation and Understanding

Changtong Zan, Liang Ding, Li Shen, Yu Cao, Weifeng Liu, Dacheng Tao

专题命中 安全评测 :alignment(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.08922 2022-09-22 cs.LG 57%

FLAME: Federated Learning Across Multi-device Environments

Hyunsung Cho, Akhil Mathur, Fahim Kawsar

专题命中 安全评测 :alignment(abstract);分类 cs.LG

Journal ref Proc. ACM Interact. Mob. Wearable Ubiquitous Technol. 6, 3, Article 107 (September 2022)

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.09666 2022-09-21 cs.SE cs.AI 57%

Documenting use cases in the affective computing domain using Unified Modeling Language

Isabelle Hupont, Emilia Gomez

专题命中 安全评测 :trustworthy(abstract);分类 cs.AI

Comments 8 pages, 5 figures, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏