arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~
arXiv 1910.13875cs.CRcs.DCcs.LG

Fault Tolerance of Neural Networks in Adversarial Settings

  • Indraprastha Institute of Information Technology(英迪拉普拉斯信息技术研究所)
  • Scientific Analysis Group(科学分析小组)
  • Institute for Systems Studies and Analyses(系统研究与分析研究所)
  • Aurel Vlaicu University of Arad(阿拉德奥雷尔弗拉伊库大学)

机构由 AI 辅助整理,请以论文原文为准。

Vasisht Duddu, N. Rajesh Pillai, D. Vijay Rao, Valentina E. Balas

更新

英文摘要:

Artificial Intelligence systems require a through assessment of different pillars of trust, namely, fairness, interpretability, data and model privacy, reliability (safety) and robustness against against adversarial attacks. While these research problems have been extensively studied in isolation, an understanding of the trade-off between different pillars of trust is lacking. To this extent, the trade-off between fault tolerance, privacy and adversarial robustness is evaluated for the specific case of Deep Neural Networks, by considering two adversarial settings under a security and a privacy threat model. Specifically, this work studies the impact of the fault tolerance of the Neural Network on training the model by adding noise to the input (Adversarial Robustness) and noise to the gradients (Differential Privacy). While training models with noise to inputs, gradients or weights enhances fault tolerance, it is observed that adversarial robustness and fault tolerance are at odds with each other. On the other hand, ($ε,δ$)-Differentially Private models enhance the fault tolerance, measured using generalisation error, theoretically has an upper bound of $e^ε - 1 + δ$. This novel study of the trade-off between different elements of trust is pivotal for training a model which satisfies the requirements for different pillars of trust simultaneously.

补充信息

↑