[1610.00768] Technical Report on the CleverHans v2.1.0 Adversarial Examples Library
CleverHans is a software library that provides standardized reference implementations of adversarial example construction techniques and adversarial training. The library may be used to develop more robust machine learning models and to provide standardized benchmarks of models' performance in the adversarial setting. Benchmarks constructed without a standardized implementation of adversarial example construction are not comparable to each other, because a good result may indicate a robust model or it may merely indicate a weak implementation of the adversarial example construction procedure. This technical report is structured as follows. Section 1 provides an overview of adversarial examples in machine learning and of the CleverHans software. Section 2 presents the core functionalities of the library: namely the attacks based on adversarial examples and defenses to improve the robustness of machine learning models to these attacks. Section 3 describes how to report benchmark results using the library. Section 4 describes the versioning system.
Technical Report on the cleverhans v2.1.0 Adversarial Examples Library arXiv:1610.00768v6 [cs.LG] 27 Jun 2018 Nicolas Papernot∗1,3 , Fartash Faghri5,3 , Nicholas Carlini2,3 , Ian Goodfellow†3 , Reuben Feinman4 , Alexey Kurakin3 , Cihang Xie6 , Yash Sharma7 , Tom Brown3 , Aurko Roy3 , Alexander Matyasko8 , Vahid Behzadan9 , Karen Hambardzumyan10 , Zhishuai Zhang6 ,…
saved by
related reading
- Some Lessons from Adversarial Machine Learning | FAR.AIfar.ai
- GitHub - SoyGema/pulling_acegithub.com
- Nicholas Carlininicholas.carlini.com
- Adversarial Examples Are Not Bugs, They Are Features – gradient sciencegradientscience.org
- [2003.01690] Reliable evaluation of adversarial robustness with an ensemble of diverse parameter-free attacksarxiv.org
- GitHub - requie/AI-Red-Teaming-Guide: A comprehensive guide to adversarial testing and security evaluation of AI systems, helping organizations identify vulnerabilities before attackers exploit them.github.com
- MAI-Thinking-1: Building a Hill-Climbing Machinemicrosoft.ai
- Goodfire AIgoodfire.ai
- The Dimpled Manifold Model of Adversarial Examples in Machine Learningarxiv.org
- 2306.15447.pdfarxiv.org
- Solving adversarial attacks in computer vision as a baby version of general AI alignment | Stanislav Fortstanislavfort.com
- Is BERT Really Robust? A Strong Baseline for Natural Language Attack on Text Classification and Entailmentarxiv.org