OpenAI Blog

Attacking machine learning with adversarial examples

Read the original on OpenAI Blog →

Adversarial examples are inputs to machine learning models that an attacker has intentionally designed to cause the model to make a mistake; they’re like optical illusions for machines. In this post we’ll show how adversarial examples work across different mediums, and will discuss why securing systems against them can be difficult.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at OpenAI Blog.

arXiv AI
Aug 25

A New Type of Adversarial Examples

The paper introduces a new class of adversarial examples that are markedly different from original inputs yet produce the same model output. It presents algorithms such as NI-FGSM, NI-FGM, and their momentum variants (NMI-FGSM, NMI-FGM) to generate these examples. The authors demonstrate that these adversarial examples are not confined to the vicinity of training data but are spread throughout the sample space.

By Xingyang Nie, Caoliang Zhang, Su Pan, Biao Wang, Huilin Ge, Tao Fang
arXiv Machine Learning
Jun 26

Over-parameterization and Adversarial Robustness in Neural Networks: An Overview and Empirical Analysis

arXiv:2406. 10090v3 Announce Type: replace Abstract: Thanks to their extensive capacity, over-parameterized neural networks exhibit superior predictive capabilities and generalization.

By Srishti Gupta, Zhang Chen, Luca Demetrio, Fabio Brau, Xiaoyi Feng, Zhaoqiang Xia, Antonio Emanuele Cin\`a, Maura Pintor, Luca Oneto, Ambra Demontis, Battista Biggio, Fabio Roli