← Back to all news
arXiv Machine Learning June 2, 2026 By Kai Wang

Sensitivity as a Double-Edged Sword: A Trade-off Between Discriminability and Adversarial Robustness

Read the original on arXiv Machine Learning →

arXiv:2606. 01746v1 Announce Type: cross Abstract: Modern neural networks are highly susceptible to adversarial perturbations.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.

  • fine-tuning
  • benchmarks
  • safety

Related stories

arXiv AI
Jun 2

CEAR: Certified Ensemble Adversarial Robustness in DNNs

arXiv:2606. 01437v1 Announce Type: cross Abstract: Deep Neural Networks (DNNs) are highly susceptible to adversarial perturbations, leading to extensive research on robustness for safety-critical applications.

By Daniel Sadig, Mohammadreza Maleki, Hamed Karimi, Reza Samavi
benchmarkssafety
More like this →
arXiv Machine Learning
Jul 7

Robustness Meets Uncertainty: Evidential Adversarial Training for Robust Selective Classification

arXiv:2607. 03075v1 Announce Type: new Abstract: Safety-critical applications require classifiers that are both robust and reliable.

By Nicolas Sournac, Ahmed Baha Ben Jmaa, Bertrand Braeckeveldt
benchmarkssafety
More like this →
arXiv AI
Jun 29

Improving Adversarial Robustness via Activation Amplification and Attenuation

arXiv:2606. 27784v1 Announce Type: cross Abstract: The existence of adversarial attacks is often attributed to the presence of non-robust features in neural networks.

By Ta\"iga Gon\c{c}alves, Yongsong Huang, Tomo Miyazaki, Shinichiro Omachi
efficiencysafety
More like this →
arXiv AI
Jul 1

Improving Certified Robustness via Adversarial Distillation

arXiv:2606. 31653v1 Announce Type: cross Abstract: Certified training aims to produce models whose predictions can be formally verified against adversarial perturbations, typically by optimising upper bounds on the worst-case loss over an allowed perturbation set.

By Matteo Melis, Jesus Martinez Del Rincon, Vishal Sharma
efficiencybenchmarkssafety
More like this →
arXiv Machine Learning
Jul 7

Binary Iterative Method for Non-targeted Adversarial Attack

arXiv:2607. 04145v1 Announce Type: new Abstract: Adversarial attacks guide and provide additional training and test data for both adversarial training and adversarial robustness validation, and expose the 'piecewise linearity' of deep learning based models.

By Naman Goyal, Milan Chaudhari
safety
More like this →
arXiv Machine Learning
Aug 4

DeepDefense: Robust Learning via Layer-Wise Gradient-Feature Alignment

arXiv:2511. 13749v2 Announce Type: replace Abstract: Deep neural networks are known to be vulnerable to adversarial perturbations, which are small, carefully crafted inputs that lead to incorrect predictions.

By Ci Lin, Tet Yeap, Iluju Kiringa
safety
More like this →