Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Better Robustness by More Coverage: Adversarial Training with Mixup Augmentation for Robust Fine-tuning

Dec 31, 2020

Chenglei Si, Zhengyan Zhang, Fanchao Qi, Zhiyuan Liu, Yasheng Wang, Qun Liu, Maosong Sun

Figure 1 for Better Robustness by More Coverage: Adversarial Training with Mixup Augmentation for Robust Fine-tuning

Figure 2 for Better Robustness by More Coverage: Adversarial Training with Mixup Augmentation for Robust Fine-tuning

Figure 3 for Better Robustness by More Coverage: Adversarial Training with Mixup Augmentation for Robust Fine-tuning

Figure 4 for Better Robustness by More Coverage: Adversarial Training with Mixup Augmentation for Robust Fine-tuning

Share this with someone who'll enjoy it:

Abstract:Pre-trained language models (PLMs) fail miserably on adversarial attacks. To improve the robustness, adversarial data augmentation (ADA) has been widely adopted, which attempts to cover more search space of adversarial attacks by adding the adversarial examples during training. However, the number of adversarial examples added by ADA is extremely insufficient due to the enormously large search space. In this work, we propose a simple and effective method to cover much larger proportion of the attack search space, called Adversarial Data Augmentation with Mixup (MixADA). Specifically, MixADA linearly interpolates the representations of pairs of training examples to form new virtual samples, which are more abundant and diverse than the discrete adversarial examples used in conventional ADA. Moreover, to evaluate the robustness of different models fairly, we adopt a challenging setup, which dynamically generates new adversarial examples for each model. In the text classification experiments of BERT and RoBERTa, MixADA achieves significant robustness gains under two strong adversarial attacks and alleviates the performance degradation of ADA on the original data. Our source codes will be released to support further explorations.

* 9 pages

View paper on

Share this with someone who'll enjoy it:

Title:Better Robustness by More Coverage: Adversarial Training with Mixup Augmentation for Robust Fine-tuning

Paper and Code