arXiv · 2104.05847
Targeted Adversarial Training for Natural Language Understanding
Abstract
We present a simple yet effective Targeted Adversarial Training (TAT) algorithm to improve adversarial training for natural language understanding. The key idea is to introspect current mistakes and prioritize adversarial training steps to where the model errs the most. Experiments show that TAT can significantly improve accuracy over standard adversarial training on GLUE and attain new state-of-the-art zero-shot results on XNLI. Our code will be released at: https://github.com/namisan/mt-dnn.
Explore related subjects
Keep this discovery
Lis Pereira, Xiaodong Liu, Hao Cheng, Hoifung Poon, Jianfeng Gao, Ichiro Kobayashi. 2021-04-12. Targeted Adversarial Training for Natural Language Understanding. https://arxiv.org/abs/2104.05847
Cite the original work for its findings. Save a collection to share your selection of sources.