Improving the Reliability of Deep Neural Networks in NLP: A Review

被引:106
作者
Alshemali, Basemah [1 ,2 ]
Kalita, Jugal [2 ]
机构
[1] Taibah Univ, Al Medina, Saudi Arabia
[2] Univ Colorado, Colorado Springs, CO 80907 USA
关键词
Adversarial examples; Adversarial texts; Natural language processing; ADVERSARIAL ATTACKS;
D O I
10.1016/j.knosys.2019.105210
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Deep learning models have achieved great success in solving a variety of natural language processing (NLP) problems. An ever-growing body of research, however, illustrates the vulnerability of deep neural networks (DNNs) to adversarial examples - inputs modified by introducing small perturbations to deliberately fool a target model into outputting incorrect results. The vulnerability to adversarial examples has become one of the main hurdles precluding neural network deployment into safety-critical environments. This paper discusses the contemporary usage of adversarial examples to foil DNNs and presents a comprehensive review of their use to improve the robustness of DNNs in NLP applications. In this paper, we summarize recent approaches for generating adversarial texts and propose a taxonomy to categorize them. We further review various types of defensive strategies against adversarial examples, explore their main challenges, and highlight some future research directions. (C) 2019 Elsevier B.V. All rights reserved.
引用
收藏
页数:19
相关论文
共 134 条
[1]   Defense against Universal Adversarial Perturbations [J].
Akhtar, Naveed ;
Liu, Jian ;
Mian, Ajmal .
2018 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR), 2018, :3389-3398
[2]  
Akula S, 2017, J HIGH ENERGY PHYS, DOI 10.1007/JHEP11(2017)051
[3]  
Alshemali B., 2019, International Journal of Computer Applications, V178, P1
[4]  
Alzantot M, 2018, 2018 CONFERENCE ON EMPIRICAL METHODS IN NATURAL LANGUAGE PROCESSING (EMNLP 2018), P2890
[5]  
[Anonymous], 2018, 6 INT C LEARN REPR I
[6]  
[Anonymous], P C N AM CHAPT ASS C
[7]  
[Anonymous], INT WORKSH SPOK LANG
[8]  
[Anonymous], 2018, INT C LEARN REPR
[9]  
[Anonymous], T ASS COMPUT LINGUIS
[10]  
[Anonymous], 2013, P 2013 C EMPIRICAL M