Evaluating Defensive Distillation For Defending Text Processing Neural Networks Against Adversarial Examples

Soll, Marcus; Hinz, Tobias; Magg, Sven; Wermter, Stefan

Computer Science > Computation and Language

arXiv:1908.07899 (cs)

[Submitted on 21 Aug 2019]

Title:Evaluating Defensive Distillation For Defending Text Processing Neural Networks Against Adversarial Examples

Authors:Marcus Soll, Tobias Hinz, Sven Magg, Stefan Wermter

View PDF

Abstract:Adversarial examples are artificially modified input samples which lead to misclassifications, while not being detectable by humans. These adversarial examples are a challenge for many tasks such as image and text classification, especially as research shows that many adversarial examples are transferable between different classifiers. In this work, we evaluate the performance of a popular defensive strategy for adversarial examples called defensive distillation, which can be successful in hardening neural networks against adversarial examples in the image domain. However, instead of applying defensive distillation to networks for image classification, we examine, for the first time, its performance on text classification tasks and also evaluate its effect on the transferability of adversarial text examples. Our results indicate that defensive distillation only has a minimal impact on text classifying neural networks and does neither help with increasing their robustness against adversarial examples nor prevent the transferability of adversarial examples between neural networks.

Comments:	Published at the International Conference on Artificial Neural Networks (ICANN) 2019
Subjects:	Computation and Language (cs.CL); Cryptography and Security (cs.CR); Machine Learning (cs.LG); Neural and Evolutionary Computing (cs.NE)
Cite as:	arXiv:1908.07899 [cs.CL]
	(or arXiv:1908.07899v1 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.1908.07899

Submission history

From: Tobias Hinz [view email]
[v1] Wed, 21 Aug 2019 14:50:13 UTC (69 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.CL

< prev | next >

new | recent | 2019-08

Change to browse by:

cs
cs.CR
cs.LG
cs.NE

References & Citations

DBLP - CS Bibliography

listing | bibtex

Marcus Soll
Tobias Hinz
Sven Magg
Stefan Wermter

export BibTeX citation

Computer Science > Computation and Language

Title:Evaluating Defensive Distillation For Defending Text Processing Neural Networks Against Adversarial Examples

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Evaluating Defensive Distillation For Defending Text Processing Neural Networks Against Adversarial Examples

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators