Odyssey: Creation, Analysis and Detection of Trojan Models

Edraki, Marzieh; Karim, Nazmul; Rahnavard, Nazanin; Mian, Ajmal; Shah, Mubarak

Computer Science > Computer Vision and Pattern Recognition

arXiv:2007.08142 (cs)

[Submitted on 16 Jul 2020 (v1), last revised 8 Dec 2020 (this version, v2)]

Title:Odyssey: Creation, Analysis and Detection of Trojan Models

Authors:Marzieh Edraki, Nazmul Karim, Nazanin Rahnavard, Ajmal Mian, Mubarak Shah

View PDF

Abstract:Along with the success of deep neural network (DNN) models, rise the threats to the integrity of these models. A recent threat is the Trojan attack where an attacker interferes with the training pipeline by inserting triggers into some of the training samples and trains the model to act maliciously only for samples that contain the trigger. Since the knowledge of triggers is privy to the attacker, detection of Trojan networks is challenging. Existing Trojan detectors make strong assumptions about the types of triggers and attacks. We propose a detector that is based on the analysis of the intrinsic DNN properties; that are affected due to the Trojaning process. For a comprehensive analysis, we develop Odysseus, the most diverse dataset to date with over 3,000 clean and Trojan models. Odysseus covers a large spectrum of attacks; generated by leveraging the versatility in trigger designs and source to target class map**s. Our analysis results show that Trojan attacks affect the classifier margin and shape of decision boundary around the manifold of clean data. Exploiting these two factors, we propose an efficient Trojan detector that operates without any knowledge of the attack and significantly outperforms existing methods. Through a comprehensive set of experiments we demonstrate the efficacy of the detector on cross model architectures, unseen Triggers and regularized models.

Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2007.08142 [cs.CV]
	(or arXiv:2007.08142v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2007.08142

Submission history

From: Marzieh Edraki [view email]
[v1] Thu, 16 Jul 2020 06:55:00 UTC (5,643 KB)
[v2] Tue, 8 Dec 2020 08:09:51 UTC (828 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Odyssey: Creation, Analysis and Detection of Trojan Models

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Odyssey: Creation, Analysis and Detection of Trojan Models

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators