Have You Poisoned My Data? Defending Neural Networks against Data Poisoning

De Gaspari, Fabio; Hitaj, Dorjan; Mancini, Luigi V.

Computer Science > Machine Learning

arXiv:2403.13523 (cs)

[Submitted on 20 Mar 2024]

Title:Have You Poisoned My Data? Defending Neural Networks against Data Poisoning

Authors:Fabio De Gaspari, Dorjan Hitaj, Luigi V. Mancini

View PDF HTML (experimental)

Abstract:The unprecedented availability of training data fueled the rapid development of powerful neural networks in recent years. However, the need for such large amounts of data leads to potential threats such as poisoning attacks: adversarial manipulations of the training data aimed at compromising the learned model to achieve a given adversarial goal.
This paper investigates defenses against clean-label poisoning attacks and proposes a novel approach to detect and filter poisoned datapoints in the transfer learning setting. We define a new characteristic vector representation of datapoints and show that it effectively captures the intrinsic properties of the data distribution. Through experimental analysis, we demonstrate that effective poisons can be successfully differentiated from clean points in the characteristic vector space. We thoroughly evaluate our proposed approach and compare it to existing state-of-the-art defenses using multiple architectures, datasets, and poison budgets. Our evaluation shows that our proposal outperforms existing approaches in defense rate and final trained model performance across all experimental settings.

Comments:	Paper accepted for publication at European Symposium on Research in Computer Security (ESORICS) 2024
Subjects:	Machine Learning (cs.LG); Artificial Intelligence (cs.AI); Cryptography and Security (cs.CR)
Cite as:	arXiv:2403.13523 [cs.LG]
	(or arXiv:2403.13523v1 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2403.13523

Submission history

From: Fabio De Gaspari [view email]
[v1] Wed, 20 Mar 2024 11:50:16 UTC (1,985 KB)

Computer Science > Machine Learning

Title:Have You Poisoned My Data? Defending Neural Networks against Data Poisoning

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Have You Poisoned My Data? Defending Neural Networks against Data Poisoning

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators