Safe Sample Screening for Support Vector Machines

Ogawa, Kohei; Suzuki, Yoshiki; Suzumura, Shinya; Takeuchi, Ichiro

Statistics > Machine Learning

arXiv:1401.6740 (stat)

[Submitted on 27 Jan 2014]

Title:Safe Sample Screening for Support Vector Machines

Authors:Kohei Ogawa, Yoshiki Suzuki, Shinya Suzumura, Ichiro Takeuchi

View PDF

Abstract:Sparse classifiers such as the support vector machines (SVM) are efficient in test-phases because the classifier is characterized only by a subset of the samples called support vectors (SVs), and the rest of the samples (non SVs) have no influence on the classification result. However, the advantage of the sparsity has not been fully exploited in training phases because it is generally difficult to know which sample turns out to be SV beforehand. In this paper, we introduce a new approach called safe sample screening that enables us to identify a subset of the non-SVs and screen them out prior to the training phase. Our approach is different from existing heuristic approaches in the sense that the screened samples are guaranteed to be non-SVs at the optimal solution. We investigate the advantage of the safe sample screening approach through intensive numerical experiments, and demonstrate that it can substantially decrease the computational cost of the state-of-the-art SVM solvers such as LIBSVM. In the current big data era, we believe that safe sample screening would be of great practical importance since the data size can be reduced without sacrificing the optimality of the final solution.

Comments:	A preliminary version was presented at ICML2013
Subjects:	Machine Learning (stat.ML)
Cite as:	arXiv:1401.6740 [stat.ML]
	(or arXiv:1401.6740v1 [stat.ML] for this version)
	https://doi.org/10.48550/arXiv.1401.6740

Submission history

From: Ichiro Takeuchi Prof. [view email]
[v1] Mon, 27 Jan 2014 04:41:37 UTC (2,225 KB)

Statistics > Machine Learning

Title:Safe Sample Screening for Support Vector Machines

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Statistics > Machine Learning

Title:Safe Sample Screening for Support Vector Machines

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators