Toward unsupervised, multi-object discovery in large-scale image collections

Vo, Huy V.; Pérez, Patrick; Ponce, Jean

Computer Science > Computer Vision and Pattern Recognition

arXiv:2007.02662 (cs)

[Submitted on 6 Jul 2020 (v1), last revised 25 Aug 2020 (this version, v2)]

Title:Toward unsupervised, multi-object discovery in large-scale image collections

Authors:Huy V. Vo, Patrick Pérez, Jean Ponce

View PDF

Abstract:This paper addresses the problem of discovering the objects present in a collection of images without any supervision. We build on the optimization approach of Vo et al. (CVPR'19) with several key novelties: (1) We propose a novel saliency-based region proposal algorithm that achieves significantly higher overlap with ground-truth objects than other competitive methods. This procedure leverages off-the-shelf CNN features trained on classification tasks without any bounding box information, but is otherwise unsupervised. (2) We exploit the inherent hierarchical structure of proposals as an effective regularizer for the approach to object discovery of Vo et al., boosting its performance to significantly improve over the state of the art on several standard benchmarks. (3) We adopt a two-stage strategy to select promising proposals using small random sets of images before using the whole image collection to discover the objects it depicts, allowing us to tackle, for the first time (to the best of our knowledge), the discovery of multiple objects in each one of the pictures making up datasets with up to 20,000 images, an over five-fold increase compared to existing methods, and a first step toward true large-scale unsupervised image interpretation.

Comments:	Accepted for publication in European Conference on Computer Vision (ECCV) 2020
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2007.02662 [cs.CV]
	(or arXiv:2007.02662v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2007.02662

Submission history

From: Van Huy Vo [view email]
[v1] Mon, 6 Jul 2020 11:43:47 UTC (6,630 KB)
[v2] Tue, 25 Aug 2020 11:11:31 UTC (6,526 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Toward unsupervised, multi-object discovery in large-scale image collections

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Toward unsupervised, multi-object discovery in large-scale image collections

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators