3D AffordanceNet: A Benchmark for Visual Object Affordance Understanding

Deng, Shengheng; Xu, Xun; Wu, Chaozheng; Chen, Ke; Jia, Kui

Computer Science > Computer Vision and Pattern Recognition

arXiv:2103.16397 (cs)

[Submitted on 30 Mar 2021 (v1), last revised 31 Mar 2021 (this version, v2)]

Title:3D AffordanceNet: A Benchmark for Visual Object Affordance Understanding

Authors:Shengheng Deng, Xun Xu, Chaozheng Wu, Ke Chen, Kui Jia

View PDF

Abstract:The ability to understand the ways to interact with objects from visual cues, a.k.a. visual affordance, is essential to vision-guided robotic research. This involves categorizing, segmenting and reasoning of visual affordance. Relevant studies in 2D and 2.5D image domains have been made previously, however, a truly functional understanding of object affordance requires learning and prediction in the 3D physical domain, which is still absent in the community. In this work, we present a 3D AffordanceNet dataset, a benchmark of 23k shapes from 23 semantic object categories, annotated with 18 visual affordance categories. Based on this dataset, we provide three benchmarking tasks for evaluating visual affordance understanding, including full-shape, partial-view and rotation-invariant affordance estimations. Three state-of-the-art point cloud deep learning networks are evaluated on all tasks. In addition we also investigate a semi-supervised learning setup to explore the possibility to benefit from unlabeled data. Comprehensive results on our contributed dataset show the promise of visual affordance understanding as a valuable yet challenging benchmark.

Comments:	CVPR2021 accepted paper
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2103.16397 [cs.CV]
	(or arXiv:2103.16397v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2103.16397

Submission history

From: Shengheng Deng [view email]
[v1] Tue, 30 Mar 2021 14:46:27 UTC (17,710 KB)
[v2] Wed, 31 Mar 2021 09:59:28 UTC (17,710 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:3D AffordanceNet: A Benchmark for Visual Object Affordance Understanding

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:3D AffordanceNet: A Benchmark for Visual Object Affordance Understanding

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators