Affine-Transformation-Invariant Image Classification by Differentiable Arithmetic Distribution Module

Tan, Zijie; Dong, Guanfang; Zhao, Chenqiu; Basu, Anup

Computer Science > Computer Vision and Pattern Recognition

arXiv:2309.00752 (cs)

[Submitted on 1 Sep 2023 (v1), last revised 12 Dec 2023 (this version, v2)]

Title:Affine-Transformation-Invariant Image Classification by Differentiable Arithmetic Distribution Module

Authors:Zijie Tan, Guanfang Dong, Chenqiu Zhao, Anup Basu

View PDF HTML (experimental)

Abstract:Although Convolutional Neural Networks (CNNs) have achieved promising results in image classification, they still are vulnerable to affine transformations including rotation, translation, flip and shuffle. The drawback motivates us to design a module which can alleviate the impact from different affine transformations. Thus, in this work, we introduce a more robust substitute by incorporating distribution learning techniques, focusing particularly on learning the spatial distribution information of pixels in images. To rectify the issue of non-differentiability of prior distribution learning methods that rely on traditional histograms, we adopt the Kernel Density Estimation (KDE) to formulate differentiable histograms. On this foundation, we present a novel Differentiable Arithmetic Distribution Module (DADM), which is designed to extract the intrinsic probability distributions from images. The proposed approach is able to enhance the model's robustness to affine transformations without sacrificing its feature extraction capabilities, thus bridging the gap between traditional CNNs and distribution-based learning. We validate the effectiveness of the proposed approach through ablation study and comparative experiments with LeNet.

Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2309.00752 [cs.CV]
	(or arXiv:2309.00752v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2309.00752

Submission history

From: Guanfang Dong [view email]
[v1] Fri, 1 Sep 2023 22:31:32 UTC (1,482 KB)
[v2] Tue, 12 Dec 2023 20:10:04 UTC (416 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Affine-Transformation-Invariant Image Classification by Differentiable Arithmetic Distribution Module

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Affine-Transformation-Invariant Image Classification by Differentiable Arithmetic Distribution Module

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators