Demonstrating the Efficacy of Kolmogorov-Arnold Networks in Vision Tasks

Cheon, Minjong

Computer Science > Computer Vision and Pattern Recognition

arXiv:2406.14916 (cs)

[Submitted on 21 Jun 2024]

Title:Demonstrating the Efficacy of Kolmogorov-Arnold Networks in Vision Tasks

Authors:Minjong Cheon

View PDF HTML (experimental)

Abstract:In the realm of deep learning, the Kolmogorov-Arnold Network (KAN) has emerged as a potential alternative to multilayer projections (MLPs). However, its applicability to vision tasks has not been extensively validated. In our study, we demonstrated the effectiveness of KAN for vision tasks through multiple trials on the MNIST, CIFAR10, and CIFAR100 datasets, using a training batch size of 32. Our results showed that while KAN outperformed the original MLP-Mixer on CIFAR10 and CIFAR100, it performed slightly worse than the state-of-the-art ResNet-18. These findings suggest that KAN holds significant promise for vision tasks, and further modifications could enhance its performance in future evaluations.Our contributions are threefold: first, we showcase the efficiency of KAN-based algorithms for visual tasks; second, we provide extensive empirical assessments across various vision benchmarks, comparing KAN's performance with MLP-Mixer, CNNs, and Vision Transformers (ViT); and third, we pioneer the use of natural KAN layers in visual tasks, addressing a gap in previous research. This paper lays the foundation for future studies on KANs, highlighting their potential as a reliable alternative for image classification tasks.

Subjects:	Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
Cite as:	arXiv:2406.14916 [cs.CV]
	(or arXiv:2406.14916v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2406.14916

Submission history

From: Minjong Cheon [view email]
[v1] Fri, 21 Jun 2024 07:20:34 UTC (13 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Demonstrating the Efficacy of Kolmogorov-Arnold Networks in Vision Tasks

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Demonstrating the Efficacy of Kolmogorov-Arnold Networks in Vision Tasks

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators