ChatGPT and general-purpose AI count fruits in pictures surprisingly well

Mengsuwan, Konlavach; Palacio, Juan Camilo Rivera; Ryo, Masahiro

Computer Science > Computer Vision and Pattern Recognition

arXiv:2404.08515 (cs)

[Submitted on 12 Apr 2024]

Title:ChatGPT and general-purpose AI count fruits in pictures surprisingly well

Authors:Konlavach Mengsuwan, Juan Camilo Rivera Palacio, Masahiro Ryo

View PDF

Abstract:Object counting is a popular task in deep learning applications in various domains, including agriculture. A conventional deep learning approach requires a large amount of training data, often a logistic problem in a real-world application. To address this issue, we examined how well ChatGPT (GPT4V) and a general-purpose AI (foundation model for object counting, T-Rex) can count the number of fruit bodies (coffee cherries) in 100 images. The foundation model with few-shot learning outperformed the trained YOLOv8 model (R2 = 0.923 and 0.900, respectively). ChatGPT also showed some interesting potential, especially when few-shot learning with human feedback was applied (R2 = 0.360 and 0.460, respectively). Moreover, we examined the time required for implementation as a practical question. Obtaining the results with the foundation model and ChatGPT were much shorter than the YOLOv8 model (0.83 hrs, 1.75 hrs, and 161 hrs). We interpret these results as two surprises for deep learning users in applied domains: a foundation model with few-shot domain-specific learning can drastically save time and effort compared to the conventional approach, and ChatGPT can reveal a relatively good performance. Both approaches do not need coding skills, which can foster AI education and dissemination.

Comments:	12 pages, 3 figures
Subjects:	Computer Vision and Pattern Recognition (cs.CV); Image and Video Processing (eess.IV)
Cite as:	arXiv:2404.08515 [cs.CV]
	(or arXiv:2404.08515v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2404.08515

Submission history

From: Konlavach Mengsuwan [view email]
[v1] Fri, 12 Apr 2024 14:54:34 UTC (1,335 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:ChatGPT and general-purpose AI count fruits in pictures surprisingly well

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:ChatGPT and general-purpose AI count fruits in pictures surprisingly well

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators