MTLSegFormer: Multi-task Learning with Transformers for Semantic Segmentation in Precision Agriculture

Goncalves, Diogo Nunes; Junior, Jose Marcato; Zamboni, Pedro; Pistori, Hemerson; Li, Jonathan; Nogueira, Keiller; Goncalves, Wesley Nunes

Computer Science > Computer Vision and Pattern Recognition

arXiv:2305.02813 (cs)

[Submitted on 4 May 2023]

Title:MTLSegFormer: Multi-task Learning with Transformers for Semantic Segmentation in Precision Agriculture

Authors:Diogo Nunes Goncalves, Jose Marcato Junior, Pedro Zamboni, Hemerson Pistori, Jonathan Li, Keiller Nogueira, Wesley Nunes Goncalves

View PDF

Abstract:Multi-task learning has proven to be effective in improving the performance of correlated tasks. Most of the existing methods use a backbone to extract initial features with independent branches for each task, and the exchange of information between the branches usually occurs through the concatenation or sum of the feature maps of the branches. However, this type of information exchange does not directly consider the local characteristics of the image nor the level of importance or correlation between the tasks. In this paper, we propose a semantic segmentation method, MTLSegFormer, which combines multi-task learning and attention mechanisms. After the backbone feature extraction, two feature maps are learned for each task. The first map is proposed to learn features related to its task, while the second map is obtained by applying learned visual attention to locally re-weigh the feature maps of the other tasks. In this way, weights are assigned to local regions of the image of other tasks that have greater importance for the specific task. Finally, the two maps are combined and used to solve a task. We tested the performance in two challenging problems with correlated tasks and observed a significant improvement in accuracy, mainly in tasks with high dependence on the others.

Comments:	Accepted 4th Agriculture-Vision Workshop - CVPRW
Subjects:	Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI)
Cite as:	arXiv:2305.02813 [cs.CV]
	(or arXiv:2305.02813v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2305.02813

Submission history

From: Keiller Nogueira [view email]
[v1] Thu, 4 May 2023 13:19:43 UTC (9,346 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:MTLSegFormer: Multi-task Learning with Transformers for Semantic Segmentation in Precision Agriculture

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:MTLSegFormer: Multi-task Learning with Transformers for Semantic Segmentation in Precision Agriculture

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators