Look Twice: A Generalist Computational Model Predicts Return Fixations across Tasks and Species

Zhang, Mengmi; Armendariz, Marcelo; Xiao, Will; Rose, Olivia; Bendtz, Katarina; Livingstone, Margaret; Ponce, Carlos; Kreiman, Gabriel

doi:10.1371/journal.pcbi.1010654

Computer Science > Computer Vision and Pattern Recognition

arXiv:2101.01611 (cs)

[Submitted on 5 Jan 2021 (v1), last revised 14 Oct 2022 (this version, v2)]

Title:Look Twice: A Generalist Computational Model Predicts Return Fixations across Tasks and Species

Authors:Mengmi Zhang, Marcelo Armendariz, Will Xiao, Olivia Rose, Katarina Bendtz, Margaret Livingstone, Carlos Ponce, Gabriel Kreiman

View PDF

Abstract:Primates constantly explore their surroundings via saccadic eye movements that bring different parts of an image into high resolution. In addition to exploring new regions in the visual field, primates also make frequent return fixations, revisiting previously foveated locations. We systematically studied a total of 44,328 return fixations out of 217,440 fixations. Return fixations were ubiquitous across different behavioral tasks, in monkeys and humans, both when subjects viewed static images and when subjects performed natural behaviors. Return fixations locations were consistent across subjects, tended to occur within short temporal offsets, and typically followed a 180-degree turn in saccadic direction. To understand the origin of return fixations, we propose a proof-of-principle, biologically-inspired and image-computable neural network model. The model combines five key modules: an image feature extractor, bottom-up saliency cues, task-relevant visual features, finite inhibition-of-return, and saccade size constraints. Even though there are no free parameters that are fine-tuned for each specific task, species, or condition, the model produces fixation sequences resembling the universal properties of return fixations. These results provide initial steps towards a mechanistic understanding of the trade-off between rapid foveal recognition and the need to scrutinize previous fixation locations.

Comments:	9 main figs and 24 supp figs, accepted in PLOS Computational Biology
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2101.01611 [cs.CV]
	(or arXiv:2101.01611v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2101.01611
Related DOI:	https://doi.org/10.1371/journal.pcbi.1010654

Submission history

From: Mengmi Zhang [view email]
[v1] Tue, 5 Jan 2021 15:53:39 UTC (6,852 KB)
[v2] Fri, 14 Oct 2022 12:26:50 UTC (5,170 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Look Twice: A Generalist Computational Model Predicts Return Fixations across Tasks and Species

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Look Twice: A Generalist Computational Model Predicts Return Fixations across Tasks and Species

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators