Residual Connections Encourage Iterative Inference

Jastrzębski, Stanisław; Arpit, Devansh; Ballas, Nicolas; Verma, Vikas; Che, Tong; Bengio, Yoshua

Computer Science > Computer Vision and Pattern Recognition

arXiv:1710.04773 (cs)

[Submitted on 13 Oct 2017 (v1), last revised 8 Mar 2018 (this version, v2)]

Title:Residual Connections Encourage Iterative Inference

Authors:Stanisław Jastrzębski, Devansh Arpit, Nicolas Ballas, Vikas Verma, Tong Che, Yoshua Bengio

View PDF

Abstract:Residual networks (Resnets) have become a prominent architecture in deep learning. However, a comprehensive understanding of Resnets is still a topic of ongoing research.
A recent view argues that Resnets perform iterative refinement of features. We attempt to further expose properties of this aspect. To this end, we study Resnets both analytically and empirically. We formalize the notion of iterative refinement in Resnets by showing that residual connections naturally encourage features of residual blocks to move along the negative gradient of loss as we go from one block to the next. In addition, our empirical analysis suggests that Resnets are able to perform both representation learning and iterative refinement. In general, a Resnet block tends to concentrate representation learning behavior in the first few layers while higher layers perform iterative refinement of features. Finally we observe that sharing residual layers naively leads to representation explosion and counterintuitively, overfitting, and we show that simple existing strategies can help alleviating this problem.

Comments:	First two authors contributed equally. Published in ICLR 2018
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:1710.04773 [cs.CV]
	(or arXiv:1710.04773v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.1710.04773

Submission history

From: Stanisław Jastrzębski [view email]
[v1] Fri, 13 Oct 2017 01:39:32 UTC (464 KB)
[v2] Thu, 8 Mar 2018 18:45:27 UTC (872 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Residual Connections Encourage Iterative Inference

Submission history

Access Paper:

References & Citations

1 blog link

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Residual Connections Encourage Iterative Inference

Submission history

Access Paper:

References & Citations

1 blog link

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators