Graph explorer

Entangled Residual Mappings

Residual mappings have been shown to perform representation learning in the first layers and iterative feature refinement in higher layers. This interplay, combined with their stabilizing effect on the gradient norms, enables them to train very deep networks. In this paper, we take a step further and introduce entangled residual mappings to generalize the structure of the residual connections and evaluate their role in iterative learning representations. An entangled residual mapping replaces the identity skip connections with specialized entangled mappings such as orthogonal, sparse, and structural correlation matrices that share key attributes (eigenvalues, structure, and Jacobian norm) with identity mappings. We show that while entangled mappings can preserve the iterative refinement of features across various deep models, they influence the representation learning process in convolutional networks differently than attention-based models and recurrent neural networks. In general, we find that for CNNs and Vision Transformers entangled sparse mapping can help generalization while orthogonal mappings hurt performance. For recurrent networks, orthogonal residual mappings form an inductive bias for time-variant sequences, which degrades accuracy on time-invariant tasks.

11 nodes23 linksoverview previewEntangled Residual Mappings
11 nodes23 links
Entangled Residual Mappings11 visible / 11 total nodes / 41 links
Related contextRelated contextRelated contextCo-authorshipCo-authorshipCo-authorshipWorks onWorks onCo-authorshipCo-authorshipCo-authorshipCo-authorshipCo-authorshipCo-authorshipCo-authorshipCo-authorshipCo-authorshipCo-authorshipCo-authorshipCo-authorshipCo-authorshipCo-authorshipCo-authorshipCo-authorshipCo-authorshipCo-authorshipAuthorshipWorks onWorks onWorks onWorks onWorks onAuthorshipAuthorshipAuthorshipTopic signalTopic signalTopic signalAuthorshipAuthorshipAuthorshipWEntangled Residual Mappingspreprint / 2022AMathias LechnerResearcherARamin HasaniResearcherAZahra BabaieeResearcherARadu GrosuResearcherTMachine Learning49008 worksTArtificial Intelligence22915 worksTNeural and Evolutionary...2839 worksADaniela RusResearcherAThomas A. HenzingerResearcherASepp HochreiterResearcher
PaperSignal 1010 links

Entangled Residual Mappings

preprint / 2022

Open