Constraining Implicit Space with Minimum Description Length: An
  Unsupervised Attention Mechanism across Neural Network Layers

Lin, Baihan

research

Constraining Implicit Space with Minimum Description Length: An Unsupervised Attention Mechanism across Neural Network Layers

Authors: Baihan Lin
Publication date: 10 September 2020
Publisher
Doi

Abstract

Inspired by the adaptation phenomenon of neuronal firing, we propose the regularity normalization (RN) as an unsupervised attention mechanism (UAM) which computes the statistical regularity in the implicit space of neural networks under the Minimum Description Length (MDL) principle. Treating the neural network optimization process as a partially observable model selection problem, UAM constrains the implicit space by a normalization factor, the universal code length. We compute this universal code incrementally across neural network layers and demonstrated the flexibility to include data priors such as top-down attention and other oracle information. Empirically, our approach outperforms existing normalization methods in tackling limited, imbalanced and non-stationary input distribution in image classification, classic control, procedurally-generated reinforcement learning, generative modeling, handwriting generation and question answering tasks with various neural network architectures. Lastly, UAM tracks dependency and critical learning stages across layers and recurrent time steps of deep networks

Similar works

Full text

Available Versions

Directory of Open Access Journals

oai:doaj.org/article:b0f5dfc2d...

Last time updated on 06/04/2022

Multidisciplinary Digital Publishing Institute

oai:mdpi.com:/1099-4300/24/1/5...

Last time updated on 21/10/2022