Training products of expert capsules with mixing by dynamic routing (1907.11643v1)

Published 26 Jul 2019 in cs.LG, cs.NE, and stat.ML

Abstract: This study develops an unsupervised learning algorithm for products of expert capsules with dynamic routing. Analogous to binary-valued neurons in Restricted Boltzmann Machines, the magnitude of a squashed capsule firing takes values between zero and one, representing the probability of the capsule being on. This analogy motivates the design of an energy function for capsule networks. In order to have an efficient sampling procedure where hidden layer nodes are not connected, the energy function is made consistent with dynamic routing in the sense of the probability of a capsule firing, and inference on the capsule network is computed with the dynamic routing between capsules procedure. In order to optimize the log-likelihood of the visible layer capsules, the gradient is found in terms of this energy function. The developed unsupervised learning algorithm is used to train a capsule network on standard vision datasets, and is able to generate realistic looking images from its learned distribution.

Summary

We haven't generated a summary for this paper yet.

Summarize Now

Related Papers

Capsule networks with non-iterative cluster routing (2021)
Routing Towards Discriminative Power of Class Capsules (2021)
Training capsules as a routing-weighted product of expert neurons (2019)
Attention routing between capsules (2019)
Neural Network Encapsulation (2018)