Frosting Weights for Better Continual Training (2001.01829v1)

Published 7 Jan 2020 in cs.LG, cs.NE, and stat.ML

Abstract: Training a neural network model can be a lifelong learning process and is a computationally intensive one. A severe adverse effect that may occur in deep neural network models is that they can suffer from catastrophic forgetting during retraining on new data. To avoid such disruptions in the continuous learning, one appealing property is the additive nature of ensemble models. In this paper, we propose two generic ensemble approaches, gradient boosting and meta-learning, to solve the catastrophic forgetting problem in tuning pre-trained neural network models.

Citations (4)

View on Semantic Scholar

Summary

We haven't generated a summary for this paper yet.

Summarize Now

Frosting Weights for Better Continual Training (2001.01829v1)

Summary

Related Papers