Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Bayesian Nonparametric Weight Factorization for Continual Learning

Apr 21, 2020

Nikhil Mehta, Kevin J Liang, Lawrence Carin

Figure 1 for Bayesian Nonparametric Weight Factorization for Continual Learning

Figure 2 for Bayesian Nonparametric Weight Factorization for Continual Learning

Figure 3 for Bayesian Nonparametric Weight Factorization for Continual Learning

Figure 4 for Bayesian Nonparametric Weight Factorization for Continual Learning

Share this with someone who'll enjoy it:

Abstract:Naively trained neural networks tend to experience catastrophic forgetting in sequential task settings, where data from previous tasks are unavailable. A number of methods, using various model expansion strategies, have been proposed recently as possible solutions. However, determining how much to expand the model is left to the practitioner, and typically a constant schedule is chosen for simplicity, regardless of how complex the incoming task is. Instead, we propose a principled Bayesian nonparametric approach based on the Indian Buffet Process (IBP) prior, letting the data determine how much to expand the model complexity. We pair this with a factorization of the neural network's weight matrices. Such an approach allows us to scale the number of factors of each weight matrix to the complexity of the task, while the IBP prior imposes weight factor sparsity and encourages factor reuse, promoting positive knowledge transfer between tasks. We demonstrate the effectiveness of our method on a number of continual learning benchmarks and analyze how weight factors are allocated and reused throughout the training.

View paper on

Share this with someone who'll enjoy it:

Title:Bayesian Nonparametric Weight Factorization for Continual Learning

Paper and Code