In plain words: Rather than training a network on one task at a time, they cycle it through tasks, shaping a starting point that helps with the next one. This teaches the network how to learn and keeps old skills from being overwritten better than usual single-task training.
Abstract
Current training regimes for deep learning usually involve exposure to a single task / dataset at a time. Here we start from the observation that in this context the trained model is not given any knowledge of anything outside its (single-task) training distribution, and has thus no way to learn parameters (i.e., feature detectors or policies) that could be helpful to solve other tasks, and to limit future interference with the acquired knowledge, and thus catastrophic forgetting. Here we show that catastrophic forgetting can be mitigated in a meta-learning context, by exposing a neural network to multiple tasks in a sequential manner during training. Finally, we present SeqFOMAML, a meta-learning algorithm that implements these principles, and we evaluate it on sequential learning problems composed by Omniglot and MiniImageNet classification tasks.
Giacomo Spigler
arXiv:1909.04170 · cs.AI, cs.LG, stat.ML · submitted Sep 9, 2019 · updated Feb 9, 2020
abstract · pdf · html