Self-augmentation: Generalizing deep networks to unseen classes for few-shot learning

Citations

WEB OF SCIENCE

42
Citations

SCOPUS

47

초록

Few-shot learning aims to classify unseen classes with a few training examples. While recent works have shown that standard mini-batch training with carefully designed training strategies can improve generalization ability for unseen classes, well-known problems in deep networks such as memorizing training statistics have been less explored for few-shot learning. To tackle this issue, we propose self-augmentation that consolidates self-mix and self-distillation. Specifically, we propose a regional dropout technique called self-mix, in which a patch of an image is substituted into other values in the same image. With this dropout effect, we show that the generalization ability of deep networks can be improved as it prevents us from learning specific structures of a dataset. Then, we employ a backbone network that has auxiliary branches with its own classifier to enforce knowledge sharing. This sharing of knowledge forces each branch to learn diverse optimal points during training. Additionally, we present a local representation learner to further exploit a few training examples of unseen classes by generating fake queries and novel weights. Experimental results show that the proposed method outperforms the state-of-the-art methods for prevalent few-shot benchmarks and improves the generalization ability. © 2021 Elsevier Ltd

키워드

ClassificationFew-shot learningGeneralizationKnowledge distillationDistillationBack-bone networkGeneralization abilityKnowledge-sharingOptimal pointsState-of-the-art methodsTraining exampleTraining strategyDeep learningarticleclassifierdistillationhumanlearning
제목
Self-augmentation: Generalizing deep networks to unseen classes for few-shot learning
저자
Seo, J.-W.Jung, H.-G.Lee, S.-W.
DOI
10.1016/j.neunet.2021.02.007
발행일
2021-06
유형
Article
저널명
Neural Networks
138
페이지
140 ~ 149