arXiv (Cornell University) · 2018 · 199 citations · 25 references
Few-shot LearningMachine VisionMachine LearningData ScienceMetric ScalingPattern RecognitionEngineeringCertain MetricsZero-shot LearningFeature LearningMulti-task LearningComputer ScienceRobot LearningDeep LearningImproved Few-shot LearningComputer Vision
Few‑shot learning is essential for building models that generalize from very limited data. The authors aim to enhance few‑shot learning by demonstrating the importance of metric scaling and task‑dependent conditioning, and by proposing a task‑dependent metric space. They introduce a simple conditioning method that learns a task‑dependent metric space and an end‑to‑end optimization procedure using auxiliary task co‑training. Metric scaling changes parameter updates, boosts mini‑Imagenet 5‑way 5‑shot accuracy by up to 14%, achieves state‑of‑the‑art results, and the gains transfer to a new CIFAR‑100 few‑shot dataset, with the code publicly available.
Few-shot learning has become essential for producing models that generalize from few examples. In this work, we identify that metric scaling and metric task conditioning are important to improve the performance of few-shot algorithms. Our analysis reveals that simple metric scaling completely changes the nature of few-shot algorithm parameter updates. Metric scaling provides improvements up to 14% in accuracy for certain metrics on the mini-Imagenet 5-way 5-shot classification task. We further propose a simple and effective way of conditioning a learner on the task sample set, resulting in learning a task-dependent metric space. Moreover, we propose and empirically test a practical end-to-end optimization procedure based on auxiliary task co-training to learn a task-dependent metric space. The resulting few-shot learning model based on the task-dependent scaled metric achieves state of the art on mini-Imagenet. We confirm these results on another few-shot dataset that we introduce in this paper based on CIFAR100. Our code is publicly available at https://github.com/ElementAI/TADAM.
25
Deep Residual Learning for Image Recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren et al. · 2016 · 214.9K citations · Full text
Image Classification, Deep Neural Networks, Machine Vision +14
Distilling the Knowledge in a Neural Network
Geoffrey E. Hinton, Oriol Vinyals · arXiv (Cornell University) · 2015 · 13.9K citations · Full text
Prototypical Networks for Few-shot Learning
Jake Snell, Kevin Swersky, Richard S. Zemel · arXiv (Cornell University) · 2017 · 5.2K citations · Full text
Dimensionality Reduction by Learning an Invariant Mapping
Raia Hadsell, Sumit Chopra, Yann LeCun · 2006 · 5.1K citations · Full text