Scinovex
articleTop 1% cited

One-shot learning of object categories

Li Fei-FeiRob FergusPietro Perona

Abstract

Learning visual models of object categories notoriously requires hundreds or thousands of training examples. We show that it is possible to learn much information about a category from just one, or a handful, of images. The key insight is that, rather than learning from scratch, one can take advantage of knowledge coming from previously learned categories, no matter how different these categories might be. We explore a Bayesian implementation of this idea. Object categories are represented by probabilistic models. Prior knowledge is represented as a probability density function on the parameters of these models. The posterior model for an object category is obtained by updating the prior in the light of one or more observations. We test a simple implementation of our algorithm on a database of 101 diverse object categories. We compare category models learned by an implementation of our Bayesian approach to models learned from by Maximum Likelihood (ML) and Maximum A Posteriori (MAP) methods. We find that on a database of more than 100 categories, the Bayesian approach produces informative models when the number of training examples is too small for other methods to operate successfully.

Domain Adaptation and Few-Shot LearningAdvanced Image and Video Retrieval TechniquesImage Retrieval and Classification TechniquesArtificial intelligenceComputer scienceObject (grammar)Machine learningBayesian probabilityMaximum a posteriori estimationProbabilistic logicCognitive neuroscience of visual object recognitionPattern recognition (psychology)Mathematics

MeSH terms

AlgorithmsArtificial IntelligenceBayes TheoremComputer SimulationImage EnhancementImage Interpretation, Computer-AssistedModels, BiologicalPattern Recognition, AutomatedSensitivity and SpecificityReproducibility of ResultsModels, StatisticalCluster AnalysisInformation Storage and RetrievalImaging, Three-Dimensional
Citations
3,040
FWCI
67.96
field-weighted impact
References
52
Percentile
100%
vs. same field & year
Citations per year
Cited by
Deep Learning for Generic Object Detection: A Survey
International Journal of Computer Vision · 2019 · 2,702 citations
LabelMe: A Database and Web-Based Tool for Image Annotation
International Journal of Computer Vision · 2007 · 4,112 citations
A Database and Evaluation Methodology for Optical Flow
International Journal of Computer Vision · 2010 · 2,210 citations
Deep Multitask Learning for Railway Track Inspection
IEEE Transactions on Intelligent Transportation Systems · 2016 · 425 citations
Generalizing from a Few Examples
ACM Computing Surveys · 2020 · 2,549 citations
Pedestrian Detection: An Evaluation of the State of the Art
IEEE Transactions on Pattern Analysis and Machine Intelligence · 2011 · 3,229 citations
References
Saliency, Scale and Image Description
International Journal of Computer Vision · 2001 · 1,197 citations
Detecting Pedestrians Using Patterns of Motion and Appearance
International Journal of Computer Vision · 2005 · 2,103 citations
Pictorial Structures for Object Recognition
International Journal of Computer Vision · 2004 · 2,188 citations
Maximum Likelihood from Incomplete Data Via the <i>EM</i> Algorithm
Journal of the Royal Statistical Society Series B (Statistical Methodology) · 1977 · 49,286 citations
Gradient-based learning applied to document recognition
Proceedings of the IEEE · 1998 · 57,014 citations
Recognition-by-components: A theory of human image understanding.
Psychological Review · 1987 · 5,528 citations
Comparing images using the Hausdorff distance
IEEE Transactions on Pattern Analysis and Machine Intelligence · 1993 · 4,396 citations
Neural network-based face detection
IEEE Transactions on Pattern Analysis and Machine Intelligence · 1998 · 3,513 citations
Citation Network

How this paper connects to the literature. Drag to explore, click any node to open that paper.

One-shot learning of object categories · Scinovex