arXiv · 1909.08950
Count, Crop and Recognise: Fine-Grained Recognition in the Wild
Abstract
The goal of this paper is to label all the animal individuals present in every frame of a video. Unlike previous methods that have principally concentrated on labelling face tracks, we aim to label individuals even when their faces are not visible. We make the following contributions: (i) we introduce a 'Count, Crop and Recognise' (CCR) multistage recognition process for frame level labelling. The Count and Recognise stages involve specialised CNNs for the task, and we show that this simple staging gives a substantial boost in performance; (ii) we compare the recall using frame based labelling to both face and body track based labelling, and demonstrate the advantage of frame based with CCR for the specified goal; (iii) we introduce a new dataset for chimpanzee recognition in the wild; and (iv) we apply a high-granularity visualisation technique to further understand the learned CNN features for the recognition of chimpanzee individuals.
Explore related subjects
Keep this discovery
Max Bain, Arsha Nagrani, Daniel Schofield, Andrew Zisserman. 2019-09-19. Count, Crop and Recognise: Fine-Grained Recognition in the Wild. https://arxiv.org/abs/1909.08950
Cite the original work for its findings. Save a collection to share your selection of sources.