[1711.08189] An Analysis of Scale Invariance in Object Detection - SNIP
An analysis of different techniques for recognizing and detecting objects under extreme scale variation is presented. Scale specific and scale invariant design of detectors are compared by training them with different configurations of input data. By evaluating the performance of different network architectures for classifying small objects on ImageNet, we show that CNNs are not robust to changes in scale. Based on this analysis, we propose to train and test detectors on the same scales of an image-pyramid. Since small and large objects are difficult to recognize at smaller and larger scales respectively, we present a novel training scheme called Scale Normalization for Image Pyramids (SNIP) which selectively back-propagates the gradients of object instances of different sizes as a function of the image scale. On the COCO dataset, our single model performance is 45.7% and an ensemble of 3 networks obtains an mAP of 48.3%. We use off-the-shelf ImageNet-1000 pre-trained models and only train with bounding box supervision. Our submission won the Best Student Entry in the COCO 2017 challenge. Code will be made available at \url{this http URL}.
An Analysis of Scale Invariance in Object Detection – SNIP Bharat Singh Larry S. Davis University of Maryland, College Park {bharat,lsd}@cs.umd.edu arXiv:1711.08189v2 [cs.CV] 25 May 2018 Abstract An analysis of…
saved by
related reading
- An Analysis of Scale Invariance in Object Detection SNIP - Singh_An_Analysis_of_CVPR_2018_paper.pdfopenaccess.thecvf.com
- Locally Scale-Invariant Convolutional NeuralNetworksarxiv.org
- Scale-Invariant CNNarxiv.org
- Scale-invariant-CNNs/docs/Report/Implementation.md at master · ruoqizzz/Scale-invariant-CNNsgithub.com
- Scale-Invariant Recognition by Weight-Shared CNNs in Parallelproceedings.mlr.press
- [1810.13128] The Effect of Learning Strategy versus Inherent Architecture Properties on the Ability of Convolutional Neural Networks to Develop Transformation Invariancearxiv.org
- GitHub - xingyizhou/ExtremeNet: Bottom-up Object Detection by Grouping Extreme and Center Pointsgithub.com
- Meauring Invariances in Deep Networksai.stanford.edu
- GitHub - david8862/keras-YOLOv3-model-set: end-to-end YOLOv4/v3/v2 object detection pipeline, implemented on tf.keras with different technologiesgithub.com
- [1905.11946] EfficientNet: Rethinking Model Scaling for Convolutional Neural Networksarxiv.org
- ImageNet Classification with Deep Convolutional Neural Networksproceedings.neurips.cc
- Aman's AI Journal • Primers • Ilya Sutskever's Top 30aman.ai