Large well building is important remote sensing object, and the research on object detection of large well buildings is of great significance to national defense. At the data level, due to the small number of large well buildings samples, there is currently no valid data set available for the object detection. Building effective datasets is of great value for the research in related fields. At the algorithm level,the different resolutions of the remote sensing images result in multi-scale characteristics of the large well buildings, which is one of the difficulties in solving the object detection problem. Based on the above analysis, firstly, this article built the first large well buildings object detection dataset using Google Earth. Then an effective detection model was designed for large well building object detection task. The model in this paper fully integrates the object's multi-scale features and contextual information, and detects the object through the multi-stage cascade network. The model can effectively detect large well buildings, and the detection effect is better than the results of the current mainstream algorithms.
[1] 凯迪网网络. 美国洲际弹道导弹发射阵地坐标全部公开[EB/OL]. (2017-09-27)[2020-03-18]. http://m.kd-net.net/share-12432513.html.
[2] 铁血社区. 网友曝光美国战略导弹发射井精确位置[EB/OL]. (2014-01-10)[2020-03-18]. https://bbs.tiex-ue.net/post_7039175_1.html.
[3] 铁血社区. 神秘的美国导弹发射井[EB/OL]. (2015-05-16)[2020-03-18]. https://bbs.tiexue.net/post_877-7935_1.html.
[4] Lowe D G. Distinctive image features from scale-invariant keypoints[J]. International Journal of Computer Vision, 2004, 60(2):91-110.
[5] Dalal N, Triggs B. Histograms of oriented gradients for human detection[C]//2005 IEEE Computer Society Conference on Computer Vision and Pattern Recognition (CVPR'05). June 20-25,2005, San Diego, CA, USA. IEEE, 2005:886-893.
[6] Ojala T, Pietikainen M, Maenpaa T. Multiresolution gray-scale and rotation invariant texture classification with local binary patterns[J]. IEEE Transactions on Pattern Analysis and Machine Intelligence, 2002, 24(7):971-987.
[7] Lienhart R, Maydt J. An extended set of Haar-like features for rapid object detection[C]//Proceedings. International Conference on Image Processing. September 22-25, 2002, Rochester NY, USA. IEEE, 2002(I):900-903.
[8] Cherkassky V. The nature of statistical learning theory[J]. IEEE Transactions on Neural Networks, 1997, 8(6):1564.
[9] Freund Y, Schapire R E. Experiments with a new boosting algorithm[C]//International Conference on Machine Learning. Bari:ACM, 1996:148-156.
[10] Girshick R, Donahue J, Darrell T, et al. Rich feature hierarchies for accurate object detection and semantic segmentation[C]//2014 IEEE Conference on Computer Vision and Pattern Recognition. June 23-28, 2014, Columbus, OH, USA. IEEE, 2014:580-587.
[11] He K M, Zhang X Y, Ren S Q, et al. Spatial pyramid pooling in deep convolutional networks for visual recognition[J]. IEEE Transactions on Pattern Analysis and Machine Intelligence, 2015, 37(9):1904-1916.
[12] Girshick R. Fast R-CNN[C]//2015 IEEE International Conference on Computer Vision(ICCV). December 7-13,2015, Santiago, Chile. IEEE, 2015:1440-1448.
[13] Liu W, Anguelov D, Erhan D, et al. SSD:single shot multibox detector[C]//European Conference on Computer Vision. Amsterdam:Springer, 2016:21-37.
[14] Shrivastava A, Sukthankar R, Malik J, et al. Beyond skip connections:top-down modulation for object detection[EB/OL]. arXiv:1612.06851v1[cs.CV]. (2016-12-20)[2020-02-20]. https://www.researchgate.net/publication/311769895.
[15] Dai J F, Li Y, He K M, et al. R-FCN:object detection via region-based fully convolutional networks[EB/OL]. arXiv:1605.06409[cs.CV].(2016-06-21)[2020-02-20]. https://arxiv.org/abs/1605.06409.
[16] Kong T, Sun F C, Yao A B, et al. Ron:Reverse connection with objectness prior networks for object detection[C]//2017 IEEE Conference on Computer Vision and Pattern Recognition(CVRR). July 21-26,2017, Honolulu, HI, USA. IEEE, 2017:5936-5944.
[17] Ren S Q, He K M, Girshick R, et al. Faster R-CNN:towards real-time object detection with region proposal networks[C]//IEEE Transactions on Pattern Analysis and Machine Intelligence. IEEE,1137-1149.
[18] Lin T Y, Dolla'r P, Girshick R, et al. Feature pyramid networks for object detection[C]//2017 IEEE Conference on Computer Vision and Pattern Recognition(CVPR). July 21-26, 2017, Honolulu, HI, USA. IEEE, 2017:936-944.
[19] Redmon J, Divvala S, Girshick R, et al. You only look once:Unified, real-time object detection[C]//2016 IEEE Conference on Computer Vision and Pattern Recognition(CVPR). June 27-30, 2016, Las Vegas, NV, USA. IEEE, 2016:779-788.
[20] Cai Z W, Vasconcelos N. Cascade R-CNN:delving into high quality object detection[C]//2018 IEEE/CVF Conference on Computer Vision and Pattern Recognition. June 18-23, 2018, Salt Lake City, UT, USA. IEEE, 2018:6154-6162.
[21] Han X B, Zhong Y F, Zhang L P. An efficient and robust integrated geospatial object detection framework for high spatial resolution remote sensing imagery[J]. Remote Sensing, 2017, 9(7):666.
[22] Xu Z Z, Xu X, Wang L, et al. Deformable ConvNet with aspect ratio constrained NMS for object detection in remote sensing imagery[J]. Remote Sensing, 2017, 9(12):1312.
[23] Dai J F, Qi H Z, Xiong Y W, et al. Deformable Convolutional Networks[C]//2017 IEEE International Conference on Computer Vision(ICCV). October 22-29, 2017, Venice, Italy. IEEE, 2017:764-773.
[24] Ren Y, Zhu C R, Xiao S P. Deformable faster R-CNN with aggregating multi-layer features for partially occluded object detection in optical remote sensing images[J]. Remote Sensing, 2018, 10(9):1470.
[25] Everingham M, Van Gool L, Williams C K I, et al. The pascal visual object classes (VOC) challenge[J]. International Journal of Computer Vision, 2010, 88(2):303-338.
[26] Lin T Y, Goyal P, Girshick R, et al. Focal loss for dense object detection[C]//2017 IEEE International Conference on Computer Vision(ICCV). October 22-29, 2017, Venice, Italy. IEEE, 2017:2999-3007.
[27] He K M, Zhang X Y, Ren S Q, et al. Deep residual learning for image recognition[C]//2016 IEEE Conference on Computer Vision and Pattern Recognition(CVPR). June 27-30, 2016, Las Vegas, NV, USA. IEEE, 2016:770-778.
[28] Azimi S M, Vig E, Bahmanyar R, et al. Towards multi-class object detection in unconstrained remote sensing imagery[C]//Asian Conference on Computer Vision. Perth:Springer, 2018:150-165.
[29] Xia G S, Bai X, Ding J, et al. DOTA:a large-scale dataset for object detection in aerial images[C]//2018 IEEE/CVF Conference on Computer Vision and Pattern Recognition. June 18-23, Salt Lake City, UT, UAS. IEEE, 2018:3974-3983.