Cross-view image geo-localization (CVGL) is a key technique for achieving geographic location matching through multi-source images, with broad applications in fields such as robot navigation and autonomous driving. However, existing methods typically assume that images are directionally aligned, making them less effective in real-world scenarios with directional deviations. To address this challenge, this paper proposes a ranking aggregation-based method for enhancing directional robustness. By introducing random directional shifts to ground-level images and employing mean and min aggregation strategies, the proposed method improves the matching accuracy and robustness. Experiments conducted on the CVUSA_360 and CVACT_360 datasets demonstrate that this approach significantly enhances the performance of various models under directionally misaligned conditions, with particularly outstanding results in terms of the R@1 metric. Visual analysis further verifies the advantages of the method in strengthening directional robustness, suppressing noise, and improving matching accuracy, thus providing an efficient solution for geo-localization tasks in complex scenarios.
SHENG Yining
,
ZHAO Lijun
,
ZHANG Zheng
,
TANG Ping
. A ranking aggregation-based method for enhancing directional robustness in cross-view image geo-localization[J]. Journal of University of Chinese Academy of Sciences, 0
: 7
.
DOI: 10.7523/j.ucas.2025.019
[1] Lin T Y, Belongie S, Hays J.Cross-view image geolocalization[C]//2013 IEEE Conference on Computer Vision and Pattern Recognition. June 23-28, 2013, Portland, OR, USA. IEEE, 2013: 891-898. DOI: 10.1109/CVPR.2013.120.
[2] Downes L M, Kim D K, Steiner T J, et al.City-wide street-to-satellite image geolocalization of a mobile ground agent[C]//2022 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS). October 23-27, 2022, Kyoto, Japan. IEEE, 2022: 11102-11108. DOI: 10.1109/IROS47612.2022.9981996.
[3] McManus C, Churchill W, Maddern W, et al. Shady dealings: robust, long-term visual localisation using illumination invariance[C]//2014 IEEE International Conference on Robotics and Automation (ICRA). May 31 - June 7, 2014, Hong Kong, China. IEEE, 2014: 901-906. DOI: 10.1109/ICRA.2014.6906961.
[4] Häne C, Heng L, Lee G H, et al.3D visual perception for self-driving cars using a multi-camera system: calibration, mapping, localization, and obstacle detection[J]. Image and Vision Computing, 2017, 68: 14-27. DOI: 10.1016/j.imavis.2017.07.003.
[5] Middelberg S, Sattler T, Untzelmann O, et al.Scalable 6-DOF localization on mobile devices[C]// Computer Vision - ECCV 2014. Cham: Springer International Publishing, 2014: 268-283. DOI: 10.1007/978-3-319-10605-2_18.
[6] Workman S, Souvenir R, Jacobs N.Wide-area image geolocalization with aerial reference imagery[C]//2015 IEEE International Conference on Computer Vision (ICCV). December 7-13, 2015, Santiago, Chile. IEEE, 2015: 3961-3969. DOI: 10.1109/ICCV.2015.451.
[7] Liu L, Li H D.Lending orientation to neural networks for cross-view geo-localization[C]//2019 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR). June 15-20, 2019, Long Beach, CA, USA. IEEE, 2019: 5617-5626. DOI: 10.1109/CVPR.2019.00577.
[8] Zhu S J, Shah M, Chen C.TransGeo: transformer is all you need for cross-view image geo-localization[C]//2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR). June 18-24, 2022, New Orleans, LA, USA. IEEE, 2022: 1152-1161. DOI: 10.1109/CVPR52688.2022.00123.
[9] Zhu Y Y, Yang H J, Lu Y X, et al. Simple, effective and general: a new backbone for cross-view image geo-localization[EB/OL]. arXiv:2302.01572.(2023-02-03)[ 2025-01-17] https://arxiv.org/abs/2302.01572v1.
[10] Deuser F, Habel K, Oswald N.Sample4Geo: hard negative sampling for cross-view geo-localisation[C]//2023 IEEE/CVF International Conference on Computer Vision (ICCV). October 1-6, 2023, Paris, France. IEEE, 2023: 16801-16810. DOI: 10.1109/ICCV51070.2023.01545.
[11] Zhang X H, Li X Y, Sultani W, et al.GeoDTR+: toward generic cross-view geolocalization via geometric disentanglement[J]. IEEE Transactions on Pattern Analysis and Machine Intelligence, 2024, 46(12): 10419-10433. DOI: 10.1109/TPAMI.2024.3443652.
[12] Mi L, Xu C, Castillo-Navarro J, et al.ConGeo: robust cross-view geo-localization across ground view variations[C]// Computer Vision - ECCV 2024. Cham: Springer Nature Switzerland, 2025: 214-230. DOI: 10.1007/978-3-031-72630-9_13.
[13] Zhang X H, Li X Y, Sultani W, et al.Cross-view geo-localization via learning disentangled geometric layout correspondence[J]. Proceedings of the AAAI Conference on Artificial Intelligence, 2023, 37(3): 3480-3488. DOI: 10.1609/aaai.v37i3.25457.