Search papers, labs, and topics across Lattice.
This paper introduces UAVSat-Deg, a comprehensive benchmark for assessing the robustness of UAV-satellite geo-localization methods against various image degradations, including 27 corruption types across three severity levels. The authors present ReLATE, a Reliable Evidence Learning framework that adaptively fuses reliable visual evidence to construct robust cross-view descriptors, significantly improving performance under challenging conditions. Benchmark results show that ReLATE outperforms existing methods in corrupted scenarios while maintaining competitive accuracy on clean images, highlighting the critical need for robustness in real-world applications.
ReLATE achieves superior robustness in UAV-satellite geo-localization, outperforming existing methods even under severe image degradations.
Unmanned aerial vehicle (UAV)-satellite cross-view geo-localization matches UAV images against satellite imagery and has achieved impressive accuracy on clean (non-degraded) image benchmarks. In real-world flights, however, UAV observations are frequently affected by adverse weather, illumination changes, platform motion, sensor noise, and compression, while the robustness of existing methods under such degradations remains largely unexamined. In this paper, we present UAVSat-Deg, a large-scale robustness benchmark for degraded UAV-satellite geo-localization, comprising University-1652-Deg and SUES-200-Deg. UAVSat-Deg covers 27 corruption types, including 19 core and 8 compound corruptions, at three severity levels, supports bidirectional drone-to-satellite and satellite-to-drone retrieval as well as multi-height UAV acquisition, and contains more than 11.7 million pre-generated corrupted test images. Benchmarking representative methods under this protocol reveals substantial robustness gaps, particularly under severe and compound corruptions. To address this problem, we propose ReLATE, a Reliable Evidence Learning framework with Adaptive Token Evidence Regulation, which realizes reliability-adaptive feature fusion during descriptor construction. ReLATE estimates a structure-smoothed reliability field over visual tokens, aggregates trustworthy local evidence, and adaptively integrates it into query-derived representations; the regulated query representations are then combined with the CLS-token and GeM-pooled branches to form the final cross-view descriptor. Across both test sets and retrieval directions, ReLATE achieves the best average corrupted-test performance among the compared methods while maintaining competitive accuracy on clean images. The code and dataset will be available at https://github.com/JHC626/ReLATE.