Skip to content
Review Open access

GENCP: GAN-Based Ground Control Point Generation for Satellite Image Georeferencing

Jul 2026 · Remote Sensing · Vol 18, pp. 2356 · 0 citations · 27 references

Abstract

Remote sensing has become a core technology for environmental and climate monitoring, supported by expanding sensor constellations, advanced processing capabilities, and coordination frameworks established by the European Space Agency (ESA), Global Earth Observation System of Systems (GEOSS), and Committee on Earth Observation Satellites (CEOS). Ensuring consistency across missions requires robust geometric and radiometric calibration and validation. However, traditional reliance on ground control points (GCPs) is limited by sparse global coverage, temporal instability, and dependence on surveyed accuracy. While alternative geospatial datasets, including satellite and aerial imagery, Light Detection and Ranging (LiDAR) point clouds, and vector databases, can serve as references, challenges remain in data access, automation, and cross-sensor applicability. This study proposes a generative adversarial network (GAN)-based approach to generate geometrically consistent image chips from vector maps. Two models were trained at 50 cm and 10 m resolution within the ESA-supported Generative Ground Control Point (GenCP) study, using Sentinel-2 and very-high-resolution RGB imagery. The generated GenCP image chips are evaluated using image similarity (radiometric consistency), as well as geometric and model-performance metrics. The results demonstrate their suitability for automated Cal/Val workflows and their potential as scalable, fit-for-purpose reference datasets.

Read PDF

Similar papers

Review Open access Jul 2026

Bundle-Adjusted Initialization for Efficient Earth Observation Gaussian Splatting

Abstract. Satellite imagery offers a distinct advantage in Earth observation by providing expansive coverage and enabling the monitoring of inaccessible regions without physical on-site intervention, serving as a significantly more cost-effective and scalable alternative to traditional aerial or ground-based surveys. The task of 3D reconstruction from multi-view satellite images has therefore been a pivotal point of research at the intersection of photogrammetry and remote sensing. Recently, novel-view synthesis techniques such as Neural Radiance Fields (NeRF) and 3D Gaussian Splatting (3DGS) have accelerated the accuracy and speed of topographic modeling. Among these, Earth Observation Gaussian Splatting (EOGS) has emerged as a state-of-the-art approach by adapting 3DGS to handle the unique geometric and radiometric characteristics of satellite data, including Rational Polynomial Coefficients (RPCs) and varying solar conditions. However, the standard EOGS pipeline relies on stochastic initialization, where Gaussians are distributed uniformly within a volumetric bounding box, leading to high computational overhead and dependency on aggressive pruning that can inadvertently remove critical geometric features, particularly in areas with complex urban structures. To address these limitations, we propose Bundle-Adjusted Initialization for Earth Observation Gaussian Splatting, which leverages sparse point clouds from bundle adjustment as geometric priors for Gaussian initialization. Combined with an adaptive densification strategy, our method achieves faster convergence and improved DSM accuracy on the DFC2019 dataset compared to the EOGS baseline.

Jiyong Kim, Shuang Song, Rongjun Qin · 0 citations
Open access Jul 2026

Evaluation of VGGT with ALS Point Clouds for Large-Scale Dense Mapping

Abstract. We present a framework that integrates ground-level imagery with Airborne Laser Scanning (ALS) point clouds. While the Visual Geometry Grounded Transformer (VGGT) enables dense geometry estimation from uncalibrated images, its application is limited by non-metric results and high GPU requirements. By leveraging publicly available, georeferenced ALS point clouds as an external metric constraint, our system restores absolute scale and global coordinates without requiring high-grade GNSS/INS or expensive on-board LiDAR systems. We introduce a confidence-weighted Sim (3) registration algorithm that utilizes a learned confidence mask to filter out unreliable points in dense street-level reconstructions. Experimental evaluations conducted on large-scale urban datasets demonstrate the average check point errors of 0.77 meters in Hong Kong dataset and 0.69 meters in Wuhan dataset, showing great potentials of feed-forward models in large-scale outdoor dense mapping.

Yandi Yang, N. El-Sheimy · 0 citations
Open access Jul 2026

Automated Monitoring of Geolocation Consistency in Micro-satellite SAR Imagery

Abstract. High revisit-rate Synthetic Aperture Radar (SAR) constellations generate large volumes of imagery that require consistent geolocation accuracy to support applications such as change detection and interferometry. However, variations in orbit determination, attitude knowledge, and external factors such as Global Navigation Satellite System (GNSS) interference can introduce geolocation errors that vary across acquisitions, making large-scale validation challenging. This study presents an automated approach to detect and quantify geolocation offsets in ICEYE SAR imagery by aligning orthorectified scenes with reference images using feature-based matching and correlation-based refinement. The method is validated against independently derived absolute geolocation measurements from corner reflector calibration sites in the United States, Canada, Australia, and Poland. Evaluation across 726 acquisitions demonstrates strong agreement with reference measurements, achieving an overall root-mean-square error (RMSE) of 1.39 m, with RMSE values of 1.18 m for Spotlight mode and 1.93 m for Stripmap mode. Operational applicability is demonstrated through large-scale acquisition campaigns, including nationwide Stripmap coverage over Japan and coherent image stack analysis. The results show that the proposed method can reliably estimate geolocation offsets, detect anomalies, and monitor geometric consistency across large SAR archives, providing a practical and scalable solution for automated geolocation quality control in micro-satellite SAR constellations.

A. Johnsy, Eyrin Kim, Qiaoping Zhang et al. · 0 citations
Open access Jul 2026

Rapid Georeferencing of Sensor-Limited Helicopter Imagery for Wildfire Response

Abstract. In the initial response to wildfires, securing rapid and accurate geographic information is essential. However, helicopter imagery acquired on-site often lacks precise sensor metadata, such as camera pose and internal parameters, making the application of georeferencing difficult. In particular, obliquely captured wildfire imagery presents additional registration challenges due to severe viewpoint changes, scale variations, and low-texture environments. This study proposes an automated georeferencing pipeline capable of operating under these constraints. The proposed method consists of five stages: preprocessing, image retrieval, feature extraction and matching, Exterior Orientation Parameters (EOP) estimation, and orthomosaic generation. An initial Area of Interest (AOI) is defined using inaccurate initial position data, and the Region of Interest (ROI) within the reference map is obtained through a ResNet50-based image retrieval approach. Subsequently, virtual Ground Control Points (GCPs) are generated through deep learning-based feature matching. Elevation data is then assigned using a Digital Elevation Model (DEM), and EOP are estimated via Perspective-n-Point (PnP) and RANSAC algorithms. Intermediate frames are initialized via interpolation and refined through bundle adjustment to produce the final orthomosaic. Experimental results demonstrated that utilizing SuperGlue and LightGlue complementarily increased the number of successfully georeferenced intervals from 5 to 9. Furthermore, a minimum RMSE of 28.30 m was achieved in the most accurate interval. This method proves that by automating the feature-based georeferencing process, practical geographic information can be rapidly provided for initial disaster response, even in sensor-limited environments.

Seongyun Kim, Jeonghyo Oh, J. Cheon et al. · 0 citations
Review Open access Aug 2026

From Multisensor Fusion to Intelligent Geospatial Monitoring: Emerging Architectures for Geotechnical Hazard Assessment

Geotechnical hazards such as landslides, subsidence, slope instability, and infrastructure deformation threaten rapidly urbanising and environmentally stressed regions worldwide, intensifying the need for scalable and intelligent monitoring systems capable of continuously observing complex Earth surface dynamics. Although multisensor remote sensing fusion has substantially expanded the observational capabilities of modern geotechnical monitoring through the integration of Synthetic Aperture Radar (SAR), optical imagery, Light Detection and Ranging (LiDAR), and environmental data, existing fusion pipelines remain subject to several well-documented constraints, including weak semantic alignment, limited temporal reasoning, and poor transferability across heterogeneous environmental conditions. This review synthesises the emerging transition from conventional sensor-centric fusion toward intelligent geospatial monitoring architectures centred on deep multimodal representation learning, transformer-based temporal reasoning, self-supervised learning, and geospatial foundation models. Particular emphasis is placed on how recent architectures are designed to better preserve coherent spatial, temporal, and contextual environmental relationships within unified latent representation spaces rather than through downstream handcrafted integration. The review further examines the growing role of multimodal transformers, masked autoencoders, contrastive learning, and large-scale geospatial foundation models in enabling scalable environmental reasoning, adaptive multimodal learning, and transferable geospatial intelligence across sensing modalities and geographic domains. Finally, remaining challenges involving uncertainty, explainability, computational scalability, and environmental generalisation are discussed alongside future research directions involving continual learning, physics-aware artificial intelligence, and autonomous geotechnical monitoring systems. Together, the reviewed literature suggests that multimodal Earth observation is evolving from passive environmental sensing toward adaptive geospatial intelligence systems capable of scalable hazard reasoning and autonomous environmental understanding.

Meghdad Bagheri, Thalosang Tshireletso, Seyed Ali Ghorashi · 0 citations
Open access Jul 2026

Challenges in automated 4D Point Cloud Generation for Glacier Calving Monitoring at high temporal Resolution

Abstract. To robustly support glacier calving monitoring at high temporal resolution and enable future AI-based calving forecasts, this study presents an optimized Multi-Epoch Multi-Imagery (MEMI) strategy for automated 4D point cloud model generation. To date, the dataset comprises over 160,000 images acquired since December 2024 by an autonomous multi-camera system operating at 30 min intervals at Glacier Perito Moreno (GPM), Argentina. Despite high scene variability and harsh environmental conditions, the proposed MEMI workflow effectively addresses constraints imposed by continuous glacier motion and image degradation. The enhanced strategy aims to generate precise dense clouds with high alignment accuracy and computational efficiency, forming the basis for subsequent analysis of glacier front evolution. To achieve this, various parameter configurations are evaluated, including AI-based image masking and adaptive, optimized alignment-adjustment settings. Results from a representative eight-day subset show that variations in the tie point computation strategy lead to measurable differences in alignment-adjustment efficiency, with the best configuration being about 11 % faster than the least efficient one. By contrast, adaptive alignment-adjustment consistently improves alignment accuracy. Moreover, masking enhances both image quality checking and reconstruction quality, and, albeit modestly, improves pre-failure deformation analysis. Furthermore, daily seasonal responses to alignment are observed, as accuracy varies with solar illumination relative to the camera positions. Applying the optimal configuration to 260 MEMI projects in under 42 h produced 518 high-precision dense clouds and detected calving retreat magnitudes of up to 17.5m, demonstrating the robustness and scalability of the proposed MEMI strategy for high-temporal-resolution 4D point cloud generation.

Laura Camila Duran Vergara, Xabier Blanch Górriz, Bindusara Nagathihalli Lokesh et al. · 1 citation