Next Article in Journal
Utilizing Multimodal Logic Fusion to Identify the Types of Food Waste Sources
Previous Article in Journal
MFPNet: A Semantic Segmentation Network for Regular Tunnel Point Clouds Based on Multi-Scale Feature Perception
 
 
Font Type:
Arial Georgia Verdana
Font Size:
Aa Aa Aa
Line Spacing:
Column Width:
Background:
Article

Method for Improving Positioning Accuracy of Rotating Scanning Satellite Images via Multi-Source Satellite Data Fusion

1
School of Aerospace, Harbin Institute of Technology, Shenzhen 518055, China
2
Key Laboratory of Aerospace RS Big-Data Intelligent Processing and Application of Guangdong Higher Education Institutes, Harbin Institute of Technology (Shenzhen), Shenzhen 518055, China
3
Space Engineering University, Beijing 101416, China
*
Author to whom correspondence should be addressed.
Sensors 2026, 26(3), 850; https://doi.org/10.3390/s26030850
Submission received: 25 December 2025 / Revised: 18 January 2026 / Accepted: 26 January 2026 / Published: 28 January 2026
(This article belongs to the Section Navigation and Positioning)

Highlights

What are the main findings?
  • A multi-source collaborative positioning framework integrates ZY-3 and GF-2 imagery to achieve a planar accuracy of 4.01 m and an edge matching RMSE of 2.52 m.
  • The proposed grid-based feature extraction and joint adjustment method successfully attains meter-level positioning accuracy (4.68 m and 5.22 m) for simulated ultra-wide rotating scanning imagery.
What are the implications of the main findings?
  • The approach effectively mitigates complex geometric distortions in rotating scanning systems without dense ground control points or strict physical models.
  • This study provides a robust solution for generating seamless, high-precision orthophoto products from ultra-wide swath satellite data using multi-source fusion.

Abstract

Rotating scanning systems are capable of acquiring ultra-wide swath satellite imagery, but they suffer from significant positioning accuracy degradation due to complex geometric distortions and the difficulty of obtaining ground control points (GCPs) over vast areas. To address these issues, this paper proposes a precise positioning method based on multi-source satellite data fusion. By comprehensively utilizing high-resolution images from ZY-3 and GF-2 satellites alongside DEM data, we establish a framework that integrates grid-based feature point extraction, high-precision matching, and multi-image joint adjustment. Specifically, we introduce a matching strategy combining geometric constraints with Least Squares Minimization (LSM) and a robust joint adjustment model to suppress geometric distortions. Experimental validation was conducted using a dataset covering the Beijing area. The results demonstrate that after joint adjustment, the planar accuracy of the imagery reached 4.01 m, and the edge matching Root Mean Square Error (RMSE) between adjacent images was 2.52 m. Furthermore, the cooperative positioning accuracy for segmented simulation data achieved 4.68 m in mountainous areas and 5.22 m in plain areas, meeting the requirements for meter-level positioning. These results verify the effectiveness of multi-source cooperative adjustment in correcting geometric distortions and significantly improving the positioning accuracy of rotating scanning imagery.

1. Introduction

With the expansion of remote sensing applications, requirements for information acquisition have evolved remarkably. Critical applications, such as global change monitoring [1] and natural resource surveys, now demand a dual capability: rapid, large-area coverage (i.e., wide swath) combined with high spatial resolution [2]. In recent years, high-resolution observation technology has advanced rapidly, providing rich information resources for the geosciences [3]. However, constraints inherent to existing imaging mechanisms—such as camera fields of view and satellite orbital regression cycles—create a mutual trade-off between spatial resolution and swath width in optical remote sensing satellites [4]. Consequently, simultaneously achieving meter-level resolution and thousand-kilometer-level coverage with a single satellite remains a formidable challenge [5].
To surmount these limitations, the rotating scanning imaging system has been proposed. Unlike the mainstream static pushbroom imaging employed by most optical satellites, the CMOS sensor in a rotating scanning system is aligned along the flight direction but rotates continuously 360° around an axis perpendicular to the trajectory. This mechanism allows the satellite to sweep the ground continuously, acquiring image strips that are subsequently stitched into a seamless panoramic view of the target area.
However, satellite remote sensing imagery obtained via this imaging method faces several key challenges. First, the imaging geometry results from the complex coupling between the satellite’s orbital motion and the high-speed rotation of the sensor, introducing high-order and time-varying geometric distortions. Such dynamic deformation renders traditional static geometric models and standard Rational Polynomial Coefficient (RPC) correction methods inadequate for achieving high-precision geolocation, often leading to model instability. Second, rotating-scanning satellites typically cover ultra-wide swaths, often on the order of thousands of kilometers [2], making it practically impossible to obtain a sufficient number of well-distributed, field-surveyed ground control points (GCPs) across the entire coverage area. To bridge this gap, this paper proposes a multi-source collaborative positioning framework that leverages existing high-precision satellite data (e.g., ZY-3) as a control skeleton to automate accuracy transfer, bypassing the GCP acquisition bottleneck.
With these research challenges and opportunities, this study aims to propose a positioning framework via multi-source satellite data fusion. The main contributions of this paper are as follows:
  • Establishment of a Positioning Framework via Multi-Source Satellite Data Fusion: We have constructed a comprehensive framework for the precise positioning of ultra-wide swath imagery. By synergistically utilizing multi-source high-resolution optical satellite imagery and DEM data, we established a joint adjustment model that proficiently facilitates error compensation and accuracy transfer across heterogeneous data sources.
  • Development of High-Precision Tie Point Extraction Technology: Addressing the severe deformation characteristics of ultra-wide images, we developed a robust automatic tie point extraction technique. This method integrates an image blocking and gridding strategy with geometric constraints and Least Squares Minimization (LSM) optimization, ensuring a balanced distribution and high-precision extraction of feature points.
  • Comprehensive Experimental Validation: We designed a rigorous validation scheme using simulated wide-swath data derived from real satellite imagery. Through quantitative evaluation of both absolute positioning accuracy and relative edge matching accuracy, we verified the method’s effectiveness and engineering practicality for rotating scanning satellite applications.
This study presents significant advancements in the geometric rectification of spaceborne sensors, primarily through two key contributions: (1) To address the critical challenge of acquiring Ground Control Points (GCPs) in ultra-wide swath imaging, an automated control information transfer mechanism utilizing multi-source reference imagery has been developed, thereby overcoming the technical bottleneck associated with control data scarcity for swaths spanning thousands of kilometers. (2) Targeting the severe geometric instability inherent in rotating scanning sensors, a joint adjustment model is proposed to successfully suppress kilometer-level distortions. By specifically modeling and compensating for high-order time-varying errors, this model achieves a substantial reduction in absolute positioning error from the kilometer-scale commonly reported in the literature to within 10 m.

2. Related Work

The geometric processing of optical satellite imagery has evolved notably, driven by the diversity of imaging mechanisms and the demand for high-precision geospatial products. This section reviews the current state of research in three key areas relevant to this study: geometric modeling of whiskbroom/scanning systems, jitter detection and correction, and advanced computational processing for remote sensing data.

2.1. Geometric Modeling and Calibration of Whiskbroom Imaging Systems

While geometric correction for static pushbroom cameras is relatively mature [6,7], the modeling of whiskbroom scanning systems presents unique challenges due to the complex coupling of scanning motion and satellite flight. Early research by Breuer and Albertz established foundational concepts for airborne whiskbroom scanners, emphasizing the need for hybrid auxiliary data to correct panoramic and motion-induced distortions [8]. Building on these principles, Uto et al. developed low-cost whiskbroom imagers using optical fiber bundles, demonstrating that hardware-specific geometric distortions can be mitigated through precise optical coupling and registration algorithms [9].
In the context of spaceborne platforms, the “rotating scanning” or “conical scanning” mechanism has been increasingly explored to achieve ultra-wide swaths. Wang et al. proposed a conceptual rotational mode for optical conical scanning imaging small satellites, analyzing the imaging degradation caused by the compound motion of satellite flight and sensor rotation [10]. To mitigate image blur in such dynamic systems, Sun et al. introduced an image motion compensation method using two-axis fast steering mirrors, establishing a quantitative relationship between the scanning trajectory and compensation rates [11].
Specific missions have driven further methodological innovations. For the SDGSAT-1 Thermal Infrared Spectrometer (TIS), Hu et al. detailed the system design which utilizes a 1-D scanning mirror to achieve a 300 km swath [12], while Li et al. proposed a three-step in-orbit calibration strategy decoupling errors from exterior orientation and scanning compensation parameters [13]. Similarly, for the HJ-2 A/B satellites, Zhang et al. introduced a multi-focal-plane-array joint calibration method to improve band-to-band registration accuracy [14]. Additionally, Sun developed a geometric calibration model for linear-array whiskbroom satellites based on look-angle corrections, utilizing cubic polynomial surfaces to fit correction quantities [15].
Most relevant to this study, Liu et al. constructed a rigorous imaging model specifically for linear array sensors performing circular scanning perpendicular to the orbit, proposing a correction method based on orbital attitude parameters and spline function fitting [16]. Furthermore, Li et al. applied rigorous photogrammetric processing to Tianwen-1 HiRIC imagery, demonstrating that bundle adjustment with high-frequency jitter compensation can achieve sub-pixel accuracy even in complex deep-space environments [17]. Zhang et al. also developed a calibration method for airborne linear-array multi-camera systems, utilizing bundle adjustment with orientation constraints to ensure high-precision stitching [18].
Despite these advances in imaging models, existing methods primarily focus on sensor-specific calibration or standard linear scanning mechanisms. They often lack a unified solution for the complex, high-order geometric distortions characteristic of continuous 360° rotating scanning systems over ultra-wide swaths. To address this, this study builds upon these rigorous geometric principles, proposing a refined model capable of integrating external constraints to suppress the specific non-linear deformations inherent in rotating scanning imagery.

2.2. Attitude Jitter Detection and Correction

For high-resolution satellites, micro-vibrations or “jitter” can remarkably deteriorate geometric quality. Rotating scanning systems are particularly susceptible to high-frequency attitude variations during the scan cycle. Zhang et al. addressed this by deriving an image-based jitter inversion model that accounts for the Time Delay Integration (TDI) effect [19]. Additionally, Zhang et al. conducted a comprehensive simulation of vibration influences on rotating scanning satellites, analyzing how different vibration frequencies affect geometric positioning and providing a theoretical basis for payload structural design [20].
However, while these studies successfully characterize vibration impacts, practical correction of these high-frequency attitude variations remains challenging, especially in scenarios lacking dense GCPs. Therefore, distinct from purely internal jitter detection, this paper proposes to mitigate these geometric instabilities by incorporating multi-source reference data into a joint adjustment framework, thereby compensating for positioning errors induced by dynamic sensor motion.

2.3. Multi-Source Fusion and Advanced Geometric Processing

The synergy of multi-source data is critical for improving mapping accuracy. Silva et al. demonstrated a fusion framework combining GEDI, ICESat-2, and NISAR data, proving that integrating sparse lidar samples with wall-to-wall radar data can yield spatially comprehensive biomass maps [21]. To reduce dependence on ground control points (GCPs), Afsharnia et al. proposed a geometric correction method based on DEM matching, correcting orbital parameters without ground control [22]. Similarly, Tatar introduced a geolocation bias correction for Cartosat-1 using virtual GCPs generated from open-source DEMs [23]. Zhou et al. achieved high-accuracy georeferencing for GF-6 WFV images using a similar reference-based approach [24].
In terms of advanced processing, deep learning and implicit neural representations are reshaping geometric pipelines. Song et al. evaluated deep learning-based features against handcrafted features for satellite image matching, finding that learning-based methods offer superior robustness [25]. Lu et al. proposed SatMVS, transferring Multi-View Stereo (MVS) neural networks to satellite imagery for robust height estimation [26]. Furthermore, Marí et al. explored the application of Neural Radiance Fields (NeRF) to multi-date Earth observation data, demonstrating that implicit representations can model geometry and appearance changes effectively [27]. Finally, Liu et al. proposed a parallel optimization approach on DCU clusters to handle the massive data volume generated by wide-swath satellites [28].
Nevertheless, directly applying these general fusion strategies to rotating scanning systems is hindered by the substantial differences in resolution and viewing geometry compared to standard reference satellites (e.g., ZY-3, GF-2). Bridging this gap, our work develops a robust multi-source fusion pipeline. While learning-based methods mentioned above show promise in general scenarios, they often require extensive training on domain-specific datasets (e.g., rotating scanning distortions) and can lack the rigorous geometric interpretability required for metric-level joint adjustment. Therefore, this study prioritizes a classical yet rigorous photogrammetric approach to ensure deterministic sub-pixel precision.

3. Methods

This study employs an integrated collaborative positioning framework to rectify the geometric distortions of rotating scanning satellite imagery by fusing multi-source high-resolution data (Figure 1). A high-precision control grid is first constructed by extracting feature points from reference ZY-3 and GF-2 imagery using a grid-based strategy and determining their 3D coordinates via a reference DEM. To address the radiometric and geometric disparities between heterogeneous sensors, a coarse-to-fine matching approach combining Normalized Cross-Correlation (NCC) and Least Squares Matching (LSM) is utilized to generate reliable tie points. Subsequently, for the target rotating scanning images, a Multi-Directional Spatial Context Information (MSCI) matching method is applied to accurately transfer control information. Finally, a multi-image joint adjustment model, reinforced by robust estimation, is implemented to optimize the orientation parameters and generate high-precision Digital Orthophoto Maps (DOM).

3.1. Multi-Source Control Point Grid Generation Method

3.1.1. Geometric Control Principle of Heterogeneous Sensor Integration

Integrating data from diverse platforms (e.g., GF-1, GF-2, ZY-3) presents inherent challenges due to differences in spatial resolution, spectral response, and initial positioning accuracy. To address the aforementioned challenges, this section employs a heterogenous sensor integration strategy centered on constructing a hierarchical, robust joint adjustment system. Its fundamental principle is as follows: first, all images to be fused are treated as a unified photogrammetric regional network. Second, high-precision connection point extraction establishes geometric links between all overlapping image pairs. Finally, within a unified adjustment model, the systematic error model parameters for all images and the three-dimensional ground coordinates for all connection points are jointly solved. In this progress, we adopt a “hierarchical constraint” strategy within a unified photogrammetric block adjustment. The core principle is to assign higher weights to images with superior initial positioning accuracy (such as ZY-3), allowing them to function as a “control skeleton”. Images with lower initial accuracy are then geometrically constrained to these anchors. Through this mechanism, geometric accuracy is proficiently transferred and distributed throughout the entire block, unifying the geometric datum across the dataset.

3.1.2. High-Precision Tie Point Extraction Strategy

We employ a robust two-step matching strategy: “Normalized Cross-Correlation (NCC) Coarse Matching” followed by “Least Squares Matching (LSM) Fine Matching”.
NCC Coarse Matching To mitigate the effects of geometric deformation and radiometric differences, NCC is used for initial matching. The NCC coefficient is calculated as follows:
C = i = 1 m j = 1 n f ( i , j ) f ¯ g ( i , j ) g ¯ i = 1 m j = 1 n f ( i , j ) f ¯ 2 i = 1 m j = 1 n g ( i , j ) g ¯ 2 ,
where m , n represents the dimensions of the template window, f ( i , j ) and g ( i , j ) denote the pixel gray values at corresponding positions in the template and search images, respectively, and f ¯ and g ¯ represent the average gray values of the two windows. The location yielding the maximum NCC coefficient is thus regarded as the approximate position of the corresponding point.
To enhance efficiency and robustness against rotation, we implement two optimization strategies:
  • Pyramid Strategy: Matching initiates at the top level of a Gaussian pyramid (low resolution) and progressively refines to the bottom level, considerably reducing the search space.
  • Rotation Template Generation: Using the RPC files, we estimate the relative rotation angle θ between images. The template is then rotated to compensate for deformation:
    x y = cos θ sin θ sin θ cos θ x x 0 y y 0 + x 0 y 0 ,
    where ( x , y ) represents the original template coordinates; ( x , y ) represents the rotated coordinates; and ( x c , y c ) represents the rotation center.
LSM Fine Matching: For sub-pixel accuracy, LSM is employed to solve for geometric and radiometric parameters. The fundamental observation equation is:
x x 0 = f a 1 ( X X S ) + b 1 ( Y Y S ) + c 1 ( Z Z S ) a 3 ( X X S ) + b 3 ( Y Y S ) + c 3 ( Z Z S ) y y 0 = f a 2 ( X X S ) + b 2 ( Y Y S ) + c 2 ( Z Z S ) a 3 ( X X S ) + b 3 ( Y Y S ) + c 3 ( Z Z S )
where ( X , Y , Z ) are the object space coordinates; ( X S , Y S , Z S ) denote the sensor’s perspective center; f is the effective focal length; and a i , b i , c i are the elements of the rotation matrix R determined by the platform’s roll, pitch, and yaw.
We utilize an affine transformation model to describe the geometric relationship between windows, solving for 6 geometric parameters and 2 radiometric parameters. The linearized error equation for each pixel is
v = g x d a + g x x d b + g x y d c + g y d d + g y x d e + g y y d f + h 0 + h 1 g ( x , y ) Δ g
This system is solved iteratively using the least squares method, X = ( A T P A ) 1 A T P L , until the parameter corrections fall below a preset threshold.

3.1.3. Control Point Library Generation via Moravec Operator

To generate the control point library, we utilize the Moravec corner detection operator [24]. While more modern detectors like Harris, SIFT, or learning-based methods offer higher rotation invariance, the Moravec operator is selected for this framework due to its notably lower computational complexity and high processing speed—critical for handling the ultra-wide swath data involved in rotating scanning imagery. This involves computing the Sum of Squared Differences (SSD) in four directions (horizontal, vertical, diagonal, and anti-diagonal):
V = ( f ( x i , y i ) f ( x i + u , y i + v ) ) 2
The “Interest Value” for each pixel is determined by the minimum SSD among the four directions: I V = min ( V 1 , V 2 , V 3 , V 4 ) . Following non-maximum suppression to retain only local maxima, the extracted corners are mapped to object space coordinates using the reference DOM and DSM, forming a high-precision control information library.

3.2. Collaborative Positioning Based on Semantic Features

Once the control library is established, the next phase is the collaborative positioning of the target rotating scanning imagery.

3.2.1. Control Point Transfer Using Spatial Context Features

Due to the complex geometric distortions and non-linear radiometric differences between the reference and rotating scanning images, traditional intensity-based matching is often insufficient. We therefore propose a matching method based on Multi-Directional Spatial Context Information (MSCI). Instead of using raw pixel values, this method aggregates the gray value relationships between a central pixel and its neighbors to construct structural features, enhancing robustness against contrast variations.
As depicted in the schema in Figure 2, the local neighborhood around a feature point is partitioned into multiple cells. This method aggregates the gray value relationships between a central pixel I ( 0 , 0 ) and its neighbors ( I ( 1 , 0 ) , I ( 1 , 1 ) , I ( 0 , 1 ) ) to construct structural features. The spatial relationship and intensity contrast between these cells are visualized in Figure 2.
By linking each directional cell in Figure 2 to a specific dimension in the formula, the MSCI model ensures that the matching cost calculated in Equation (3) reflects the actual spatial structure of the satellite imagery.
To suppress noise inherent in multi-source data (e.g., SAR or LiDAR intensity), Gaussian filtering is applied to the feature channels. Subsequently, the features are normalized to create the MSCI descriptor, enhancing robustness against contrast variations:
M S C I ( x ) = exp C ( x ) C ( y ) 2 σ 2
Finally, matching is performed in the frequency domain. By calculating the Cross-Power Spectrum via Fast Fourier Transform (FFT), we obtain a pulse function where the peak location indicates the precise translation shift.

3.2.2. Multi-Image Joint Adjustment and Robust Estimation

We employ a multi-image joint adjustment model based on collinearity condition equations to optimize the orientation parameters. The process includes extracting tie points between overlapping strips and transferring control points from the library [24]. Error equations are constructed for both control points (using affine parameters as unknowns) and tie points (using object coordinates as unknowns). To eliminate gross errors generated during automatic matching, we apply robust estimation using the Iteratively Reweighted Least Squares (IRLS) method. The weight matrix is updated in each iteration using the Huber weight function:
P i i = 1 | v i | k σ 0 k σ 0 / | v i | | v i | > k σ 0
This ensures that observations with large residuals are down-weighted, preventing them from distorting the final adjustment solution.

4. Results

4.1. Experiment Area and Data Sources

To validate the effectiveness of the proposed method across different geomorphological conditions, the test area was selected as the experimental site (N 39.465° to N 40.461°, E 115.672° to E 116.496°). This region provides a natural laboratory for comparative analysis due to its distinct topographic dichotomy: the northwestern part is dominated by rugged mountainous terrain, while the southeastern part consists of the flat North Plain.
The experimental dataset consists of six multi-source high-resolution optical satellite images:
  • Reference Data: Two scenes of ZY-3 panchromatic imagery (2.1 m resolution) and four scenes of GF-2 satellite imagery with high internal geometric stability served as the control backbone;
  • Target Data: Rotating scanning satellite simulation data from the research by Xue [1]. To simulate the rotating scanning imagery, the original GF-2 push-broom data was re-projected onto a cylindrical surface. We applied a transformation based on the scanning angle θ ( t ) = ω t + δ θ , where ω is the nominal scan rate and δ θ represents simulated attitude jitter. This ensures the resulting imagery contains the characteristic S-shape distortion and scale variation typical of rotating sensors. For this comparative study, 80 representative chips were selected: 40 chips covering the western mountainous areas and 40 chips covering the southeastern plains;
  • Auxiliary Data: SRTM DEM data were utilized to provide elevation constraints for the joint adjustment and orthorectification processes.

4.2. Control Point Extraction and Joint Adjustment Accuracy

Using the grid-based feature extraction and MSCI matching strategy, a dense network of control and tie points was established. As illustrated in Figure 3, the study area is divided into regular grids, and tie points are extracted within each grid to ensure an even spatial distribution. This distribution serves as the physical basis for the ‘control skeleton’ mentioned above, ensuring that the high-precision constraints from ZY-3 are uniformly propagated through the adjustment equations discussed in Section 3.2.2. On average, 2046 points per chip were extracted for the mountainous region and 2475 points per chip for the plain region. The higher point density in the plain region is attributed to the abundance of man-made structures (roads and building corners), which are conducive to feature matching.
Following the multi-image joint adjustment, we evaluated the positioning accuracy using high-precision checkpoints. Table 1 summarizes the positioning accuracy after multi-image joint adjustment. The results demonstrate that the systematic errors across the heterogeneous sensors were efficiently unified, achieving a consistent planar accuracy level. The average planar positioning accuracy (RMSE) after adjustment reached 4.01 m. Considering the reference imagery itself has a planar accuracy of approximately 5 m (leading to a theoretical error propagation of 6.41 m), the adjustment result of 4.01 m significantly outperforms the theoretical value, demonstrating the efficacy of the error compensation model.
Furthermore, we assessed the relative consistency between adjacent images. As shown in Table 2, the average edge matching RMSE between all overlapping image pairs was 2.52 m. This high level of consistency indicates that systematic errors between multi-source images were successfully eliminated, satisfying the requirements for seamless stitching.

4.3. Collaborative Positioning Performance on Ultra-Wide Images

The core of the experiment involved evaluating the cooperative positioning accuracy of the simulated ultra-wide rotating scanning image segments. We conducted a rigorous comparison between the original direct positioning and the proposed collaborative adjustment for both mountainous and plain sub-regions.
To compare the model’s performance in mountainous and plain regions, the data for these areas will now be processed and analyzed separately. Table 3 presents the initial direct positioning accuracy of the images from mountainous areas, with Table 4 for plain areas. Due to the high-order distortions inherent in the rotating scanning mechanism, the original imagery exhibited severe geometric displacement.
The initial errors were heavily dominated by the X-direction (cross-track/scanning direction), reaching over 80 m in mountainous areas. This is consistent with the physical characteristics of rotating scanning, where scanning non-linearity and terrain-induced relief displacement are most pronounced in the cross-track direction. Mountainous regions exhibited higher original errors (82.22 m) compared to plains (78.87 m) because the extreme elevation changes further amplified the relief displacement in the scanning geometry.
After applying the proposed collaborative positioning correction based on multi-source control data, the accuracy improved drastically. As detailed in Table 5 and Table 6, the experimental data reveals several key insights:
  • Effective Geometric Recovery: The planar accuracy (RMSE XY) was reduced to under 5.5 m for all regions, meeting the requirements for meter-level positioning.
  • Terrain Adaptability: Interestingly, the correction accuracy in mountainous areas (4.68 m) was slightly superior to that in plain areas (5.22 m). This demonstrates that the MSCI descriptor and DEM-assisted adjustment are highly effective at capturing stable topographic structural features in rugged terrain, which are less susceptible to temporal land-use changes.
  • Error Isotropy: Following the adjustment, the extreme X-direction bias was eliminated. The residual errors in the X and Y directions are now within the same order of magnitude (3–4 m), indicating that the systematic scanning distortions were successfully modeled and suppressed.

4.4. Digital Orthophoto Map (DOM) Generation

Based on the high-precision orientation parameters obtained from the joint adjustment and the external DEM, we generated a seamless Digital Orthophoto Map (DOM) for the entire experimental area using pixel-by-pixel rectification. Visual inspection of the generated DOM products reveals clear image textures and accurate planar positions (Figure 4). The seamless mosaic further validates the high edge-matching accuracy achieved, proving the method’s engineering practicality for producing high-quality mapping products from rotating scanning satellite data.
Visual inspection of the generated DOM products reveals clear image textures and accurate planar positions. The seamless mosaic further validates the high edge-matching accuracy achieved, proving the method’s engineering practicality for producing high-quality mapping products from rotating scanning satellite data.

4.5. Computational Cost and Operational Applicability

To evaluate the operational feasibility of the proposed framework, the processing time was recorded on a Dell 7040 workstation equipped with two Intel Xeon Gold 5118 processors (2.3 GHz) and 64 GB of RAM. For a typical image chip, the automated grid-based feature extraction and MSCI matching require approximately 45.2 s, while the joint adjustment with robust estimation converges in less than 2.5 s. The total processing time for the ultra-wide swath simulation (80 chips) was approximately 64 min. This efficiency indicates that the method is suitable for large-scale, near-real-time satellite data processing without manual intervention.

5. Discussion

The results of this study demonstrate that the proposed multi-source data fusion framework proficiently addresses the geometric challenges of rotating scanning satellites. A critical interpretation of the results reveals a counter-intuitive finding: despite the higher initial geometric complexity in mountainous areas (82.22 m initial error), the final collaborative positioning accuracy (4.68 m) is superior to that of the plain areas (5.22 m). This disparity highlights the relative contribution of the MSCI descriptor. In urban plains, high-frequency temporal changes in buildings and shadows introduce ‘textural noise’ that slightly degrades matching precision. In contrast, the rugged terrain provides stable, unique topographic skeletons (e.g., ridges and valleys) that the MSCI descriptor captures with higher semantic fidelity. This suggests that the proposed method is particularly robust for remote, unpopulated mountainous regions where GCPs are hardest to obtain.
Despite these promising results, several avenues for improvement remain:
  • Integration of On-board Data: Future work should incorporate satellite attitude and Inertial Measurement Unit (IMU) data to construct a more rigorous physical imaging model, thereby reducing error accumulation at the source.
  • Robustness in Complex Terrain: While the method performed well in the test area, extracting reliable tie points in extreme terrains (e.g., high relief mountains) or urban canyons remains challenging. Developing adaptive filtering and matching algorithms for these scenarios is necessary.
  • Multi-Modal Fusion: To further enhance positioning capabilities, particularly in varying weather conditions, we aim to explore the online collaborative processing of heterogeneous data, including Optical, SAR, and LiDAR, to achieve intelligent, near-real-time earth observation processing.
  • Transferability and Limitations in Other Terrains: While the experimental results in demonstrate high accuracy across mountainous and plain areas, the transferability of the proposed method to other global terrains warrants further discussion:
    • High-Relief Mountains: The synergy between MSCI semantic features and DEM-based elevation constraints allows the model to proficiently rectify non-linear relief displacements. However, in extreme high-relief areas (e.g., the Himalayas), excessive radar/optical shadowing or snow cover may reduce the number of available tie points, potentially requiring higher-resolution reference DEMs to maintain sub-pixel accuracy.
    • Deserts and Low-Texture Regions: A primary limitation arises in extremely homogeneous landscapes such as vast sand deserts or consistent ice sheets. Because the MSCI descriptor relies on local spatial contextual structures, the absence of distinct textural gradients in these regions would lead to sparse or unreliable feature matching. In such cases, the framework would rely more heavily on the satellite’s initial ephemeris data or would require the integration of multi-modal data (e.g., SAR intensity features) which are less dependent on optical texture.
    • Dependency on Reference Quality: The final absolute positioning accuracy is intrinsically capped by the precision of the reference ‘skeleton’ (e.g., ZY-3 imagery). In remote regions where high-precision reference imagery or GCP-calibrated base maps are unavailable, the absolute accuracy may degrade toward the level of the target satellite’s raw orientation parameters.

6. Conclusions

This study established a collaborative positioning framework for ultra-wide rotating scanning satellite imagery. By utilizing ZY-3 imagery as a geometric skeleton and incorporating DEM constraints, we achieved a meter-level positioning accuracy of 4.68 m (RMSE XY) for mountainous areas and 5.22 m for plain areas.
Validation Conditions and Dependencies: The results demonstrate that the method efficiently mitigates high-order scanning distortions in both mountainous (4.68 m) and plain (5.22 m) environments. However, the final accuracy remains highly dependent on the precision of the reference imagery and the vertical quality of the auxiliary DEM.
Limitations and Future Work: A primary limitation is the sensitivity to significant temporal land-use changes between multi-source sensors, which can introduce matching outliers in rapidly developing urban areas. Future research will focus on integrating multi-temporal robust filtering to further enhance reliability in dynamically changing landscapes.

Author Contributions

Conceptualization, L.W., P.W., Y.W. and B.C.; methodology, L.W., P.W. and Y.Z.; validation, L.W., P.W., Y.Z. and Y.W.; formal analysis, L.W.; investigation, L.W.; resources, B.C.; data curation, L.W.; writing—original draft preparation, L.W.; writing—review and editing, B.C.; visualization, L.W. and Y.W.; supervision, B.C.; project administration, B.C.; funding acquisition, B.C. All authors have read and agreed to the published version of the manuscript.

Funding

This research was funded by National Key Research and Development Program of China, grant number 2022YFF0503900; National Key Research and Development Program of China, grant number 2022YFD2401200; Shenzhen Higher Education Institutions Stabilization Support Program Project, grant number GXWD20220811163556003; National Natural Science Foundation of China, grant number NSFC62202127; and National Natural Science Foundation of Shenzhen, grant number JCYJ20241202123731040.

Data Availability Statement

The original contributions presented in this study are included in the article. Further inquiries can be directed to the corresponding author.

Conflicts of Interest

The authors declare no conflicts of interest.

References

  1. Xue, W.; Wang, P.; Zhong, L. Geometric Processing of New Ultra-large Swath Optical Remote Sensing Satellite Images. Remote Sens. Inf. 2021, 36, 60–65. [Google Scholar]
  2. Liu, X.; Xue, W.; Wang, P. Construction and accuracy assessment of rational function model for perpendicular-orbit circular scanning satellite images. Opt. Precis. Eng. 2023, 31, 2898–2909. [Google Scholar] [CrossRef]
  3. Wang, Y.; Fu, T.; Zhou, Y.; Kong, Q.; Yu, W.; Liu, J.; Wang, Y.; Chen, B. TENet: Attention-Frequency Edge-Enhanced 3D Texture Enhancement Network. Sensors 2025, 25, 715. [Google Scholar] [CrossRef] [PubMed]
  4. Xue, W.; Liu, X.; Zhao, L.; Wang, P.; Zhang, X.; Li, W. Research on High-Precision Geometric Correction of Agile Optical Satellite Images. J. Geo-Inf. Sci. 2025, 27, 2701–2712. [Google Scholar]
  5. Zhang, T.; Li, Y. rpcPRF: Generalizable MPI Neural Radiance Field for Satellite Camera. arXiv 2023, arXiv:2310.07179. [Google Scholar] [CrossRef]
  6. Zhang, G.; Xu, K.; Zhang, Q.; Li, D. Correction of Pushbroom Satellite Imagery Interior Distortions Independent of Ground Control Points. Remote Sens. 2018, 10, 98. [Google Scholar] [CrossRef]
  7. Tao, C.V.; Hu, Y. A comprehensive study of the rational function model for photogrammetric processing. Photogramm. Eng. Remote Sens. 2001, 67, 1347–1357. [Google Scholar]
  8. Breuer, M.; Albertz, J. Geometric correction of airborne whiskbroom scanner imagery using hybrid auxiliary data. Int. Arch. Photogramm. Remote Sens. 2000, 33, 93–100. [Google Scholar]
  9. Uto, K.; Seki, H.; Saito, G.; Kosugi, Y.; Komatsu, T. Development of a low-cost hyperspectral whiskbroom imager using an optical fiber bundle, a swing mirror, and compact spectrometers. IEEE J. Sel. Top. Appl. Earth Obs. Remote Sens. 2016, 9, 3909–3925. [Google Scholar] [CrossRef]
  10. Wang, J.; Dai, W.; Yang, S.; Xiao, L.; Li, X. Conceptual rotational mode design for optical conical scanning imaging small satellites. IEEE Access 2020, 8, 126592–126603. [Google Scholar] [CrossRef]
  11. Sun, J.; Chen, S. Conceptual design and image motion compensation rate analysis of two-axis fast steering mirror for dynamic scanning. Opt. Express 2021, 29, 5083–5097. [Google Scholar] [CrossRef]
  12. Hu, Z.; Li, X.; Li, L.; Su, X.; Yang, L.; Zhang, Y.; Chen, F. Wide-swath and high-resolution whisk-broom imaging and on-orbit performance of SDGSAT-1 thermal infrared spectrometer. Remote Sens. Environ. 2024, 300, 113887. [Google Scholar] [CrossRef]
  13. Li, X.; Li, L.; Zhao, L.; Jiao, J.; Jiang, L.; Yang, L.; Sun, S. In-Orbit Geometric Calibration for Long-Linear-Array and Wide-Swath Whisk-Broom TIS of SDGSAT-1. IEEE Trans. Geosci. Remote Sens. 2023, 61, 1–14. [Google Scholar] [CrossRef]
  14. Zhang, H.; Qin, W.; Wang, K.; Wang, Q.; Tao, P. On-Orbit Geometric Calibration of the HJ-2 A/B Satellites’ Infrared Sensors. ISPRS Ann. Photogramm. Remote Sens. Spat. Inf. Sci. 2024, 10, 305–311. [Google Scholar] [CrossRef]
  15. Sun, Y. Geometric calibration of the linear-array whiskbroom optical satellites based on look-angle corrections. J. Phys. Conf. Ser. 2024, 2724, 012032. [Google Scholar] [CrossRef]
  16. Xue, W.; Wang, P.; Zhong, L.Y. Geometric correction of optical remote sensing satellite images captured by linear array sensors circular scanning perpendicular to the orbit. Opt. Precis. Eng. 2021, 29, 2924–2934. [Google Scholar] [CrossRef]
  17. Li, D.; Yan, W.; Geng, X. Photogrammetric Processing of Tianwen-1 HiRIC Imagery. Photogramm. Eng. Remote Sens. 2022, 88, 511–518. [Google Scholar] [CrossRef]
  18. Zhang, Z.; Wang, T.; Zhang, G. Calibrating an airborne linear-array multi-camera system on the master focal plane with existing bundle adjustment. ISPRS J. Photogramm. Remote Sens. 2025, 221, 12–26. [Google Scholar] [CrossRef]
  19. Zhang, H.; Xie, B.; Liu, S.; Ding, R.; Ye, Z.; Tong, X. Detection and Correction of Jitter Effect for Satellite TDICCD Imagery. Int. Arch. Photogramm. Remote Sens. Spat. Inf. Sci. 2022, 43, 79–84. [Google Scholar] [CrossRef]
  20. Zhang, Y.; Zhang, K.; Wang, S. Imaging geometric simulation and vibration influence analysis of rotary scanning remote sensing satellites. Sci. Rep. 2025, 15, 3612. [Google Scholar] [CrossRef]
  21. Silva, C.A.; Duncanson, L.; Hancock, S.; Neuenschwander, A.; Thomas, N.; Hofton, M.; Dubayah, R. Fusing simulated GEDI, ICESat-2 and NISAR data for regional aboveground biomass mapping. Remote Sens. Environ. 2021, 253, 112234. [Google Scholar] [CrossRef]
  22. Afsharnia, M.; Arefi, H.; Alizadeh, A. Geometric correction of satellite stereo images by DEM matching without ground control points and map projection. Geocarto Int. 2022, 37, 8051–8072. [Google Scholar]
  23. Tatar, N. Geolocation bias correction for CARTOSAT-1 stereo images through virtual ground control points generated from open source DEMs. Adv. Space Res. 2024, 73, 5727–5741. [Google Scholar] [CrossRef]
  24. Zhou, Y.; Li, S.; Liu, Y. High Accuracy Georeferencing of GF-6 Wide Field of View Imagery. Remote Sens. 2023, 15, 654. [Google Scholar] [CrossRef]
  25. Song, Z.; Zhang, J.; Liu, J. Deep learning meets satellite images—An evaluation on handcrafted and learning-based features for matching. Geo-Spat. Inf. Sci. 2025, 28, 50–67. [Google Scholar]
  26. Lu, J.; Li, Y.; Zuo, Z. SatMVS: A Novel 3D Reconstruction Pipeline for Remote Sensing Satellite Imagery. In Proceedings of the International Conference on Aerospace System Science and Engineering 2021, Moscow, Russia, 14–16 July 2021; pp. 521–538. [Google Scholar]
  27. Marí, R.; Facciolo, G.; Ehret, T. Multi-Date Earth Observation NeRF: The Detail Is in the Shadows. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) Workshops, Vancouver, BC, Canada, 18–22 June 2023; pp. 2035–2045. [Google Scholar]
  28. Liu, R.; Wang, S. Parallel Optimization of Remote Sensing Image Geometric Correction on DCU. In Proceedings of the 2024 5th International Conference on Computer Engineering and Application (ICCEA), Hangzhou, China, 12–14 April 2024; pp. 865–869. [Google Scholar]
Figure 1. Overview of the workflow for multi-source collaborative positioning.
Figure 1. Overview of the workflow for multi-source collaborative positioning.
Sensors 26 00850 g001
Figure 2. Diagram of Spatial Relationships Between Pixels. (a) Schematic diagram of original pixel locations. (b) Schematic diagram of spatial contextual information in four directions.
Figure 2. Diagram of Spatial Relationships Between Pixels. (a) Schematic diagram of original pixel locations. (b) Schematic diagram of spatial contextual information in four directions.
Sensors 26 00850 g002
Figure 3. Schematic diagram of generated connection points and control points. (a) Distribution map of extracted tie points. (b) Distribution map of control points.
Figure 3. Schematic diagram of generated connection points and control points. (a) Distribution map of extracted tie points. (b) Distribution map of control points.
Sensors 26 00850 g003
Figure 4. Multi-image registration correction results. (a) Coverage map of segmented orthophotos. (b) Edge matching detail of segmented orthophotos.
Figure 4. Multi-image registration correction results. (a) Coverage map of segmented orthophotos. (b) Edge matching detail of segmented orthophotos.
Sensors 26 00850 g004
Table 1. Joint Adjustment Positioning Accuracy Statistics (Unit: m).
Table 1. Joint Adjustment Positioning Accuracy Statistics (Unit: m).
Image IDCheck PointsRMSE XRMSE YRMSE XY
GF2_38678452.222.673.48
GF2_386613063.221.883.73
GF2_968711782.372.753.63
GF2_968910452.812.253.6
ZY3_237428183.263.164.54
ZY3_791826784.22.494.88
Overall3.082.564.01
Table 2. Edge Matching Accuracy Between Images (Unit: m).
Table 2. Edge Matching Accuracy Between Images (Unit: m).
Image PairRMSE XRMSE YRMSE XY
GF2_3867GF2_96871.241.131.68
GF2_3867GF2_38660.500.500.70
GF2_3867ZY3_23742.671.152.91
GF2_9687GF2_96890.940.521.07
GF2_9687GF2_38661.120.961.48
GF2_9687ZY3_23742.931.443.27
GF2_9687ZY3_79182.861.453.20
GF2_9689GF2_38661.361.241.84
GF2_9689ZY3_23742.771.453.13
GF2_9689ZY3_79183.031.343.31
GF2_3866ZY3_23742.521.833.11
GF2_3866ZY3_79182.731.473.10
ZY3_2374ZY3_79181.401.201.85
Overall2.52
Table 3. Original Direct Positioning Accuracy Statistics for Mountainous Areas (Unit: m).
Table 3. Original Direct Positioning Accuracy Statistics for Mountainous Areas (Unit: m).
Image IDCheck PointsRMSE XRMSE YRMSE XY
2_1_1226680.382.3980.41
2_1_2257880.3280.33
2_1_3268784.243.1984.3
2_2_6260182.931.8282.95
2_4_4244183.682.383.71
2_3_1235983.552.4483.58
Overall82.133.5082.22
Table 4. Original Direct Positioning Accuracy Statistics for Plain Areas (Unit: m).
Table 4. Original Direct Positioning Accuracy Statistics for Plain Areas (Unit: m).
Image IDCheck PointsRMSE XRMSE YRMSE XY
2_5_8232177.223.0977.28
2_6_8252478.183.5378.26
2_6_7225979.574.2779.69
2_7_2226378.622.8278.68
2_7_3241979.652.9779.71
2_7_4224980.092.3480.12
Overall78.783.4578.87
Table 5. Collaborative Positioning Accuracy Statistics for Mountainous Areas (Unit: m).
Table 5. Collaborative Positioning Accuracy Statistics for Mountainous Areas (Unit: m).
Image IDCheck PointsRMSE XRMSE YRMSE XY
2_1_125764.454.926.64
2_1_226145.5146.81
2_1_326523.612.384.33
2_2_625931.72.653.15
2_4_423932.071.662.65
2_3_124101.771.722.47
Overall3.393.144.68
Table 6. Collaborative Positioning Accuracy Statistics for Plain Areas (Unit: m).
Table 6. Collaborative Positioning Accuracy Statistics for Plain Areas (Unit: m).
Image IDCheck PointsRMSE XRMSE YRMSE XY
2_5_825973.972.724.81
2_6_821814.042.024.51
2_6_724993.422.064
2_7_2259325153.684.06
2_7_3239325443.733.39
2_7_4241023373.712.81
Overall4.133.075.22
Disclaimer/Publisher’s Note: The statements, opinions and data contained in all publications are solely those of the individual author(s) and contributor(s) and not of MDPI and/or the editor(s). MDPI and/or the editor(s) disclaim responsibility for any injury to people or property resulting from any ideas, methods, instructions or products referred to in the content.

Share and Cite

MDPI and ACS Style

Wang, L.; Wang, P.; Zhang, Y.; Wang, Y.; Chen, B. Method for Improving Positioning Accuracy of Rotating Scanning Satellite Images via Multi-Source Satellite Data Fusion. Sensors 2026, 26, 850. https://doi.org/10.3390/s26030850

AMA Style

Wang L, Wang P, Zhang Y, Wang Y, Chen B. Method for Improving Positioning Accuracy of Rotating Scanning Satellite Images via Multi-Source Satellite Data Fusion. Sensors. 2026; 26(3):850. https://doi.org/10.3390/s26030850

Chicago/Turabian Style

Wang, Liwei, Peng Wang, Yamin Zhang, Yi Wang, and Bo Chen. 2026. "Method for Improving Positioning Accuracy of Rotating Scanning Satellite Images via Multi-Source Satellite Data Fusion" Sensors 26, no. 3: 850. https://doi.org/10.3390/s26030850

APA Style

Wang, L., Wang, P., Zhang, Y., Wang, Y., & Chen, B. (2026). Method for Improving Positioning Accuracy of Rotating Scanning Satellite Images via Multi-Source Satellite Data Fusion. Sensors, 26(3), 850. https://doi.org/10.3390/s26030850

Note that from the first issue of 2016, this journal uses article numbers instead of page numbers. See further details here.

Article Metrics

Back to TopTop