USGS ScienceSearch

USGS · 70226588

Incorporating interpreter variability into estimation of the total variance of land cover area estimates under simple random sampling

Abstract

Area estimates of land cover and land cover change are often based on reference class labels determined by analysts interpreting satellite imagery and aerial photography. Different interpreters may assign different reference class labels to the same sample unit. This interpreter variability is typically not accounted for in variance estimators applied to area estimates of land cover. A simple measurement model provides the basis for an estimator of the total variance ( V Total ) that takes into account both sampling variance and interpreter variance. This method requires two or more reference class interpretations (i.e., repeated measurements) obtained by analysts, working independently of each other, for the full sample or a random subsample of the full sample. Estimators of the total variance ( V ̂ Total "> V̂Total ) and the variance component attributable to interpreters ( V ̂ 1 "> V̂1 ) were obtained for the case of two reference class interpretations per repeated sample unit. To evaluate the effect of interpreter variability on variance estimation, we used land cover reference data interpreted by seven analysts who each interpreted the same 300 sample pixels from a region of the Pacific Northwest of the United States. From these data, we estimated the contribution of interpreter variance to the total variance (i.e., V ̂ 1 / V ̂ Total "> V̂1/V̂Total ) and the relative bias of the standard simple random sampling variance estimator ( V ̂ stand "> V̂stand ) as an estimator of V Total , defined as 100%*( V ̂ stand − V ̂ Total "> V̂stand−V̂Total )/ V ̂ Total "> V̂Total . For each of five land cover classes, we computed V ̂ 1 "> V̂1 , V ̂ Total "> V̂Total , and V ̂ stand "> V̂stand using the sample data from each of the 21 possible pairwise combinations of the seven interpreters, and then calculated the mean of V ̂ 1 / V ̂ Total "> V̂1/V̂Total and the mean of the estimated relative bias of V ̂ stand "> V̂stand over these 21 pairs. Based on the mean of V ̂ 1 / V ̂ Total "> V̂1/V̂Total per class, interpreter variance contributed from 25% (cropland) to 76% (grass/shrub) of the total variance, indicating that interpreter variance was a non-negligible component of the total variance. Typically, the standard variance estimator, V ̂ stand "> V̂stand , underestimated the total variance with the mean estimated relative bias ranging from −3% (cropland) to −33% (grass/shrub). Classes with greater inconsistency between pairs of interpreters had larger contributions of interpreter variance to the total variance ( V ̂ 1 / V ̂ Total "> V̂1/V̂Total ) and larger negative estimated relative bias of V ̂ stand "> V̂stand . Given that interpreter variance can contribute substantially to the total variance, the repeated measurements approach offers a practical way to incorporate this variability into an estimator of the total variance.

Explore related subjects

90° N90° S · 180° W ← longitude → 180° E
Source-reported bounding extent: 46.92025531537451° to 49.009050809382046° latitude; -123.46435546875° to -121.53076171875° longitude. This indicates report coverage, not an exact sampling location. View area on OpenStreetMap.

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Stephen V. Stehman, John Mousoupetros, Ronald E. McRoberts, Erik Naesset, Bruce Pengra, Dingfan Xing, Josephine Horton. 2022. Incorporating interpreter variability into estimation of the total variance of land cover area estimates under simple random sampling. https://doi.org/10.1016/j.rse.2021.112806

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related USGS reports

On-demand global Landsat evapotranspiration product: Development, evaluation, and dissemination

Global actual evapotranspiration (ET) is one of the essential climate variables needed to understand and manage the relationships among food, energy, and water resources. The U.S. Geological Survey Earth Resources Observation and Science (EROS) Center launched a provisional ET product in 2020, offering on-demand, field-scale global coverage derived from Landsat data through the EROS Science Processing Architecture (ESPA) platform. The ESPA interface provides ET data for cloud-free Landsat overpasses starting in 1982 with Landsat 4 through the current Landsat 9. The ET data are delivered as a Provisional Level-3 Science product created using the Operational Simplified Surface Energy Balance (SSEBop) model. Landsat surface temperature and reference ET are the main model drivers along with vegetation index and net radiation for model parameterization. A large volume of Landsat-based ET orders (e.g., over 1,200,000 images from June 2020 through December 2025) around the world indicate increasing awareness and application of the ET data. The ESPA platform enables land and water resource managers and researchers to access a first-order ET product without requiring advanced knowledge of remote sensing technology or evapotranspiration modeling. We present the methodology and workflow of the on-demand Landsat ET product and its performance evaluations over diverse hydro-climatic settings. The product can help estimate field-scale consumptive water use and thus quickly and consistently assess historical water use, allocation, and budget to inform water management under changing environments. Future ET data aggregated to monthly and seasonal time scales are expected to enhance integration with decision-making tools and procedures.

Remote Sensing of Environment

A framework for integrating spatiotemporal deep learning methods with landsat for annual land cover and impervious surface mapping

Land cover information is essential for understanding Earth’s surface dynamics and how vegetation, water, soil, climate, and terrain interact. The National Land Cover Database (NLCD) has been the authoritative source for consistent U.S. land cover mapping. To extend NLCD’s temporal resolution and reduce production latency, we developed the Land Cover Artificial Mapping System (LCAMS)—a prototype spatiotemporal deep learning framework piloted as the foundation for the new Annual NLCD. LCAMS builds on concepts from legacy NLCD and the U.S. Geological Survey Land Change Monitoring, Assessment, and Projection (LCMAP) initiatives. It employs a loosely coupled two-stage architecture consisting of independent but functionally interdependent spatial and temporal models. Spatial models extract per-year information from Landsat data, while the temporal models refine the spatial outputs to enforce inter-annual consistency—critical for reliable land change monitoring. LCAMS produces annual 30 m resolution land cover and impervious surface outputs, with region-specific fine-tuning to generalize across diverse landscapes and temporal dynamics. Validation was conducted using an independent dataset of 1925 randomly sampled plots from five U.S. Landsat Analysis Ready Data (ARD) tiles spanning 1985-2021, selected for spatial and temporal variability. This dataset was used consistently to evaluate LCAMS, Legacy NLCD, and LCMAP. Using the NLCD legend, LCAMS achieved 72.1 ± 1.60% overall agreement, compared to 71.1 ± 1.7% agreement for Legacy NLCD. Using the LCMAP legend, LCAMS achieved 83.4 ± 1.22% agreement, compared to 84.6 ± 1.11% agreement for LCMAP. Overall, LCAMS delivers comparable accuracy while offering higher thematic resolution, longer temporal coverage, and automated production of annual 30 m CONUS land cover.

Remote Sensing of Environment

Towards global mapping of dynamic surface water extents using Sentinel-1 SAR data

We introduce a fully automated and scalable method for mapping surface water extents from single-acquisition Sentinel-1 synthetic aperture radar (SAR) imagery. This approach integrates adaptive thresholding of radiometric terrain-corrected SAR backscatter data, fuzzy-logic classification, region growing, dark land estimation, and a bimodality test to minimize false positives in low-backscattering areas and false negatives in high-backscattering areas. By combining these steps, the algorithm achieves classification accuracies exceeding 85% in detecting surface water extents across diverse environmental conditions. Accuracy was first assessed at meter scale using 52 PlanetScope scenes acquired worldwide in September–October 2019; the algorithm achieved 93% overall accuracy, 86% user's accuracy, and 94% producer's accuracy. Global robustness was then evaluated by processing every Sentinel-1 acquisition from 1 to 12 November 2023 and cross-comparing the resulting maps with 6561 temporally matched observational products for end-users from remote sensing analysis (OPERA) dynamic surface water extent from Harmonized Landsat and Sentinel-2 (DSWx-HLS) products. This large-scale test yielded 90% user's and 94% producer's accuracies, confirming reliable performance at continental extent. Additional case studies demonstrate the algorithm's ability to handle surface water extent in sand-dominated deserts, to track seasonal amplitude in Folsom Lake (California), drought-induced loss in Cerro Prieto Reservoir (Mexico), and rapid filling of the Grand Ethiopian Renaissance Dam. These results show that the method scales across local to global domains and maintains high accuracy, providing a practical tool for near-real-time monitoring of floods, droughts, and water-resource management. Because the approach is sensor-agnostic, it can be ported to forthcoming L- and S-band missions such as NASA-ISRO synthetic aperture radar (NISAR), broadening its applicability to future hydrologic observations.

Remote Sensing of Environment