Geology ReportsSearch

USGS · 70203561

Mapping cropland extent of Southeast and Northeast Asia using multi-year time-series Landsat 30-m data using Random Forest classifier on Google Earth Engine

Abstract

Cropland extent maps are useful components for assessing food security. Ideally, such products are a useful addition to countrywide agricultural statistics since they are not politically biased and can be used to calculate cropland area for any spatial unit from an individual farm to various administrative unites (e.g., state, county, district) within and across nations, which in turn can be used to estimate agricultural productivity as well as degree of disturbance on food security from natural disasters and political conflict. However, existing cropland extent maps over large areas (e.g., Country, region, continent, world) are derived from coarse resolution imagery (250 m to 1 km pixels) and have many limitations such as missing fragmented and\or small farms with mixed signatures from different crop types and\or farming practices that can be, confused with other land cover. As a result, the coarse resolution maps have limited useflness in areas where fields are small (<1 ha), such as in Southeast Asia. Furthermore, coarse resolution cropland maps have known uncertainties in both geo-precision of cropland location as well as accuracies of the product. To overcome these limitations, this research was conducted using multi-date, multi-year 30-m Landsat time-series data for 3 years chosen from 2013 to 2016 for all Southeast and Northeast Asian Countries (SNACs), which included 7 refined agro-ecological zones (RAEZ) and 12 countries (Indonesia, Thailand, Myanmar, Vietnam, Malaysia, Philippines, Cambodia, Japan, North Korea, Laos, South Korea, and Brunei). The 30-m (1 pixel = 0.09 ha) data from Landsat 8 Operational Land Imager (OLI) and Landsat 7 Enhanced Thematic Mapper (ETM+) were used in the study. Ten Landsat bands were used in the analysis (blue, green, red, NIR, SWIR1, SWIR2, Thermal, NDVI, NDWI, LSWI) along with additional layers of standard deviation of these 10 bands across 1 year, and global digital elevation model (GDEM)-derived slope and elevation bands. To reduce the impact of clouds, the Landsat imagery was time-composited over four time-periods (Period 1: January- April, Period 2: May-August, and Period 3: September-December) over 3-years. Period 4 was the standard deviation of all 10 bands taken over all images acquired during the 2015 calendar year. These four period composites, totaling 42 band data-cube, were generated for each of the 7 RAEZs. The reference training data (N = 7849) generated for the 7 RAEZ using sub-meter to 5-m very high spatial resolution imagery (VHRI) helped generate the knowledge-base to separate croplands from non-croplands. This knowledge-base was used to code and run a pixel-based random forest (RF) supervised machine learning algorithm on the Google Earth Engine (GEE) cloud computing environment to separate croplands from non-croplands. The resulting cropland extent products were evaluated using an independent reference validation dataset (N = 1750) in each of the 7 RAEZs as well as for the entire SNAC area. For the entire SNAC area, the overall accuracy was 88.1% with a producer’s accuracy of 81.6% (errors of omissions = 18.4%) and user’s accuracy of 76.7% (errors of commissions = 23.3%). For each of the 7 RAEZs overall accuracies varied from 83.2 to 96.4%. Cropland areas calculated for the 12 countries were compared with country areas reported by the United Nations Food and Agriculture Organization and other national cropland statistics resulting in an R 2 value of 0.93. The cropland areas of provinces were compared with the province statistics that showed an R 2 = 0.95 for South Korea and R 2 = 0.94 for Thailand. The cropland products are made available on an interactive viewer at www.croplands.org and for download at National Aeronautics and Space Administration’s (NASA) Land Processes Distributed Active Archive Center (LP DAAC): https://lpdaac.usgs.gov/node/1281 .

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Adam Oliphant, Prasad S. Thenkabail, Pardhasaradhi Teluguntla, Jun Xiong, Murali Krishna Gumma, Russell G. Congalton, Kamini Yadav. 2019. Mapping cropland extent of Southeast and Northeast Asia using multi-year time-series Landsat 30-m data using Random Forest classifier on Google Earth Engine. https://doi.org/10.1016/j.jag.2018.11.014

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related USGS reports

Satellite embeddings for crop type classification: A comparative examination

Embedding datasets encode complex relationships among multiple sources of Earth observation data into a compact format. Here, we evaluated the utility of a 10-m global Satellite Embedding product (SE) for classifying crop types in central California for the year 2020. We compared the classification accuracy of a random forest model based exclusively on the SE layer to an existing random forest model with multiple imagery inputs. Our results showed the SE-based classification had higher agreement with the reference dataset (California Department of Water Resources crop map) than the classification based on Landsat and National Agricultural Imagery Program inputs (94.7% versus. 91.9% overall accuracy, respectively). The performance of individual crop types was consistent across models, ranging from high agreement for rice (98.4% versus 98% accuracy) to lower agreement for pasture, grain, and fallow/young perennial classes (< 65% accuracy in both models). The SE-based workflow used three times less cloud-based computational resources and represented substantial savings of predictor development time. Geospatial embedding products can aid classification efforts by reducing predictor development time and processing demands while maintaining classification accuracy.

California

The impacts of cover crop biomass on satellite-based detectability of cover crops in Maryland

Cover crop adoption in the U.S. has increased over the past decades, increasing the need to quantify their performance and environmental benefits. While remote sensing (RS)-based approaches for detecting cover crop presence have been developed, there has been limited research on how cover crops with varied biomass and management practices influence detectability. Using unique field-level cover crop presence and biomass datasets in Maryland, U.S., we investigated how RS-based detectability changes for cover crops with varied aboveground biomass, planting, and termination dates. Specifically, we proposed a time-integrated satellite-based greenness feature from Harmonized Landsat-8 and Sentinel-2 (HLS) time series from 2017 to 2021 to estimate biomass of cover crops and evaluated their detectability using a phenology-based cover crop detection framework. The impacts of cover crop planting and termination dates on cover crop biomass were also analyzed. Our results demonstrated that Normalized Difference Vegetation Index (NDVI) estimated biomass with higher accuracy compared to other vegetation indices, and the time-integrated model estimated biomass with higher accuracy than the single-date “snapshot” linear model (R2 from 0.53 to 0.66 and RMSE from 922 kg/ha to 747 kg/ha). While the snapshot models were species sensitive, the time-integrated models showed strong robustness across different species. Detectability increased with cover crop biomass, as detected cover crops averaged 963.3 ±719.5 kg/ha compared to 297.2 ± 209.0 kg/ha for non-detected cover crops. Detection accuracy reached 96.1% for fields exceeding 500 kg/ha, compared with an overall accuracy of 62.7%. Earlier planting and later termination increased biomass and detectability, with biomass rising by 4.14 kg/ha/day (p < 0.01). This study demonstrates how management practices affect cover crop biomass and detectability via satellite time series and provides insights that can inform management of cover crops and monitoring of their effects on agroecosystems.

Maryland

Aligning legacy NLCD land cover maps based on Landsat Collection 1 to Collection 2

The transition from Landsat Collection 1 to Collection 2 introduced significant improvements in radiometric and geometric accuracy. However, the improvements cause location misalignment between the existing Landsat-derived land cover products and the new collection. The legacy National Land Cover Database (NLCD) has been used as a cornerstone land cover source for a variety of research. Therefore, a method aligning the legacy NLCD product to Collection 2 is required to ensure its continuity and consistency of service. We developed a strategy to not only align legacy NLCD to match new Collection 2 geometric locations but also improve land cover labeling in the region that was affected by the geometric shifts. The method identifies boundary pixels of homogeneous land cover patches as potential problem areas that are likely impacted by geometric shifts and generates candidate labels from 3 × 3 window with the target pixel at the center and segmentation-derived majority label. Standard phenology patterns of each candidate land cover type are established based on the random samples except boundary pixels within a 1000-pixels × 1000-pixels processing window region. The phenological distance to each standard land cover type pattern is calculated through a penalty dynamic time warping (DTW) method for each target pixel in the boundary region. Finally, the method determines the most suitable label based on the phenological distance from the candidate labels. Both visual and accuracy assessment results demonstrate that the alignment preserves the overall land cover patterns in the original legacy NLCD product while reducing the spatial discrepancies between the Landsat Collection 2 and land cover. In addition, it enhances the accuracy of land cover labeling of boundary pixels. The overall accuracy (OA) was increased by 7% in the land cover boundary regions after alignment. The quality and confusion matrix comparison between the alignment results and the original legacy NLCD confirm the reliability of the method. Our alignment method has the potential to serve as a framework for aligning other Landsat-derived land cover products to future collections.

conterminous United States