Geology ReportsSearch

USGS · 70221158

Synthesizing and analyzing long-term monitoring data: A greater sage-grouse case study

Abstract

Long-term monitoring of natural resources is imperative for increasing the understanding of ecosystem processes, services, and how to manage those ecosystems to maintain or improve function. Challenges with using these data may occur because methods of monitoring changed over time, multiple organizations collect and manage data differently, and monetary resources fluctuate, affecting many aspects of data. Because many species respond to changes in habitat conditions and predator-prey relationships across different spatial scales that span management boundaries, greater efforts for collaborating are essential. We demonstrate the challenges and methods for standardizing greater sage-grouse ( Centrocercus urophasianus ) long-term monitoring data across the species range in the western United States to inform population modeling needs identified by the Western Association of Fish and Wildlife Agencies. We used automated and repeatable methods of standardizing data via custom open-source software ( grsg_lekdb ) to improve the scientific integrity of future sage-grouse population assessments within and among states. Data standardization included reconciling uses of different terminology and expunging unusable data, resulting in the removal of 26% of data records due to database insertion errors and modifications to >1 million values to correct formatting and typing errors. Our approaches maximized the inclusion of usable data and identified data that could inform detection probabilities, population trends, and monitoring guidelines. Using sage-grouse databases as an example, we identified the importance of data management and how quality assurance and quality control measures can improve the usefulness of these data for future research needs. Our methods of using informatics and concluding recommendations can support similar endeavors of flora and fauna monitoring programs, whether those efforts are to use existing data or support new monitoring programs.

Explore related subjects

90° N90° S · 180° W ← longitude → 180° E
Source-reported bounding extent: 35.28150065789119° to 49.1242192485914° latitude; -125.20019531249999° to -103.9306640625° longitude. This indicates report coverage, not an exact sampling location. View area on OpenStreetMap.

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Michael S. O’Donnell, David R. Edmunds, Cameron L. Aldridge, Julie A. Heinrichs, Adrian P. Monroe, Peter S. Coates, Brian G. Prochazka, Thomas J Christiansen, Steve E. Hanser, Lief A. Wiechman, Avery A Cook, Shawn P. Espinosa, Lee J. Foster, Kathleen A. Griffin, Jesse L. Kolar, Katherine Miller, Ann M. Moser, Thomas E. Remington, Travis J Runia, Leslie A Schreiber, Michael A Schroeder, San J Stiver, Nyssa I Whitford, Catherine S Wightman. 2021. Synthesizing and analyzing long-term monitoring data: A greater sage-grouse case study. https://doi.org/10.1016/j.ecoinf.2021.101327

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related USGS reports

A Bayesian hierarchical modeling approach for species diversity in ecology

Species diversity is the foundation of many ecological disciplines. This metric is often approximated using species richness and evenness, even though actual richness likely exceeds observations due to imperfect sampling methods. Estimating the “true” species richness, which includes identifying the number of missing species, has intrigued ecologists for decades. We adopted a parametric model that appeared in Fisher et al. (1943), which models the numbers of individuals from different species as random samples from a negative binomial distribution, and developed a Bayesian computational approach to directly estimate the distribution model parameters. The model parameters represent species abundance and evenness, and can be used to derive species richness. We evaluated our parametric approach using (1) a simulation study and (2) three historical data sets. Furthermore, we illustrated the hierarchical modeling approach to combine data from multiple parallel studies using a biannual fishery survey data set. Our parametric model formulation is computationally efficient, and the hierarchical structure facilitates embedding diversity estimation into broader application, such as assessing spatial and temporal trends in species diversity associated with environmental stressors. Additionally, because the two parameters of the negative binomial distribution model represent species abundance and evenness of a community, this parametric approach facilitates a deeper understanding of the ecological systems under study. The negative binomial distribution model works with a wide range of species frequency distribution types. As a result, our emphasis on a parametric model can help us characterize the structure of an ecosystem and provide a greater depth of ecologically meaningful information.

Ecological Informatics

Hierarchical mixture models and high-resolution monitoring data can inform siting and operational strategies to mitigate bat fatalities at wind turbines

Bats provide critical ecosystem services, but bat fatalities due to wind energy development may imperil some bat populations. Statistical models are used to estimate the total fatalities that occur based on carcasses observed during monitoring surveys. Current models often estimate fatalities aggregated across species, time, and/or turbines, but fall short of reliably informing siting and operational collision mitigation strategies that account for species-specific fatality patterns on a fine spatiotemporal scale. We developed a hierarchical mixture model for estimating species-specific covariate effects and total fatalities per species at each turbine on weekly intervals. We applied the model to a high-resolution dataset of bat carcasses found during turbine searches across nineteen wind facilities in Iowa over two years. Our model explains species-specific variation in bat fatalities at individual wind turbines according to turbine proximity to bat habitat, turbine design specifications, seasonal trends, and weather conditions such as nightly air temperature, air pressure, and wind speed. Turbines located on the edge of wind facilities had higher fatalities, and proximity to roosting and foraging habitat accounted for variation in species-specific fatality estimates. These insights into turbine placement effects can inform siting strategies. We also discovered species-specific relationships with average nightly wind speed and air temperature, among other weather conditions, that could inform operational mitigation strategies such as smart curtailment. Our model can transform observations of carcasses found during turbine searches across multiple facilities, years, and variable search efforts into estimates of total fatalities per species associated with species-specific spatial, temporal, and environmental covariate effects.

Ecological Informatics

Two-stage approach to automatic detection with machine learning for improved surveillance of the invasive Cuban treefrog

The Cuban treefrog ( Osteopilus septentrionalis ), as an invasive species in the southern United States, presents a need for effective surveillance. Automated detection expedites processing of audio data for large-scale surveillance and monitoring programs. However, current available methods commonly used for anuran species have not been sufficient to detect Cuban treefrogs. Here, we present results from a two-stage method for automated detection that employs both cross-correlation template matching and secondary supervised learning classifiers. In the first stage, audio data are screened for initial detections using template matching, in which the detections contain both true and false positives. In the second stage, the false positives are screened out using classifier algorithms. We used this method to process 139,985 audio recordings, consisting of 596,046 total minutes, collected at 13 locations in Louisiana and Florida from 2014 to 2022. From the stage 1 template matching, we detected 83,191 Cuban treefrog signals across recordings. The stage 2 machine learning model was able to identify stage 1 false positive detections with a testing accuracy of 98.46% and a testing false positive rate of 1.116%. After pruning false positive detections, a total of 20,271 individual Cuban treefrog detections remained, distributed mainly across 3 sites in an area with known presence. Locations with presumed absence had an easily verifiable number of false positive detections ( n = 109 across all other sites). The two-stage methodology utilizing both template matching and machine learning algorithms can be integrated into wildlife surveillance or monitoring programs for species with distinctive, conserved calls as an effective way to achieve sensitive species detection with a low incidence of false positives.

Florida, Louisiana