USGS ScienceSearch

USGS · 70271501

A compilation pipeline for wildlife tracking datasets collected from ground-based and satellite-based telemetry transmission devices

Abstract

Wildlife conservation planning increasingly requires collaboration and integration of research from discrete studies spanning large geographic areas. Tracking datasets are essential for analyzing animal movements and species distributions in relation to environmental conditions and combining them can enable powerful analyses to further aid planning efforts. However, combining datasets necessitates addressing variation in study designs, tracking methodologies, location uncertainty, and data attributes. We outline a compilation pipeline to integrate ground-based and satellite-based telemetry tracking datasets, motivated from our work with greater sage-grouse ( Centrocercus urophasianus ), a highly imperiled species of western North America. Our objective was to create a database with a standardized set of attributes to facilitate filtering locations for spatial analyses. Our pipeline phases are: (1) dataset pre-processing, (2) formatting individual datasets to a common template, (3) dataset binding, (4) error checking, and (5) filtering. Our pipeline includes additional functionality to identify coordinates from recurrently visited locations (e.g., nest sites), which may be of special interest. The final compiled sage-grouse database included nearly 5 million locations collected from 53 datasets and over 19,000 birds tracked from 1980 to 2022, including over 11,000 nest locations. Our error checks flagged 3.9 % of locations as likely errors, predominantly collected from satellite-based telemetry transmissions. We demonstrate the ability of our pipeline to identify nest locations and flag erroneous locations by applying it to simulated tracking datasets. Overall, our workflow offers a transferable approach for researchers aiming to standardize wildlife telemetry datasets and conduct ecological analyses for both individual studies and large-scale collaborations.

Explore related subjects

90° N90° S · 180° W ← longitude → 180° E
Source-reported bounding extent: 36.14719079486781° to 48.97073706473478° latitude; -122.65581494478303° to -102.5194136320973° longitude. This indicates report coverage, not an exact sampling location. View area on OpenStreetMap.

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Gregory T. Wann, Ashley L. Whipple, Michael S. O’Donnell, Cameron L. Aldridge. 2025. A compilation pipeline for wildlife tracking datasets collected from ground-based and satellite-based telemetry transmission devices. https://doi.org/10.1016/j.ecoinf.2025.103220

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related USGS reports

A Bayesian hierarchical modeling approach for species diversity in ecology

Species diversity is the foundation of many ecological disciplines. This metric is often approximated using species richness and evenness, even though actual richness likely exceeds observations due to imperfect sampling methods. Estimating the “true” species richness, which includes identifying the number of missing species, has intrigued ecologists for decades. We adopted a parametric model that appeared in Fisher et al. (1943), which models the numbers of individuals from different species as random samples from a negative binomial distribution, and developed a Bayesian computational approach to directly estimate the distribution model parameters. The model parameters represent species abundance and evenness, and can be used to derive species richness. We evaluated our parametric approach using (1) a simulation study and (2) three historical data sets. Furthermore, we illustrated the hierarchical modeling approach to combine data from multiple parallel studies using a biannual fishery survey data set. Our parametric model formulation is computationally efficient, and the hierarchical structure facilitates embedding diversity estimation into broader application, such as assessing spatial and temporal trends in species diversity associated with environmental stressors. Additionally, because the two parameters of the negative binomial distribution model represent species abundance and evenness of a community, this parametric approach facilitates a deeper understanding of the ecological systems under study. The negative binomial distribution model works with a wide range of species frequency distribution types. As a result, our emphasis on a parametric model can help us characterize the structure of an ecosystem and provide a greater depth of ecologically meaningful information.

Ecological Informatics

Hierarchical mixture models and high-resolution monitoring data can inform siting and operational strategies to mitigate bat fatalities at wind turbines

Bats provide critical ecosystem services, but bat fatalities due to wind energy development may imperil some bat populations. Statistical models are used to estimate the total fatalities that occur based on carcasses observed during monitoring surveys. Current models often estimate fatalities aggregated across species, time, and/or turbines, but fall short of reliably informing siting and operational collision mitigation strategies that account for species-specific fatality patterns on a fine spatiotemporal scale. We developed a hierarchical mixture model for estimating species-specific covariate effects and total fatalities per species at each turbine on weekly intervals. We applied the model to a high-resolution dataset of bat carcasses found during turbine searches across nineteen wind facilities in Iowa over two years. Our model explains species-specific variation in bat fatalities at individual wind turbines according to turbine proximity to bat habitat, turbine design specifications, seasonal trends, and weather conditions such as nightly air temperature, air pressure, and wind speed. Turbines located on the edge of wind facilities had higher fatalities, and proximity to roosting and foraging habitat accounted for variation in species-specific fatality estimates. These insights into turbine placement effects can inform siting strategies. We also discovered species-specific relationships with average nightly wind speed and air temperature, among other weather conditions, that could inform operational mitigation strategies such as smart curtailment. Our model can transform observations of carcasses found during turbine searches across multiple facilities, years, and variable search efforts into estimates of total fatalities per species associated with species-specific spatial, temporal, and environmental covariate effects.

Ecological Informatics

Two-stage approach to automatic detection with machine learning for improved surveillance of the invasive Cuban treefrog

The Cuban treefrog ( Osteopilus septentrionalis ), as an invasive species in the southern United States, presents a need for effective surveillance. Automated detection expedites processing of audio data for large-scale surveillance and monitoring programs. However, current available methods commonly used for anuran species have not been sufficient to detect Cuban treefrogs. Here, we present results from a two-stage method for automated detection that employs both cross-correlation template matching and secondary supervised learning classifiers. In the first stage, audio data are screened for initial detections using template matching, in which the detections contain both true and false positives. In the second stage, the false positives are screened out using classifier algorithms. We used this method to process 139,985 audio recordings, consisting of 596,046 total minutes, collected at 13 locations in Louisiana and Florida from 2014 to 2022. From the stage 1 template matching, we detected 83,191 Cuban treefrog signals across recordings. The stage 2 machine learning model was able to identify stage 1 false positive detections with a testing accuracy of 98.46% and a testing false positive rate of 1.116%. After pruning false positive detections, a total of 20,271 individual Cuban treefrog detections remained, distributed mainly across 3 sites in an area with known presence. Locations with presumed absence had an easily verifiable number of false positive detections ( n = 109 across all other sites). The two-stage methodology utilizing both template matching and machine learning algorithms can be integrated into wildlife surveillance or monitoring programs for species with distinctive, conserved calls as an effective way to achieve sensitive species detection with a low incidence of false positives.

Florida, Louisiana