USGS ScienceSearch

USGS · 70198028

On the reliability of N‐mixture models for count data

Abstract

N‐mixture models describe count data replicated in time and across sites in terms of abundance N and detectability p . They are popular because they allow inference about N while controlling for factors that influence p without the need for marking animals. Using a capture–recapture perspective, we show that the loss of information that results from not marking animals is critical, making reliable statistical modeling of N and p problematic using just count data. One cannot reliably fit a model in which the detection probabilities are distinct among repeat visits as this model is overspecified. This makes uncontrolled variation in p problematic. By counter example, we show that even if p is constant after adjusting for covariate effects (the “constant p ” assumption) scientifically plausible alternative models in which N (or its expectation) is non‐identifiable or does not even exist as a parameter, lead to data that are practically indistinguishable from data generated under an N‐mixture model. This is particularly the case for sparse data as is commonly seen in applications. We conclude that under the constant p assumption reliable inference is only possible for relative abundance in the absence of questionable and/or untestable assumptions or with better quality data than seen in typical applications. Relative abundance models for counts can be readily fitted using Poisson regression in standard software such as R and are sufficiently flexible to allow controlling for p through the use covariates while simultaneously modeling variation in relative abundance. If users require estimates of absolute abundance, they should collect auxiliary data that help with estimation of p .

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Richard J. Barker, Matthew J. Schofield, William A. Link, John R. Sauer. 2017-07-03. On the reliability of N‐mixture models for count data. https://doi.org/10.1111/biom.12734

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related USGS reports

A flexible framework for N-mixture occupancy models: Applications to breeding bird surveys

Estimating species abundance under imperfect detection is a key challenge in biodiversity conservation. The N -mixture model, widely recognized for its ability to distinguish between abundance and individual detection probability without marking individuals, is constrained by its stringent closure assumption, which leads to biased estimates when violated in real-world settings. To address this limitation, we propose an extended framework based on a development of the mixed Gamma-Poisson model, incorporating a community parameter that represents the proportion of individuals consistently present throughout the survey period. This flexible framework generalizes both the zero-inflated type occupancy model and the standard N -mixture model as special cases, corresponding to community parameter values of 0 and 1, respectively. The model’s effectiveness is validated through simulations and applications to real-world datasets, specifically with 5 species from the North American Breeding Bird Survey and 46 species from the Swiss Breeding Bird Survey, demonstrating its improved accuracy and adaptability in settings where strict closure may not hold.

Biometrics

Multivariate Bayesian clustering using covariate-informed components with application to boreal vegetation sensitivity

Climate change is impacting both the distribution and abundance of vegetation, especially in far northern latitudes. The effects of climate change are different for every plant assemblage and vary heterogeneously in both space and time. Small changes in climate could result in large vegetation responses in sensitive assemblages but weak responses in robust assemblages. But, patterns and mechanisms of sensitivity and robustness are not yet well understood, largely due to a lack of long-term measurements of climate and vegetation. Fortunately, observations are sometimes available across a broad spatial extent. We develop a novel statistical model for a multivariate response based on unknown cluster-specific effects and covariances, where cluster labels correspond to sensitivity and robustness. Our approach utilizes a prototype model for cluster membership that offers flexibility while enforcing smoothness in cluster probabilities across sites with similar characteristics. We demonstrate our approach with an application to vegetation abundance in Alaska, USA, in which we leverage the broad spatial extent of the study area as a proxy for unrecorded historical observations. In the context of the application, our approach yields interpretable site-level cluster labels associated with assemblage-level sensitivity and robustness without requiring strong a priori assumptions about the drivers of climate sensitivity.

Alaska

A temporally stratified extension of space‐for‐time Cormack–Jolly–Seber for migratory animals

Understanding drivers of temporal variation in demographic parameters is a central goal of mark‐recapture analysis. To estimate the survival of migrating animal populations in migration corridors, space‐for‐time mark–recapture models employ discrete sampling locations in space to monitor marked populations as they move past monitoring sites, rather than the standard practice of using fixed sampling points in time. Because these models focus on estimating survival over discrete spatial segments, model parameters are implicitly integrated over the temporal dimension. Furthermore, modeling the effect of time‐varying covariates on model parameters is complicated by unknown passage times for individuals that are not detected at monitoring sites. To overcome these limitations, we extended the Cormack–Jolly–Seber (CJS) framework to estimate temporally stratified survival and capture probabilities by including a discretized arrival time process in a Bayesian framework. We allow for flexibility in the model form by including temporally stratified covariates and hierarchical structures. In addition, we provide tools for assessing model fit and comparing among alternative structural models for the parameters. We demonstrate our framework by fitting three competing models to estimate daily survival, capture, and arrival probabilities at four hydroelectric dams for over 200 000 individually tagged migratory juvenile salmon released into the Snake River, USA.

Biometrics