Search USGSSearch

USGS · 5224963

Uncovering a latent multinomial: Analysis of mark–recapture data with misidentification

Also available from

Abstract

Natural tags based on DNA fingerprints or natural features of animals are now becoming very widely used in wildlife population biology. However, classic capture-recapture models do not allow for misidentification of animals which is a potentially very serious problem with natural tags. Statistical analysis of misidentification processes is extremely difficult using traditional likelihood methods but is easily handled using Bayesian methods. We present a general framework for Bayesian analysis of categorical data arising from a latent multinomial distribution. Although our work is motivated by a specific model for misidentification in closed population capture-recapture analyses, with crucial assumptions which may not always be appropriate, the methods we develop extend naturally to a variety of other models with similar structure. Suppose that observed frequencies f are a known linear transformation f=A'x of a latent multinomial variable x with cell probability vector pi= pi(theta). Given that full conditional distributions [theta | x] can be sampled, implementation of Gibbs sampling requires only that we can sample from the full conditional distribution [x | f, theta], which is made possible by knowledge of the null space of A'. We illustrate the approach using two data sets with individual misidentification, one simulated, the other summarizing recapture data for salamanders based on natural marks.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

W.A. Link, J. Yoshizaki, L.L. Bailey, K. H. Pollock. 2010-03-17. Uncovering a latent multinomial: Analysis of mark–recapture data with misidentification. https://doi.org/10.1111/j.1541-0420.2009.01244.x

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related USGS reports

A flexible framework for N-mixture occupancy models: Applications to breeding bird surveys

Estimating species abundance under imperfect detection is a key challenge in biodiversity conservation. The N -mixture model, widely recognized for its ability to distinguish between abundance and individual detection probability without marking individuals, is constrained by its stringent closure assumption, which leads to biased estimates when violated in real-world settings. To address this limitation, we propose an extended framework based on a development of the mixed Gamma-Poisson model, incorporating a community parameter that represents the proportion of individuals consistently present throughout the survey period. This flexible framework generalizes both the zero-inflated type occupancy model and the standard N -mixture model as special cases, corresponding to community parameter values of 0 and 1, respectively. The model’s effectiveness is validated through simulations and applications to real-world datasets, specifically with 5 species from the North American Breeding Bird Survey and 46 species from the Swiss Breeding Bird Survey, demonstrating its improved accuracy and adaptability in settings where strict closure may not hold.

Biometrics

Multivariate Bayesian clustering using covariate-informed components with application to boreal vegetation sensitivity

Climate change is impacting both the distribution and abundance of vegetation, especially in far northern latitudes. The effects of climate change are different for every plant assemblage and vary heterogeneously in both space and time. Small changes in climate could result in large vegetation responses in sensitive assemblages but weak responses in robust assemblages. But, patterns and mechanisms of sensitivity and robustness are not yet well understood, largely due to a lack of long-term measurements of climate and vegetation. Fortunately, observations are sometimes available across a broad spatial extent. We develop a novel statistical model for a multivariate response based on unknown cluster-specific effects and covariances, where cluster labels correspond to sensitivity and robustness. Our approach utilizes a prototype model for cluster membership that offers flexibility while enforcing smoothness in cluster probabilities across sites with similar characteristics. We demonstrate our approach with an application to vegetation abundance in Alaska, USA, in which we leverage the broad spatial extent of the study area as a proxy for unrecorded historical observations. In the context of the application, our approach yields interpretable site-level cluster labels associated with assemblage-level sensitivity and robustness without requiring strong a priori assumptions about the drivers of climate sensitivity.

Alaska

A temporally stratified extension of space‐for‐time Cormack–Jolly–Seber for migratory animals

Understanding drivers of temporal variation in demographic parameters is a central goal of mark‐recapture analysis. To estimate the survival of migrating animal populations in migration corridors, space‐for‐time mark–recapture models employ discrete sampling locations in space to monitor marked populations as they move past monitoring sites, rather than the standard practice of using fixed sampling points in time. Because these models focus on estimating survival over discrete spatial segments, model parameters are implicitly integrated over the temporal dimension. Furthermore, modeling the effect of time‐varying covariates on model parameters is complicated by unknown passage times for individuals that are not detected at monitoring sites. To overcome these limitations, we extended the Cormack–Jolly–Seber (CJS) framework to estimate temporally stratified survival and capture probabilities by including a discretized arrival time process in a Bayesian framework. We allow for flexibility in the model form by including temporally stratified covariates and hierarchical structures. In addition, we provide tools for assessing model fit and comparing among alternative structural models for the parameters. We demonstrate our framework by fitting three competing models to estimate daily survival, capture, and arrival probabilities at four hydroelectric dams for over 200 000 individually tagged migratory juvenile salmon released into the Snake River, USA.

Biometrics