Skip to content

Estimating Repeatability and Reproducibility with Limited Replications

Conference/Workshop:
Joint Statistical Meetings (JSM)
Published: 2020
Primary Author: Hina Arora
Secondary Authors: Naomi Kaplan-Damary, Hal Stern
Research Area: Forensic Statistics

In many measurement settings, it is important to assess the reliability and validity of measurements. As an example, forensic examiners are called upon to assess the quality of forensic evidence and draw conclusions about the evidence (e.g., whether two fingerprints came from the same source). Reliability and validity are often assessed through “black box” studies in which examiners make judgments regarding evidence of known origin under conditions meant to imitate real investigation. An open question is whether examiners differ in their ability to assess different items of evidence, i.e., whether there are examiner-by-evidence interactions. For logistical and cost reasons it is not practical to obtain a full set of replicate measurements. We leverage a hierarchical Bayesian analysis of variance model to address this limitation and simultaneously explain the variation in the decisions both between different examiners (reproducibility) and within an examiner (repeatability). The model can be applied to continuous, binary or ordinal data. Simulation studies demonstrate the approach and the methods are applied to data from handwriting and latent print examinations.

Related Resources

Score-based Likelihood Ratios Using Stylometric Text Embeddings

Score-based Likelihood Ratios Using Stylometric Text Embeddings

We consider the problem setting in which we have two sets of texts in digital form and would like to quantify our beliefs that the two sets of texts were…
Statistics and its Applications in Forensic Science and the Criminal Justice System

Statistics and its Applications in Forensic Science and the Criminal Justice System

This presentation is from the 2024 Joint Statistical Meetings (JSM), Portland, Oregon, August 3-8, 2024.
Algorithmic matching of striated tool marks

Algorithmic matching of striated tool marks

Automatic matching algorithms for assessing the similarity between striation marks have been investigated for bullet lands and some tool marks, such as screwdrivers. We are interested in the investigation of…
Silencing the Defense Expert

Silencing the Defense Expert

In the wake of the 2009 NRC and 2016 PCAST Reports, the Firearms and Toolmark (FATM) discipline has come under increasing scrutiny. Validation studies like AMES I, Keisler, AMES II,…