Consistency Metric
Statistical evaluation of the compatibility of multiple independent measurements uses a ratio of the observed variability to the estimated internal uncertainty. In interlaboratory comparisons, the Birge ratio determines whether the reported measurement results are statistically consistent with each other. This calculation compares the weighted deviations of the laboratory averages from the reference value to the declared uncertainty of the measurements.
The ratio is not applicable when the laboratories have used correlated transfer standards.
Discrepancy Evaluation
When several facilities calibrate identical sensors, their results should agree within their stated measurement uncertainties. Calculating the Birge ratio for these results yields a value near one when the reported uncertainties are realistic and the measurements are consistent. A value much greater than one indicates either that some laboratories have underestimated their uncertainties or that systematic errors are present.
Values below one suggest that the uncertainties are overly conservative.
Uncertainty Adjustment
Correcting for underestimated uncertainties requires scaling the standard uncertainty of the reference value. If the Birge ratio is greater than one, the reference value uncertainty is multiplied by the ratio to ensure a realistic confidence interval. This adjustment prevents users of the reference value from having unjustified confidence in the result.
The adjusted value is then used in subsequent traceability chains.
Laboratory Verification
International standard bodies rely on this calculation to validate proficiency test rounds. When the ratio is too high, the test coordinators must investigate the laboratories with the largest deviations. Identifying these discrepancies helps to maintain standards of measurement accuracy.