Related Experiment Video
Updated: Apr 19, 2026

Performing Data Mining And Integrative Analysis Of Biomarker in Breast Cancer Using Multiple Publicly Accessible Databases
Published on: May 17, 2019
Reference intervals data mining: no longer a probability paper method
Alexander Katayev1, James K Fleming2, Dajie Luo2
1From Laboratory Corporation of America Holdings, Elon, NC. katayea@labcorp.com.
Objectives:
To describe the application of a data-mining statistical algorithm for calculation of clinical laboratory tests reference intervals.
Methods:
Reference intervals for eight different analytes and different age and sex groups (a total of 11 separate reference intervals) for tests that are unlikely to be ordered during routine screening of disease-free populations were calculated using the modified algorithm for data mining of test results stored in the laboratory database and compared with published peer-reviewed studies that used direct sampling. The selection of analytes was based on the predefined criteria that include comparability of analytical methods with a statistically significant number of observations.
Results:
Of the 11 calculated reference intervals, having upper and lower limits for each, 21 of 22 reference interval limits were not statistically different from the reference studies.
Conclusions:
The presented statistical algorithm is shown to be an accurate and practical tool for reference interval calculations.
Related Concept Videos
Interval Level of Measurement
Data measured using the interval scale are similar to ordinal level data because they have a definite arrangement. However, in the interval level of measurement, the differences between data values are meaningful even though the data does not have a starting point.
Temperature is measured using the interval scale. It is measurable data, and the difference between...
Quartile
1; 1; 2; 2; 4; 6; 6.8; 7.2; 8; 8.3; 9; 10; 10; 11.5
The median or second quartile is seven. The lower half of the...
Midrange
Simply put, the midrange is half of the data set’s range. Similar to the mean, the midrange is sensitive to the extreme values and hence the prospective outliers. However, unlike the mean, the midrange is not sensitive to all the values of the data set that lie in the middle. Thus, it is prone to...
Prediction Intervals
However, the point estimate is most likely not the exact value of the population parameter, but close to it. After calculating point estimates, we construct interval estimates, called confidence intervals or prediction intervals. This prediction interval comprises a range of values unlike the point estimate and is a better predictor of the observed sample value, y.
Confidence Intervals
A confidence...
Interpretation of Confidence Intervals
Confidence intervals have confidence coefficients that are crucial for their interpretation. The most common confidence coefficients are 0.90, 0.95, and 0.99, which can be written as percentages–90%, 95%, and 99%, respectively.
Suppose a person calculates a confidence interval with a confidence coefficient of 0.95. In that case, they can...
