Related Experiment Video
Updated: Feb 1, 2026

An Organotypic High Throughput System for Characterization of Drug Sensitivity of Primary Multiple Myeloma Cells
Published on: July 15, 2015
Development of an Algorithm to Distinguish Smoldering Versus Symptomatic Multiple Myeloma in Claims-Based Data Sets
Mark A Fiala1, James Dukeman1, Sascha A Tuchman2
1Washington University School of Medicine, St Louis, MO.
Purpose:
The distinction of patients with symptomatic multiple myeloma (MM) from those with smoldering MM poses a challenge for researchers who use administrative databases. Historically, researchers either have included all patients or used treatment receipt as the distinguishing factor; both methods have drawbacks. We present an algorithm for distinguishing between symptomatic and smoldering MM using ICD-9-CM (International Classification of Diseases, Ninth Revision, Clinical Modification) codes for the classic defining events of symptomatic MM commonly referred to as the CRAB criteria (hypercalcemia, renal impairment, anemia, and bone lesions).
Patients And Methods:
SEER-Medicare-linked data from 4,187 patients with MM diagnosed between 2007 and 2011 were used for this analysis.
Results:
Eighty-four percent had ICD-9-CM codes consistent with CRAB criteria, whereas only 57% received treatment. Overall survival of patients with symptomatic MM defined as receipt of treatment was 32.3 months versus 26.6 months for the overall population and 22.9 months for patients with symptomatic MM defined by CRAB criteria. Conceptually, removal of patients with smoldering MM should result in a reduction in overall survival; however, the cohort of patients who received treatment tended to be younger and healthier than the overall population, which could have skewed the results.
Conclusion:
The algorithm we present resulted in a larger and more representative sample than classification by treatment status and reduced potential bias that could result from including all patients with smoldering MM in the analysis. Although this study was performed using the SEER-Medicare database, the methodology was broad enough that the algorithm could be extended to additional claims-based data sets with relative ease.
Related Concept Videos
Design Example: Setting a Curve Using Design Data
Testing a Claim about Mean: Known Population SD
Estimating a population mean requires the samples to be distributed normally. The data should be collected from the randomly selected samples having no sampling bias. The sample size needed to be higher than 30, and most importantly, the population standard deviation should be already known.
In most realistic situations, the population standard deviation is often unknown, but in rare circumstances, when it...
Testing a Claim about Population Proportion
There are two methods of testing a claim about a population proportion: (1) Using the sample proportion from the data where a binomial distribution is approximated to the normal distribution and (2) Using the binomial probabilities calculated from the data.
The first method uses normal distribution as an approximation to the binomial distribution. The requirements are as follows: sample size is large...
Testing a Claim about Standard Deviation
The hypothesis testing for the claim of population standard deviation (or variance) requires the data and samples to be random and unbiased. The population distribution also must be normal. There is no specific requirement on the sample size as the estimation is based on the chi-square distribution.
As a first step, the hypothesis (null and alternative) concerning the claim about...
Testing a Claim about Mean: Unknown Population SD
Estimating a population mean requires the samples to be approximately normally distributed. The data should be collected from the randomly selected samples having no sampling bias. There is no specific requirement for sample size. But if the sample size is less than 30, and we don't know the population standard deviation, a different approach is used;...
Trial and Error and Algorithm

