Clustering Sparse Data With Feature Correlation With Application to Discover Subtypes in Cancer

Jipeng Qiang1,2, Wei Ding2, Marieke Kuijjer3

  • 1Department of Computer Science, Yangzhou University, Yangzhou 225127, China.

IEEE Access : Practical Innovations, Open Solutions
|November 4, 2022
PubMed
Summary

This study introduces a novel network-based similarity metric to address data sparseness in high-dimensional data. The method enhances cancer subtype discovery by analyzing feature interactions, outperforming existing approaches.

Related Concept Videos

Cancer Survival Analysis01:21

Cancer Survival Analysis

Cancer survival analysis focuses on quantifying and interpreting the time from a key starting point, such as diagnosis or the initiation of treatment, to a specific endpoint, such as remission or death. This analysis provides critical insights into treatment effectiveness and factors that influence patient outcomes, helping to shape clinical decisions and guide prognostic evaluations. A cornerstone of oncology research, survival analysis tackles the challenges of skewed, non-normally...
419
Cluster Sampling Method01:20

Cluster Sampling Method

Appropriate sampling methods ensure that samples are drawn without bias and accurately represent the population. Because measuring the entire population in a study is not practical, researchers use samples to represent the population of interest.
To choose a cluster sample, divide the population into clusters (groups) and then randomly select some of the clusters. All the members from these clusters are in the cluster sample. For example, if you randomly sample four departments from your...
12.3K
Cancers Originate from Somatic Mutations in a Single Cell02:21

Cancers Originate from Somatic Mutations in a Single Cell

Cancer arises from mutations in the critical genes that allow healthy cells to escape cell cycle regulation and acquire the ability to proliferate indefinitely. Though originating from a single mutation event in one of the originator cells, cancer progresses when the mutant cell lines continue to gain more and more mutations, and finally, become malignant. For example, chronic myelogenous leukemia (CML) develops initially as a non-lethal increase in white blood cells, which progressively...
12.5K
Adaptive Mechanisms in Cancer Cells02:53

Adaptive Mechanisms in Cancer Cells

Cancer cells accumulate genetic changes at an abnormally rapid rate due to the defects in the DNA repair mechanisms. From an evolutionary perspective, such genetic instability is advantageous for cancer development. Mutant cell lines accumulate a series of beneficial mutations that contribute to their progression into cancer.
Some of the advantages that cancer cells have on normal cells include - enhanced ability to divide without terminally differentiating, induce new blood vessel formation,...
5.9K
Cancer-Critical Genes II: Tumor Suppressor Genes01:05

Cancer-Critical Genes II: Tumor Suppressor Genes

Genes usually encode proteins necessary for the proper functioning of a healthy cell. Mutations can often cause changes to the gene expression pattern, thereby altering the phenotype.
When the function of certain critical genes, especially those involved in cell cycle regulation and cell growth signaling cascades, gets disrupted, it upsets the cell cycle progression. Such cells with unchecked cell cycles start proliferating uncontrollably and eventually develop into tumors.
Such genes that act...
7.7K
Genome-wide Association Studies-GWAS01:11

Genome-wide Association Studies-GWAS

Genome-wide association studies or GWAS are used to identify whether common SNPs are associated with certain diseases. Suppose specific SNPs are more frequently observed in individuals with a particular disease than those without the disease. In that case, those SNPs are said to be associated with the disease. Chi-square analysis is performed to check the probability of the allele likely to be associated with the disease.
GWAS does not require the identification of the target gene involved in...
14.0K