Related Experiment Video
Updated: May 15, 2026

An Integrated Approach for Microprotein Identification and Sequence Analysis
Published on: July 12, 2022
Clumppling 2.0: A Clustering Alignment Program for Population Structure Analyses
Xiran Liu1, Noah A Rosenberg2, Sohini Ramachandran1,3
1Data Science Institute, Brown University, Providence, RI, 02912, USA.
Abstract:
We previously introduced Clumppling to address the "alignment problem" for multiple mixed-membership unsupervised clustering results in population structure analyses, where clusters represent latent genetic ancestries. This problem stems from three challenges-label-switching, multi-modality, and varying numbers of clusters-which Clumppling resolves in three steps: aligning results with the same number of clusters, detecting distinct solutions or "modes," and aligning modes across different numbers of clusters. Here, we present Clumppling 2.0, an update with features for visualizing the emergence of clusters, comparing aligned results from different models, and incorporating modularity of algorithmic steps. We outline the Clumppling 2.0 workflow, highlighting its improved algorithmic flexibility and visual interpretability through a graph of alignment patterns. We then demonstrate its utility on human genetic datasets that include individuals from admixed populations.
Related Concept Videos
RNA-seq
Before the discovery of RNA-seq, microarray-based methods and Sanger sequencing were used for transcriptome analysis. However, while microarray-based...
Cluster Sampling Method
To choose a cluster sample, divide the population into clusters (groups) and then randomly select some of the clusters. All the members from these clusters are in the cluster sample. For example, if you randomly sample four departments from your...

