Jove
Visualize
Contact Us
JoVE
x logofacebook logolinkedin logoyoutube logo
ABOUT JoVE
OverviewLeadershipBlogJoVE Help Center
AUTHORS
Publishing ProcessEditorial BoardScope & PoliciesPeer ReviewFAQSubmit
LIBRARIANS
TestimonialsSubscriptionsAccessResourcesLibrary Advisory BoardFAQ
RESEARCH
JoVE JournalMethods CollectionsJoVE Encyclopedia of ExperimentsArchive
EDUCATION
JoVE CoreJoVE BusinessJoVE Science EducationJoVE Lab ManualFaculty Resource CenterFaculty Site
Terms & Conditions of Use
Privacy Policy
Policies

Related Concept Videos

Selected Data About Geographic Locations01:25

Selected Data About Geographic Locations

278
Geographic Information Systems (GIS) rely on two core types of data: spatial data and attribute data.Spatial DataSpatial data defines the physical location of features within a coordinate system, typically expressed in terms of latitude and longitude. It provides precise positioning for elements like roads, rivers, or buildings.Attribute DataAttribute data complements spatial data by adding descriptive information about these features. For example, a road's spatial data includes its start and...
278
DNA Packaging00:58

DNA Packaging

112.8K
Overview
112.8K
Biodiversity and Human Values01:24

Biodiversity and Human Values

17.1K
Human civilization relies on biodiversity in many ways. Sudden changes in species biodiversity result in environmental changes that can modify weather patterns and therefore human civilizations.
17.1K
Chromatin Packaging01:32

Chromatin Packaging

19.3K
Each human somatic cell contains 6 billion base pairs of DNA. Each base pair is 0.34 nm long, meaning each diploid cell contains a staggering 2 meters of DNA. This long DNA strand is packed inside a nucleus measuring only 10-20 microns in diameter with the help of specialized DNA-binding proteins called histones. Together they form a compact DNA-protein complex called chromatin. The chromatin is further compacted into higher-order structures. The highest level of compaction is achieved during...
19.3K
Chromatin Packaging02:21

Chromatin Packaging

22.2K
Each human somatic cell contains 6 billion base-pairs of DNA. Each base-pair is 0.34 nm long, which means that each diploid cell contains a staggering 2 meters of DNA. How is such a long DNA strand packed inside a nucleus measuring only 10 - 20 microns in diameter? 
The chromatin
In combination with specialized DNA binding protein called Histones, the DNA double helix forms a compact DNA: protein complex called chromatin. The chromatin itself is further compacted into higher-order...
22.2K
Chromatin Packaging02:21

Chromatin Packaging

9.8K
9.8K

You might also read

Related Articles

Articles linked to this work by shared authors, journal, and citation graph.

Sort by
Same author

Enhancing heart and lung dose reconstruction in breast cancer radiotherapy using publicly available out-of-field dose profiles.

Physica medica : PM : an international journal devoted to the applications of physics to medicine and biology : official journal of the Italian Association of Biomedical Physics (AIFB)·2026
Same author

Meta-analysis models with group structure for pleiotropy detection at gene and variant level using summary statistics from multiple datasets.

Biostatistics (Oxford, England)·2025
Same author

Neural network analysis of the contribution of psychotropic prescription sequences to the risk of non-psychiatric adverse events in bipolar and schizophrenia spectrum disorders.

Frontiers in digital health·2025
Same author

Sleep and Trajectories of Respiratory and Allergic Symptoms Between 1 and 5.5 Years of Age in the Elfe Birth Cohort.

Journal of sleep research·2025
Same author

GCPBayes pipeline: a tool for exploring pleiotropy at the gene level.

NAR genomics and bioinformatics·2023
Same author

Machine Learning-Based Urine Peptidome Analysis to Predict and Understand Mechanisms of Progression to Kidney Failure.

Kidney international reports·2023

Related Experiment Video

Updated: Feb 5, 2026

Visualization and Quantification of High-Dimensional Cytometry Data using Cytofast and the Upstream Clustering Methods FlowSOM and Cytosplore
06:01

Visualization and Quantification of High-Dimensional Cytometry Data using Cytofast and the Upstream Clustering Methods FlowSOM and Cytosplore

Published on: December 12, 2019

8.9K

VarSelLCM: an R/C++ package for variable selection in model-based clustering of mixed-data with missing values.

Matthieu Marbac1, Mohammed Sedki2

  • 1CREST, Ensai, Bruz, France.

Bioinformatics (Oxford, England)
|September 8, 2018
PubMed
Summary

VarSelLCM performs full model selection for model-based clustering, identifying relevant features and cluster numbers. It handles mixed data types and missing values, offering data imputation via mixture models.

More Related Videos

Large-scale Reconstructions and Independent, Unbiased Clustering Based on Morphological Metrics to Classify Neurons in Selective Populations
12:27

Large-scale Reconstructions and Independent, Unbiased Clustering Based on Morphological Metrics to Classify Neurons in Selective Populations

Published on: February 15, 2017

7.4K
Calculating Heart Rate Variability from ECG Data from Youth with Cerebral Palsy During Active Video Game Sessions
08:12

Calculating Heart Rate Variability from ECG Data from Youth with Cerebral Palsy During Active Video Game Sessions

Published on: June 5, 2019

20.5K

Related Experiment Videos

Last Updated: Feb 5, 2026

Visualization and Quantification of High-Dimensional Cytometry Data using Cytofast and the Upstream Clustering Methods FlowSOM and Cytosplore
06:01

Visualization and Quantification of High-Dimensional Cytometry Data using Cytofast and the Upstream Clustering Methods FlowSOM and Cytosplore

Published on: December 12, 2019

8.9K
Large-scale Reconstructions and Independent, Unbiased Clustering Based on Morphological Metrics to Classify Neurons in Selective Populations
12:27

Large-scale Reconstructions and Independent, Unbiased Clustering Based on Morphological Metrics to Classify Neurons in Selective Populations

Published on: February 15, 2017

7.4K
Calculating Heart Rate Variability from ECG Data from Youth with Cerebral Palsy During Active Video Game Sessions
08:12

Calculating Heart Rate Variability from ECG Data from Youth with Cerebral Palsy During Active Video Game Sessions

Published on: June 5, 2019

20.5K

Area of Science:

  • Machine Learning
  • Statistical Modeling
  • Bioinformatics

Background:

  • Model-based clustering is crucial for data analysis, but selecting relevant features and the optimal number of clusters remains challenging.
  • Existing methods often require pre-processing for mixed data types and missing values, limiting their applicability.

Purpose of the Study:

  • To introduce VarSelLCM, a novel R package for comprehensive model selection in model-based clustering.
  • To provide a robust solution for handling diverse data types and missing values within a unified clustering framework.

Main Methods:

  • VarSelLCM employs classical information criteria for full model selection, encompassing feature selection and cluster number determination.
  • The package intrinsically manages continuous, integer, and categorical data, including missing values under a Missing Completely At Random (MCAR) assumption.
  • Mixture models are utilized for integrated data imputation and clustering.

Main Results:

  • VarSelLCM enables simultaneous detection of relevant features and the optimal number of clusters.
  • The method effectively handles datasets with mixed data types and missing values without pre-processing.
  • The integrated data imputation capability enhances the robustness of clustering results.

Conclusions:

  • VarSelLCM offers a powerful and flexible tool for advanced model-based clustering.
  • Its ability to perform full model selection and manage complex data structures makes it valuable for various scientific applications.