Related Experiment Video
Updated: Jan 11, 2026

06:40
A Robust Method for the Large-Scale Production of Spheroids for High-Content Screening and Analysis Applications
Published on: December 28, 2021
3.8K
Multiscale Hyperbolic Embedding for Cell Hierarchies in Large-Scale Bioinformatics Data
Mingchen Yao1,2, Anoop Praturu1,2, Tatyana Sharpee1,2
1Computational Neurobiology Laboratory, Salk Institute for Biological Studies, La Jolla, CA 92037, USA.
Biorxiv : the Preprint Server for Biology
|November 19, 2025
Summary
We developed MuH-MDS, a scalable hyperbolic multidimensional scaling method for analyzing large datasets. This novel approach efficiently captures complex hierarchical structures, outperforming existing methods in computational speed and accuracy for biological data analysis.
Area of Science:
- Computational Biology
- Data Visualization
- Machine Learning
Background:
- Large datasets present challenges for visualization and interpretation.
- Hyperbolic embedding excels at capturing hierarchical structures but struggles with scalability.
- Existing methods lack efficient scaling for massive datasets.
Purpose of the Study:
- Introduce MuH-MDS, a multiscale hyperbolic multidimensional scaling algorithm.
- Address limitations of fixed curvature and scalability in current hyperbolic embedding techniques.
- Enhance analysis of large-scale biological datasets.
Main Methods:
- Developed MuH-MDS, a novel multiscale hyperbolic multidimensional scaling algorithm.
- Utilized "adiabatic" approximation from physics for optimizing local positions.
- Implemented a method capable of handling datasets with over 80,000 samples.
Main Results:
- MuH-MDS demonstrates a 10^3 improvement in computing time compared to previous methods.
- Successfully analyzed a large-scale C. elegans embryogenesis scRNA-seq dataset (>80,000 samples).
- Uncovered intrinsic hierarchical structure, improving pseudotime inference and lineage analysis over UMAP and other methods.
Conclusions:
- MuH-MDS offers a scalable and computationally efficient solution for hyperbolic embedding.
- Preserves global hierarchy and metric accuracy, unlike methods focusing solely on local structure.
- Provides superior performance for large-scale biological data analysis, particularly in lineage and pseudotime inference.

