Related Experiment Video
Updated: Jun 9, 2026

10:16
Mining Spatial Transcriptomics Datasets using DeepSpaceDB
Published on: September 5, 2025
Trainable clustering framework for spatial transcriptomics
Riasat Azim1, Sabab Aosaf2, Swakkhar Shatabda3
1Department of Computer Science and Engineering, United International University, Dhaka, 1212, Bangladesh.
Bioinformatics Advances
|June 8, 2026
Summary
This study introduces a novel trainable clustering framework for spatial domain identification in spatial transcriptomics (ST). The method enhances understanding of tissue microenvironments by unifying multiple strategies for accurate spatial domain mapping.
Area of Science:
- Genomics
- Bioinformatics
- Computational Biology
Background:
- Spatial transcriptomics (ST) provides high-resolution tissue architecture insights by integrating gene expression with spatial data.
- Spatial domain identification is crucial for linking gene expression to tissue morphology and microenvironment analysis.
Purpose of the Study:
- To introduce a trainable clustering framework for unified spatial domain identification.
- To optimize feature learning and cluster assignments for improved spatial transcriptomics analysis.
Main Methods:
- A cohesive architecture unifying four strategies: ACT, FACT, Scatter, and Ensemble.
- Coupling autoencoder-driven feature learning with an Mclust-assisted clustering layer.
- Utilizing a trainable loss function for joint optimization of representation and cluster assignments.
Main Results:
- The framework achieves competitive accuracy across human DLPFC, mouse brain, and breast cancer datasets.
- It reliably identifies spatial domains while preserving complex tissue architecture.
- Cross-platform generalizability and robustness were evaluated using Stereo-seq and Slide-seq data.
Conclusions:
- The proposed framework offers a robust and accurate method for spatial domain identification in spatial transcriptomics.
- It advances the analysis of tissue microenvironments and cellular interactions.
- The open-source implementation facilitates broader application in biological research.
Related Concept Videos
RNA-seq
RNA sequencing, or RNA-Seq, is a high-throughput sequencing technology used to study the transcriptome of a cell. Transcriptomics helps to interpret the functional elements of a genome and identify the molecular constituents of an organism. Additionally, it also helps in understanding the development of an organism and the occurrence of diseases.
Before the discovery of RNA-seq, microarray-based methods and Sanger sequencing were used for transcriptome analysis. However, while microarray-based...
Before the discovery of RNA-seq, microarray-based methods and Sanger sequencing were used for transcriptome analysis. However, while microarray-based...
Cluster Sampling Method
Appropriate sampling methods ensure that samples are drawn without bias and accurately represent the population. Because measuring the entire population in a study is not practical, researchers use samples to represent the population of interest.
To choose a cluster sample, divide the population into clusters (groups) and then randomly select some of the clusters. All the members from these clusters are in the cluster sample. For example, if you randomly sample four departments from your...
To choose a cluster sample, divide the population into clusters (groups) and then randomly select some of the clusters. All the members from these clusters are in the cluster sample. For example, if you randomly sample four departments from your...
Improving Translational Accuracy
Base complementarity between the three base pairs of mRNA codon and the tRNA anticodon is not a failsafe mechanism. Inaccuracies can range from a single mismatch to no correct base pairing at all. The free energy difference between the correct and nearly correct base pairs can be as small as 3 kcal/ mol. With complementarity being the only proofreading step, the estimated error frequency would be one wrong amino acid in every 100 amino acids incorporated. However, error frequencies observed in...
Improving Translational Accuracy
Base complementarity between the three base pairs of mRNA codon and the tRNA anticodon is not a failsafe mechanism. Inaccuracies can range from a single mismatch to no correct base pairing at all. The free energy difference between the correct and nearly correct base pairs can be as small as 3 kcal/ mol. With complementarity being the only proofreading step, the estimated error frequency would be one wrong amino acid in every 100 amino acids incorporated. However, error frequencies observed in...
DNA Microarrays
Microarrays are high-throughput and relatively inexpensive assays that can be automated to analyze large quantities of data at a time. They are used in genome-wide studies to compare gene or protein expression under two varied conditions, such as healthy and diseased states. Microarrays consist of glass or silica slides on which probe molecules are covalently attached through surface functionalization. Most commonly, the slides are prepared through the chemisorption of silanes to silica...

