Related Experiment Video
Updated: Feb 20, 2026

12:39
A Novel Bayesian Change-point Algorithm for Genome-wide Analysis of Diverse ChIPseq Data Types
Published on: December 10, 2012
11.7K
A new alignment free genome comparison algorithm based on statistically estimated feature frequency profile
Summary
This study introduces a novel alignment-free genome comparison method using statistical analysis of word frequencies. This approach offers computational efficiency for understanding genomic properties and classifying mammalian sequences.
Area of Science:
- Bioinformatics
- Genomics
- Computational Biology
Background:
- Sequence comparison is crucial for understanding genomic biological properties.
- Traditional alignment-based methods are reliable but computationally intensive.
- Alignment-free methods are gaining traction due to their efficiency.
Purpose of the Study:
- To propose a new alignment-free genome comparison scheme.
- To leverage statistical approaches for sequence analysis.
- To numerically represent sequence characteristics through word frequency.
Main Methods:
- Estimating word frequency information from sequence components.
- Investigating the relationship between estimated and actual word frequencies.
- Developing a statistical algorithm for genome comparison.
Main Results:
- Numerical representation of sequence characteristics.
- Generation of a phylogenetic tree for mammalian sequences.
- Classification of mammalian sequences using the proposed algorithm.
Conclusions:
- The statistical alignment-free algorithm demonstrates remarkable performance.
- This method provides an efficient alternative for genome comparison.
- The approach is effective for phylogenetic analysis and sequence classification.
Related Concept Videos
Evolutionary Relationships through Genome Comparisons
7.1K
Genome comparison is one of the excellent ways to interpret the evolutionary relationships between organisms. The basic principle of genome comparison is that if two species share a common feature, it is likely encoded by the DNA sequence conserved between both species. The advent of genome sequencing technologies in the late 20th century enabled scientists to understand the concept of conservation of domains between species and helped them to deduce evolutionary relationships across diverse...
7.1K
Comparing Copy Number Variations and SNPs
18.8K
Sequencing of the human genome has opened up several best-kept secrets of the genome. Scientists have identified thousands of genome variations that exist within a population. These variations can be a single nucleotide or a larger chromosomal variation.
Copy number variations or CNVs are the structural variations that cover more than 1kb of DNA sequence. The single nucleotide polymorphism (SNP), on the other hand, is a single nucleotide change or a point mutation that is found in more than 1%...
Copy number variations or CNVs are the structural variations that cover more than 1kb of DNA sequence. The single nucleotide polymorphism (SNP), on the other hand, is a single nucleotide change or a point mutation that is found in more than 1%...
18.8K
Expected Frequencies in Goodness-of-Fit Tests
8.7K
A goodness-of-fit test is conducted to determine whether the observed frequency values are statistically similar to the frequencies expected for the dataset. Suppose the expected frequencies for a dataset are equal such as when predicting the frequency of any number appearing when casting a die. In that case, the expected frequency is the ratio of the total number of observations (n) to the number of categories (k).
8.7K
Comparing Mitochondrial, Chloroplast, and Prokaryotic Genomes
17.1K
The present-day mitochondrial and chloroplast genomes have retained some of the characteristics of their ancestral prokaryotes and also have acquired new attributes during their evolution within eukaryotic cells. Like prokaryotic genomes, mitochondrial and chloroplast genomes neither bind with histone-like proteins nor show complex packaging into chromosome-like structures, as observed in eukaryotes. Unlike mitotic cell divisions observed in eukaryotic cells, mitochondria and chloroplasts...
17.1K
Determination of Expected Frequency
2.6K
Suppose one wants to test independence between the two variables of a contingency table. The values in the table constitute the observed frequencies of the dataset. But how does one determine the expected frequency of the dataset? One of the important assumptions is that the two variables are independent, which means the variables do not influence each other. For independent variables, the statistical probability of any event involving both variables is calculated by multiplying the individual...
2.6K

