Related Experiment Video
Updated: Feb 15, 2026

Ultra-long Read Sequencing for Whole Genomic DNA Analysis
Published on: March 15, 2019
Analyzing similarities in genome sequences
I C Fonseca1, E Nogueira1, P H Figueirêdo2
1Departamento de Física, Universidade Federal da Paraíba, 58051-970, João Pessoa, PB, Brazil.
Abstract:
This article investigates aspects of similarity between complete sequences of mitochondrial DNA by determining the distribution of the relative frequencies of words with different lengths and the characteristics of their relevance throughout the sequences. The degree of similarity is obtained by comparing the distances between words contained within these sequences. Our results indicate that the best groupings among different species depend on the lengths of words and their respective relative frequencies. We also observed that the longer the word the more consistent the grouping between the sequences becomes. The application of our results, together with the perspective of analyzing DNA sequences belonging to a single biological species, may be important for the construction of phylogenetic trees, which are appropriate structures for understanding the evolutionary history of the species.
Related Concept Videos
Genomics
Causes of Similarity-Dissimilarity Effect
Factors Influencing Attraction III: Similarity
Genomic Imprinting and Inheritance
The expression of some genes depends on which parent passed the gene to the offspring, through a phenomenon known as...
Genome Size and the Evolution of New Genes
Cis-regulatory Sequences

