Related Experiment Video
Updated: Mar 8, 2026

12:00
A Practical Guide to Phylogenetics for Nonexperts
Published on: February 5, 2014
36.2K
Protein sequence-similarity search acceleration using a heuristic algorithm with a sensitive matrix
Kyungtaek Lim1, Kazunori D Yamada1,2, Martin C Frith1,3
1Artificial Intelligence Research Center, National Institute of Advanced Industrial Science and Technology (AIST), 2-4-7 Aomi, Koto-ku, Tokyo, 135-0064, Japan.
Journal of Structural and Functional Genomics
|January 14, 2017
Summary
The MIQS substitution matrix enhances protein homology detection when used with the LAST search tool. This combination offers faster and more accurate protein sequence analysis for genomics research.
Area of Science:
- Bioinformatics
- Computational Biology
- Genomics
Background:
- Protein database searching is crucial for structural and functional genomics.
- Amino acid substitution matrices significantly impact homology detection accuracy.
- Previous work introduced the MIQS matrix for distant protein homology search.
Purpose of the Study:
- To evaluate the MIQS substitution matrix combined with the LAST sequence alignment tool.
- To assess the performance of MIQS-LAST across various sensitivity parameters.
- To compare MIQS-LAST against existing homology search methods.
Main Methods:
- Utilized the MIQS substitution matrix with the LAST heuristic search tool.
- Tested LAST with a tunable sensitivity parameter (m) on a large protein database (approx. 15 million sequences).
- Compared performance against BLASTP, CS-BLAST, and SSEARCH.
Main Results:
- MIQS significantly improved LAST's homology detection and alignment quality.
- LAST with m=10^5 outperformed BLASTP and was 20x faster.
- LAST with MIQS and m=10^6 showed comparable sensitivity to CS-BLAST and SSEARCH at higher speeds.
Conclusions:
- MIQS-powered LAST provides a time-efficient method for sensitive and accurate protein homology searching.
- This approach enhances target selection in structural and functional genomics.
- The findings support the utility of MIQS for improving protein sequence analysis.
Related Concept Videos
Peptide Identification Using Tandem Mass Spectrometry
8.7K
Tandem mass spectrometry, also known as MS/MS or MS2, is an analytical technique that employs two mass analyzers. Essentially it is a series of mass spectrometers that helps isolate a particular biomolecule and then helps study its chemical properties.
This technique helps gather information regarding the protein from which the peptide was obtained and to study the peptides’ amino acid sequence. Identifying peptides from a complex mixture is an important component of the growing field of...
This technique helps gather information regarding the protein from which the peptide was obtained and to study the peptides’ amino acid sequence. Identifying peptides from a complex mixture is an important component of the growing field of...
8.7K
Evolutionary Relationships through Genome Comparisons
7.1K
Genome comparison is one of the excellent ways to interpret the evolutionary relationships between organisms. The basic principle of genome comparison is that if two species share a common feature, it is likely encoded by the DNA sequence conserved between both species. The advent of genome sequencing technologies in the late 20th century enabled scientists to understand the concept of conservation of domains between species and helped them to deduce evolutionary relationships across diverse...
7.1K
Leaky Scanning
5.8K
During most eukaryotic translation processes, the small 40S ribosome subunit scans an mRNA from its 5' end until it encounters the first start AUG codon. The large 60S ribosomal subunit then joins the smaller one to initiate protein synthesis. The location of the translation initiation is largely determined by the nucleotides near the start codon as there may be multiple translation initiation sites present on the mRNA. Marilyn Kozak discovered that the sequence RCCAUGG (where R...
5.8K
Protein Families
17.4K
Protein families are groups of homologous proteins; that is, they have similarities in amino acid sequences and three-dimensional structures. Protein families usually occur because of gene duplication, where an additional copy of a gene is inserted into the genome of an organism. Mutations that change the amino acids but still allow the protein to be properly synthesized, will lead to new protein family members. If these new proteins contain similar amino acids in key...
17.4K

