Jove
Visualize
Contact Us
JoVE
x logofacebook logolinkedin logoyoutube logo
ABOUT JoVE
OverviewLeadershipBlogJoVE Help Center
AUTHORS
Publishing ProcessEditorial BoardScope & PoliciesPeer ReviewFAQSubmit
LIBRARIANS
TestimonialsSubscriptionsAccessResourcesLibrary Advisory BoardFAQ
RESEARCH
JoVE JournalMethods CollectionsJoVE Encyclopedia of ExperimentsArchive
EDUCATION
JoVE CoreJoVE BusinessJoVE Science EducationJoVE Lab ManualFaculty Resource CenterFaculty Site
Terms & Conditions of Use
Privacy Policy
Policies

Related Concept Videos

Protein Families02:47

Protein Families

Protein families are groups of homologous proteins; that is, they have similarities in amino acid sequences and three-dimensional structures. Protein families usually occur because of gene duplication, where an additional copy of a gene is inserted into the genome of an organism.   Mutations that change the amino acids but still allow the protein to be properly synthesized, will lead to new protein family members.   If these new proteins contain similar amino acids in key locations, protein...
Protein Networks02:26

Protein Networks

An organism can have thousands of different proteins, and these proteins must cooperate to ensure the health of an organism. Proteins bind to other proteins and form complexes to carry out their functions. Many proteins interact with multiple other proteins creating a complex network of protein interactions.
These interactions can be represented through maps depicting protein-protein interaction networks, represented as nodes and edges. Nodes are circles that are representative of a protein,...

You might also read

Related Articles

Articles linked to this work by shared authors, journal, and citation graph.

Sort by
Same author

Unravelling Ovarian Cancer: an analysis of the Influence of LRP1 and PAI1 Genetic Variations.

Biochemical genetics·2026
Same author

Genome-Edited Maize Expressing Two Native Genes Confers Broad-Spectrum Resistance to Northern Corn Leaf Blight.

Molecular plant pathology·2026
Same author

Interplay between NOD1 polymorphisms and gene expression in the pathogenesis of gallstone disease.

Molecular biology reports·2026
Same author

Genetic impact of copy number variations on congenital heart defects: Current insights and future directions.

Global medical genetics·2025
Same author

Optimizing the strain engineering process for industrial-scale production of bio-based molecules.

Journal of industrial microbiology & biotechnology·2023
Same author

Author Correction: Deep mutational scanning of essential bacterial proteins can guide antibiotic development.

Nature communications·2023

Related Experiment Video

Updated: Jul 13, 2026

An Integrated Approach for Microprotein Identification and Sequence Analysis
09:37

An Integrated Approach for Microprotein Identification and Sequence Analysis

Published on: July 12, 2022

Automated protein subfamily identification and classification.

Duncan P Brown1, Nandini Krishnamurthy, Kimmen Sjölander

  • 1Department of Bioengineering, University of California, Berkeley, California, United States of America.

Plos Computational Biology
|August 22, 2007
PubMed
Summary

This study introduces SCI-PHY, a phylogenomic pipeline for accurate protein function prediction. It automates subfamily identification using Hidden Markov Models (HMMs), improving gene annotation and avoiding errors common in homology-based methods.

More Related Videos

Creating and Applying a Reference to Facilitate the Discussion and Classification of Proteins in a Diverse Group
07:49

Creating and Applying a Reference to Facilitate the Discussion and Classification of Proteins in a Diverse Group

Published on: August 16, 2017

A Protocol for Computer-Based Protein Structure and Function Prediction
16:41

A Protocol for Computer-Based Protein Structure and Function Prediction

Published on: November 3, 2011

Related Experiment Videos

Last Updated: Jul 13, 2026

An Integrated Approach for Microprotein Identification and Sequence Analysis
09:37

An Integrated Approach for Microprotein Identification and Sequence Analysis

Published on: July 12, 2022

Creating and Applying a Reference to Facilitate the Discussion and Classification of Proteins in a Diverse Group
07:49

Creating and Applying a Reference to Facilitate the Discussion and Classification of Proteins in a Diverse Group

Published on: August 16, 2017

A Protocol for Computer-Based Protein Structure and Function Prediction
16:41

A Protocol for Computer-Based Protein Structure and Function Prediction

Published on: November 3, 2011

Area of Science:

  • Bioinformatics
  • Computational Biology
  • Genomics

Background:

  • Homology-based gene function prediction is common but prone to errors.
  • Phylogenomic analysis offers accurate predictions but is difficult to automate.
  • Existing methods struggle with high-throughput annotation and error propagation.

Purpose of the Study:

  • To develop a computationally efficient pipeline for automated phylogenomic protein classification.
  • To improve the accuracy and reliability of gene function prediction.
  • To address the limitations of homology-based annotation and manual phylogenomic analysis.

Main Methods:

  • Utilized the SCI-PHY (Subfamily Classification in Phylogenomics) algorithm for automated subfamily identification.
  • Employed subfamily Hidden Markov Models (HMMs) for sequence classification.
  • Implemented logistic regression to differentiate novel subfamilies and an information-sharing protocol for HMM parameter estimation.

Main Results:

  • SCI-PHY subfamilies closely align with expert-defined functional subtypes and conserved phylogenetic clades.
  • Subfamily HMMs significantly enhance the discrimination between homologous and non-homologous proteins in database searches.
  • Achieved extremely high specificity in classification and demonstrated potential for predicting novel subtypes.

Conclusions:

  • The SCI-PHY pipeline provides an automated and efficient method for accurate protein subfamily classification.
  • Subfamily HMMs represent a significant advancement over family HMMs for precise functional annotation.
  • The SCI-PHY Web server and PhyloFacts resource offer valuable tools for the research community.