Related Experiment Video
Updated: Feb 20, 2026

16:41
A Protocol for Computer-Based Protein Structure and Function Prediction
Published on: November 3, 2011
69.9K
HashGO: hashing gene ontology for protein function prediction
Guoxian Yu1, Yingwen Zhao1, Chang Lu1
1College of Computer and Information Science, Southwest University, Chongqing 400715, China.
Computational Biology and Chemistry
|October 17, 2017
Summary
Hashing GO (HashGO) accurately predicts protein functions by exploring Gene Ontology (GO) term relationships. This method efficiently identifies protein associations and predicts missing annotations faster than existing approaches.
Area of Science:
- Bioinformatics
- Computational Biology
- Genomics
Background:
- Gene Ontology (GO) provides a standardized vocabulary for protein function, but predicting associations with thousands of terms is challenging.
- Accurate protein function prediction is crucial for understanding biological roles and cellular processes.
- Existing methods struggle with the scale and complexity of GO annotations.
Purpose of the Study:
- To develop an accurate and efficient method for predicting protein-GO term associations.
- To address the challenge of massive GO term datasets in protein function prediction.
- To improve the prediction of missing protein annotations.
Main Methods:
- Proposed Hashing GO for protein function prediction (HashGO).
- Utilized a protein-term association matrix and a graph hashing method to reduce dimensionality.
- Computed protein semantic similarity using Hamming distance on a compressed matrix.
- Predicted missing annotations based on semantic neighbors.
Main Results:
- HashGO demonstrated higher accuracy in predicting protein functions compared to related approaches.
- The method significantly improved the speed of protein function prediction.
- Experiments on Yeast and Human datasets validated HashGO's effectiveness.
Conclusions:
- HashGO offers a novel and effective solution for large-scale protein function prediction.
- The graph hashing approach efficiently captures GO term structures for improved accuracy.
- HashGO provides a faster and more accurate alternative for annotating protein functions.
Related Concept Videos
Protein-protein Interfaces
14.8K
Many proteins form complexes to carry out their functions, making protein-protein interactions (PPIs) essential for an organism's survival. Most PPIs are stabilized by numerous weak noncovalent chemical forces. The physical shape of the interfaces determines the way two proteins interact. Many globular proteins have closely-matching shapes on their surfaces, which form a large number of weak bonds. Additionally, many PPIs occur between two helices or between a surface cleft and a...
14.8K
Protein Networks
4.6K
An organism can have thousands of different proteins, and these proteins must cooperate to ensure the health of an organism. Proteins bind to other proteins and form complexes to carry out their functions. Many proteins interact with multiple other proteins creating a complex network of protein interactions.
These interactions can be represented through maps depicting protein-protein interaction networks, represented as nodes and edges. Nodes are circles that are representative of a protein,...
These interactions can be represented through maps depicting protein-protein interaction networks, represented as nodes and edges. Nodes are circles that are representative of a protein,...
4.6K
Protein Networks
2.9K
2.9K
Conservation of Protein Domains Over Different Proteins
14.8K
Protein domains are small structurally independent units that are part of a single amino acid chain. Although these domains are often structurally independent, they may rely on synergistic effects to perform their functions as part of a larger protein. Protein domains may be conserved within the same organism, as well as across different organisms.
A limited set of protein domains often duplicate and recombine during evolution. These domains can be organized in different combinations to...
A limited set of protein domains often duplicate and recombine during evolution. These domains can be organized in different combinations to...
14.8K
Protein Families
17.3K
Protein families are groups of homologous proteins; that is, they have similarities in amino acid sequences and three-dimensional structures. Protein families usually occur because of gene duplication, where an additional copy of a gene is inserted into the genome of an organism. Mutations that change the amino acids but still allow the protein to be properly synthesized, will lead to new protein family members. If these new proteins contain similar amino acids in key...
17.3K
Protein Families
4.5K
4.5K

