Related Experiment Video
Updated: Jun 15, 2025

In Vivo Modeling of the Morbid Human Genome using Danio rerio
Published on: August 24, 2013
Rapid discrimination between deleterious and benign missense mutations in the CAGI 6 experiment
Eshel Faraggi1,2, Robert L Jernigan3, Andrzej Kloczkowski4,5,6
1Research and Information Systems, LLC, 1620 E. 72nd ST., Indianapolis, IN, 46240, USA. efaraggi@gmail.com.
Abstract:
We describe the machine learning tool that we applied in the CAGI 6 experiment to predict whether single residue mutations in proteins are deleterious or benign. This tool was trained using only single sequences, i.e., without multiple sequence alignments or structural information. Instead, we used global characterizations of the protein sequence. Training and testing data for human gene mutations was obtained from ClinVar (ncbi.nlm.nih.gov/pub/ClinVar/), and for non-human gene mutations from Uniprot (www.uniprot.org). Testing was done on post-training data from ClinVar. This testing yielded high AUC and Matthews correlation coefficient (MCC) for well trained examples but low generalizability. For genes with either sparse or unbalanced training data, the prediction accuracy is poor. The resulting prediction server is available online at http://www.mamiris.com/Shoni.cagi6.
More Related Videos
Related Concept Videos
Mismatch Repair
The Mutator Protein Family Plays a Key Role in DNA Mismatch Repair
The human genome has more than 3 billion base pairs of DNA per cell. Prior to cell division, that vast amount of genetic...
Mutations
Nonsense-mediated mRNA Decay
Usually, Upf3 binds to an Exon Junction Complex (EJC) at mRNA splice sites. If a ribosome fully translates the mRNA,...

