Frequent contiguous pattern mining over biological sequences of protein misfolded diseases

Mohammad Shahedul Islam1, Md Abul Kashem Mia2, Mohammad Shamsur Rahman3

  • 1Information Communication Technology Centre, Bangabandhu Sheikh Mujibur Rahman Maritime University, Pallabi, Mirpur-12, Dhaka, Bangladesh.

BMC Bioinformatics
|September 13, 2021
PubMed
Summary

This study identifies amino acid patterns in complex protein misfolding diseases using association rule mining. The findings offer reliable rules for discovering new medicines for genetic disorders.

Related Concept Videos

Protein Folding01:25

Protein Folding

Proteins are chains of amino acids linked together by peptide bonds. Upon synthesis, a protein folds into a three-dimensional conformation, critical to its biological function. Interactions between its constituent amino acids guide protein folding, and hence the protein structure is primarily dependent on its amino acid sequence.
Protein Structure Is Critical to Its Biological Function
Proteins perform a wide range of biological functions such as catalyzing chemical reactions, providing...
9.7K
Amyloid Fibrils03:03

Amyloid Fibrils

Amyloid fibrils are aggregates of misfolded proteins.  Under most circumstances, misfolded proteins are either refolded by chaperone proteins or degraded by the proteasome. However, in the case of a mutation or a disease, these proteins can accumulate to form large clusters and often further assemble to form elongated fibers, called fibrils. 
Amyloid deposits were observed as early as 1639 in the liver and the spleen.   In 1854, Rudolph Virchow performed iodine staining,...
10.8K
Signal Sequences and Sorting Receptors01:41

Signal Sequences and Sorting Receptors

Signal sequences are short amino acid sequences that guide newly synthesized proteins to their proper location within the cell. Classical signal sequences are fifteen to sixty amino acids long and present at the N-terminus of a polypeptide chain. Each signal sequence has a conserved segment of basic residues towards their N terminus, a hydrophobic core, and a C-terminus rich in polar residues. The C-terminus also contains a signal cleavage site and features a -3 -1 sequence motif. The -3-1...
10.8K
Protein Families02:47

Protein Families

Protein families are groups of homologous proteins; that is, they have similarities in amino acid sequences and three-dimensional structures. Protein families usually occur because of gene duplication, where an additional copy of a gene is inserted into the genome of an organism.   Mutations that change the amino acids but still allow the protein to be properly synthesized, will lead to new protein family members.   If these new proteins contain similar amino acids in key...
16.1K
Protein Networks02:26

Protein Networks

An organism can have thousands of different proteins, and these proteins must cooperate to ensure the health of an organism. Proteins bind to other proteins and form complexes to carry out their functions. Many proteins interact with multiple other proteins creating a complex network of protein interactions.
These interactions can be represented through maps depicting protein-protein interaction networks, represented as nodes and edges. Nodes are circles that are representative of a protein,...
4.2K
Conserved Binding Sites01:49

Conserved Binding Sites

Many proteins’ biological role depends on their interactions with their ligands, small molecules that bind to specific locations on the protein known as ligand-binding sites. Ligand-binding sites are often conserved among homologous proteins as these sites are critical for protein function.
Binding sites are often located in large pockets, and if their location on a protein’s surface is unknown, it can be predicted using various approaches. The energetic method computationally...
4.7K