Related Experiment Video
Updated: Jul 25, 2025

Author Spotlight: Advancing Alzheimer's Research – Exploring Early Detection and Multi-Omics Approaches
Published on: December 15, 2023
Different Recognition of Protein Features Depending on Deep Learning Models: A Case Study of Aromatic Decarboxylase
Naoki Watanabe1, Yuki Kuriya1, Masahiro Murata2
1Artificial Intelligence Center for Health and Biomedical Research, National Institutes of Biomedical Innovation, Health and Nutrition, 3-17 Senrioka-shinmachi, Settsu 566-0002, Japan.
Abstract:
The number of unannotated protein sequences is explosively increasing due to genome sequence technology. A more comprehensive understanding of protein functions for protein annotation requires the discovery of new features that cannot be captured from conventional methods. Deep learning can extract important features from input data and predict protein functions based on the features. Here, protein feature vectors generated by 3 deep learning models are analyzed using Integrated Gradients to explore important features of amino acid sites. As a case study, prediction and feature extraction models for UbiD enzymes were built using these models. The important amino acid residues extracted from the models were different from secondary structures, conserved regions and active sites of known UbiD information. Interestingly, the different amino acid residues within UbiD sequences were regarded as important factors depending on the type of models and sequences. The Transformer models focused on more specific regions than the other models. These results suggest that each deep learning model understands protein features with different aspects from existing knowledge and has the potential to discover new laws of protein functions. This study will help to extract new protein features for the other protein annotations.
Related Concept Videos
Peptide Identification Using Tandem Mass Spectrometry
This technique helps gather information regarding the protein from which the peptide was obtained and to study the peptides’ amino acid sequence. Identifying peptides from a complex mixture is an important component of the growing field of...
Conservation of Protein Domains Over Different Proteins
A limited set of protein domains often duplicate and recombine during evolution. These domains can be organized in different combinations to...
¹³C NMR: Distortionless Enhancement by Polarization Transfer (DEPT)
Protein Denaturation
Covalently Linked Protein Regulators
These groups modify specific amino acids in a protein....
IR and UV–Vis Spectroscopy of Carboxylic Acids
However, the stretching absorptions for the C=O bond vary depending on the structure of carboxylic acids. The C=O bond of the free carboxylic acids shows a higher stretching frequency, 1760 cm−1, while H-bonded carboxylic acids (dimers) exhibit stretching absorptions at a lower frequency,...

