Related Experiment Video
Updated: Apr 15, 2026

Determining the Likelihood of Variant Pathogenicity Using Amino Acid-level Signal-to-Noise Analysis of Genetic Variation
Published on: January 16, 2019
Development and Validation of an Algorithm for Constructing an Amino Acid Database for Application to the Korean
1Department of Foodservice Management and Nutrition, Graduate School, Sangmyung University, Seoul 03016, Republic of Korea.
Background/Objectives:
The Korean Genome and Epidemiology Study (KoGES) is a large population-based cohort designed to investigate chronic disease risk using long-term dietary and health data. However, comprehensive amino acid information for estimating long-term intake from food frequency questionnaire (FFQ) data remains limited. This study aimed to develop and validate a standardized, rule-based algorithm for food matching and substitution and to construct an amino acid database applicable to the KoGES FFQ.
Methods:
The algorithm sequentially evaluated food name concordance, preparation forms, substitutability of similar foods, and differences in energy, macronutrients, and moisture (±20%). Amino acid composition data were derived from domestic and international food composition tables and published literature, with protein-nitrogen conversion factors applied by food group.
Results:
Amino acid information was established for 475 FFQ food items covering 19 amino acids. Of the database values, 31.0% were analytical, 64.2% were calculated, and 4.8% were substituted. Overall database coverage across all amino acid-food item combinations was 98.8%. The constructed database was applied to dietary data from the second follow-up (Phase 3) of the KoGES Ansan and Ansung community-based cohorts, showing that total amino acid intake accounted for 86.7% of total protein intake, reflecting the inclusion of non-protein nitrogen in conventional protein estimates. Based on the Estimated Average Requirement (EAR) criteria, the proportions of participants with intakes below the EAR for protein and essential amino acids varied across age and sex groups. Overall and in both men and women, lysine showed the highest proportion of participants below the EAR, whereas tryptophan showed the lowest proportion.
Conclusions:
This standardized algorithm provides a reproducible framework for constructing amino acid databases and can be applied to large-scale cohort and dietary survey data.
More Related Videos
05:08Application of I TASSER, trRosetta, UCSF Chimera, HADDOCK server, and HEX loria for De Novo and In Silico Design of Proteins
Published on: July 8, 2025
06:50Author Spotlight: A Computational Approach to Decipher Amino Acid Preferences in Multispecific Protein-Protein Interactions
Published on: January 26, 2024