Related Experiment Video
Updated: Oct 10, 2025

09:37
An Integrated Approach for Microprotein Identification and Sequence Analysis
Published on: July 12, 2022
3.6K
Conflict between Amino Acid and Nucleotide Characters
Mark P Simmons1, Helga Ochoterena2, John V Freudenstein1
1The Ohio State University Herbarium, 1315 Kinnear Road, Columbus, Ohio, 43212.
Cladistics : the International Journal of the Willi Hennig Society
|December 16, 2021
Summary
Silent substitutions in DNA sequences are more informative for phylogenetic analysis than amino acid replacements. Using amino acids can lead to misleading evolutionary trees due to composite characters and convergence.
Area of Science:
- Molecular Evolution
- Phylogenetics
- Bioinformatics
Background:
- Phylogenetic analyses traditionally favor slowly evolving characters like amino acid replacements.
- Amino acids are composite characters influenced by the degenerate genetic code, increasing susceptibility to convergent evolution.
- Previous studies have overlooked the potential phylogenetic utility of silent, or synonymous, nucleotide substitutions.
Purpose of the Study:
- To compare the phylogenetic informativeness of silent nucleotide substitutions versus amino acid replacement substitutions.
- To investigate the impact of composite characters and convergence on phylogenetic tree reconstruction using amino acid data.
- To assess the reliability of amino acid-based phylogenies in seed plants.
Main Methods:
- Analysis of the atpB and rbcL genes from 567 seed plant species.
- Phylogenetic tree construction using both nucleotide (silent substitutions) and amino acid (replacement substitutions) datasets.
- Comparison of resulting phylogenetic trees with independent biological evidence.
Main Results:
- Silent nucleotide substitutions provide more accurate phylogenetic signals than amino acid replacements.
- Phylogenetic artifacts, including composite characters and convergence, cause amino acid-based trees to conflict with nucleotide-based trees.
- Amino acid trees showed discordance with independent evidence, highlighting potential inaccuracies.
Conclusions:
- Coding nucleotide sequences solely as amino acid characters offers limited phylogenetic benefit.
- Relying on amino acid data for phylogenetic inference can yield misleading evolutionary relationships.
- Silent nucleotide substitutions are a more reliable character type for reconstructing evolutionary history in seed plants.
Related Concept Videos
Mutations
85.4K
Overview
85.4K
From DNA to Protein
19.8K
The flow of genetic information in cells from DNA to mRNA to protein is described by the central dogma, which states that genes specify the sequence of mRNAs, which in turn specify the sequence of amino acids making up all proteins. The decoding of one molecule to another is performed by specific proteins and RNAs. Because the information stored in DNA is so central to cellular function, it makes intuitive sense that the cell would make mRNA copies of this information for protein synthesis...
19.8K
tRNA Activation
20.4K
Aminoacyl-tRNA synthetases are present in both eukaryotes and bacteria. Though eukaryotes have 20 different aminoacyl-tRNA synthetases to couple to 20 amino acids, many bacteria do not have genes for all of these aminoacyl-tRNA synthetases. Despite this, they still use all 20 amino acids to synthesize their proteins. For instance, some bacteria do not have the gene encoding the enzyme that couples glutamine with its partner tRNA. In these organisms, one enzyme adds glutamic acid to all of the...
20.4K
DNA Base Pairing
29.6K
Erwin Chargaff’s rules on DNA equivalence paved the way for the discovery of base pairing in DNA. Chargaff’s rules state that in a double-stranded DNA molecule,
29.6K
Amino acids
96.1K
Amino acids are the monomers that comprise proteins. Each amino acid has the same fundamental structure, which consists of a central carbon atom, or the alpha (α) carbon, bonded to an amino group (NH2), a carboxyl group (COOH), and to a hydrogen atom. Every amino acid also has another atom or group of atoms bonded to the central atom known as the R group. There are 20 common amino acids present in proteins, each with a different R group. Variation in the amino acid sequence is responsible...
96.1K
Nucleic Acids and Nucleotides
11.0K
Nucleic acids are the most important macromolecules for the continuity of life. They carry the cell's genetic blueprint and have instructions for its functioning. The two main types of nucleic acids are deoxyribonucleic acid (DNA) and ribonucleic acid (RNA).
Deoxyribonucleic Acid (DNA)
DNA is the genetic material in all living organisms, ranging from single-celled bacteria to multicellular mammals. It is in the nucleus of eukaryotes and the organelles such as chloroplasts and mitochondria....
Deoxyribonucleic Acid (DNA)
DNA is the genetic material in all living organisms, ranging from single-celled bacteria to multicellular mammals. It is in the nucleus of eukaryotes and the organelles such as chloroplasts and mitochondria....
11.0K

