Genome-based source attribution using a One Health Escherichia coli isolate collection from 2013 to 2023 in Scotland.

Antonia Chalka1,2, Louise Crozier2, Adriana Vallejo-Trujillo1,2

  • 1The Roslin Institute and Royal (Dick) School of Veterinary Medicine, University of Edinburgh, Easter Bush Estate, Edinburgh, Scotland, EH25 9RG, UK.

Microbial Genomics
|April 20, 2026
PubMed
Summary

This study used random forest models to trace Escherichia coli (E. coli) contamination sources. A significant livestock E. coli signal was found in 15% of human clinical isolates, highlighting agri-food chain transmission risks.

Related Concept Videos

Modern Molecular Taxonomy01:29

Modern Molecular Taxonomy

Advancements in molecular biology have revolutionized the identification and characterization of bacteria, with multiple methods leveraging DNA sequencing for enhanced precision. As sequencing technologies improve and costs decline, these approaches are increasingly used in clinical, environmental, and evolutionary studies.Multilocus Sequence Typing (MLST) examines several housekeeping genes, essential chromosomal genes encoding cellular functions, to distinguish strains. Approximately...
828
Investigation of Disease Outbreaks01:23

Investigation of Disease Outbreaks

Multistate foodborne outbreaks pose significant public health risks and require meticulous investigation to identify sources and implement control measures. The Centers for Disease Control and Prevention (CDC) utilizes a dynamic seven-step process for these investigations, integrating data from laboratories, interviews, and environmental assessments to protect public health.Outbreak Detection: The detection of multistate outbreaks typically begins with PulseNet, the CDC's national laboratory...
69
Evolution of Microbial Genome01:08

Evolution of Microbial Genome

Microbial genome evolution is a highly dynamic process shaped by continual gene gain and loss across species and strains. This genomic flexibility allows microorganisms to adapt rapidly to environmental pressures and interactions with other organisms. Central to understanding this diversity is the distinction between the core and pan genomes.The core genome comprises the genes shared by all sampled strains of a species, representing essential functions needed for fundamental cellular processes.
83
Applications of Molecular Taxonomy01:20

Applications of Molecular Taxonomy

Molecular taxonomy has revolutionized the understanding and classification of bacteria, providing precise insights into their diversity, evolutionary relationships, and ecological roles. By utilizing molecular techniques such as DNA sequencing and fingerprinting, researchers have made significant strides in various fields related to bacterial studies.Resolving Taxonomic AmbiguitiesMolecular taxonomy has been instrumental in distinguishing closely related bacterial species initially thought to...
699
Evolutionary Relationships through Genome Comparisons02:54

Evolutionary Relationships through Genome Comparisons

Genome comparison is one of the excellent ways to interpret the evolutionary relationships between organisms. The basic principle of genome comparison is that if two species share a common feature, it is likely encoded by the DNA sequence conserved between both species. The advent of genome sequencing technologies in the late 20th century enabled scientists to understand the concept of conservation of domains between species and helped them to deduce evolutionary relationships across diverse...
5.8K
Genome Size and the Evolution of New Genes03:21

Genome Size and the Evolution of New Genes

While every living organism has a genome of some kind (be it RNA, or DNA), there is considerable variation in the sizes of these blueprints. One major factor that impacts genome size is whether the organism is prokaryotic or eukaryotic. In prokaryotes, the genome contains little to no non-coding sequence, such that genes are tightly clustered in groups or operons sequentially along the chromosome. Conversely, the genes in eukaryotes are punctuated by long stretches of non-coding sequence.
7.5K