Using reinforcement learning in genome assembly: in-depth analysis of a Q-learning assembler.

Kleber Padovani1, Rafael Cabral Borges2, Roberto Xavier2

  • 1Center for Higher Studies of Itacoatiara, University of the State of Amazonas, Itacoatiara, Amazonas, Brazil.

Frontiers in Bioinformatics
|September 5, 2025
PubMed
Summary

Reinforcement learning (RL) for de novo genome assembly shows poor scalability. Despite improvements, Q-learning approaches struggle with assembly quality and execution time, highlighting limitations for complex genomic tasks.

Related Concept Videos

Genome Annotation and Assembly03:36

Genome Annotation and Assembly

The genome refers to all of the genetic material in an organism. It can range from a few million base pairs in microbial cells to several billion base pairs in many eukaryotic organisms. Genome assembly refers to the process of taking the DNA sequencing data and putting it all back together in a correct order to create a close representation of the original genome. This is followed by the identification of functional elements on the newly assembled genome, a process called genome annotation.
19.3K
Protein Complex Assembly02:41

Protein Complex Assembly

Proteins can form homomeric complexes with another unit of the same protein or heteromeric complexes with different types.  Most protein complexes self-assemble spontaneously via ordered pathways, while some proteins need assembly factors that guide their proper assembly. Despite the crowded intracellular environment, proteins usually interact with their correct partners and form functional complexes.
Many viruses self-assemble into a fully functional unit using the infected host cell to...
10.8K
Assembly of Signaling Complexes01:30

Assembly of Signaling Complexes

Multiprotein signaling complexes are formed in a dynamic process involving protein-protein interactions at the cytoplasmic domain of transmembrane receptors or enzymatic and non-enzymatic proteins associated with the receptor. These complexes ensure the activation and propagation of intracellular signals that regulate cell functions.
Interaction domains in cell signaling
Interaction domains recognize exposed features of their binding partners containing post-translationally modified sequences,...
5.9K
Oligosaccharide Assembly01:24

Oligosaccharide Assembly

Protein glycosylation starts in the ER lumen and continues in the Golgi apparatus. Glycosyltransferases catalyze the addition of sugar molecules or glycosylation of proteins. Usually, these enzymes add sugars to the hydroxyl groups of selected serine or threonine residues to form O-linked glycans or the amino groups of asparagine residues to form N-linked glycans. Different positions on the same polypeptide chain can contain differently linked glycans.
Multiple sugar molecules that may or may...
3.0K
Observational Learning01:12

Observational Learning

Albert Bandura's observational learning, also known as imitation or modeling, occurs when a person observes and imitates another's behavior. It is a quicker process than operant conditioning. A well-known example is the Bobo doll study, where children who saw an adult acting aggressively towards the doll were more likely to act aggressively when left alone, compared to those who observed a nonaggressive adult. Many psychologists view observational learning as a form of latent learning...
305
Introduction to Learning01:18

Introduction to Learning

Learning is the process of acquiring knowledge or skills through practice or experience, leading to long-lasting behavioral changes. This acquisition occurs through interaction with the environment and requires practice or experience. For instance, mastering a skill such as surfing requires considerable practice and experience, highlighting the essential role of repeated interactions with the environment in learning.
In contrast to learned behaviors, unlearned behaviors such as crying, sexual...
528