Jove
Visualize
Contáctanos
JoVE
x logofacebook logolinkedin logoyoutube logo
ACERCA DE JoVE
Visión GeneralLiderazgoBlogCentro de Ayuda JoVE
AUTORES
Proceso de PublicaciónConsejo EditorialAlcance y PolíticasRevisión por ParesPreguntas FrecuentesEnviar
BIBLIOTECARIOS
TestimoniosSuscripcionesAccesoRecursosConsejo Asesor de BibliotecasPreguntas Frecuentes
INVESTIGACIÓN
JoVE JournalMethods CollectionsJoVE Encyclopedia of ExperimentsArchivo
EDUCACIÓN
JoVE CoreJoVE BusinessJoVE Science EducationJoVE Lab ManualCentro de Recursos para ProfesoresSitio de Profesores
Términos y Condiciones de Uso
Política de Privacidad
Políticas

Videos de Conceptos Relacionados

Improving Translational Accuracy02:07

Improving Translational Accuracy

11.8K
Base complementarity between the three base pairs of mRNA codon and the tRNA anticodon is not a failsafe mechanism. Inaccuracies can range from a single mismatch to no correct base pairing at all. The free energy difference between the correct and nearly correct base pairs can be as small as 3 kcal/ mol. With complementarity being the only proofreading step, the estimated error frequency would be one wrong amino acid in every 100 amino acids incorporated. However, error frequencies observed in...
11.8K
Genomics02:02

Genomics

37.4K
Genomics is the science of genomes: it is the study of all the genetic material of an organism. In humans, the genome consists of information carried in 23 pairs of chromosomes in the nucleus, as well as mitochondrial DNA. In genomics, both coding and non-coding DNA is sequenced and analyzed. Genomics allows a better understanding of all living things, their evolution, and their diversity. It has a myriad of uses: for example, to build phylogenetic trees, to improve productivity and...
37.4K
Genetic Lingo01:11

Genetic Lingo

104.5K
Overview
104.5K
Genome Annotation and Assembly03:36

Genome Annotation and Assembly

19.3K
The genome refers to all of the genetic material in an organism. It can range from a few million base pairs in microbial cells to several billion base pairs in many eukaryotic organisms. Genome assembly refers to the process of taking the DNA sequencing data and putting it all back together in a correct order to create a close representation of the original genome. This is followed by the identification of functional elements on the newly assembled genome, a process called genome annotation.
19.3K
Genome-wide Association Studies-GWAS01:11

Genome-wide Association Studies-GWAS

14.1K
Genome-wide association studies or GWAS are used to identify whether common SNPs are associated with certain diseases. Suppose specific SNPs are more frequently observed in individuals with a particular disease than those without the disease. In that case, those SNPs are said to be associated with the disease. Chi-square analysis is performed to check the probability of the allele likely to be associated with the disease.
GWAS does not require the identification of the target gene involved in...
14.1K
Comparing Copy Number Variations and SNPs02:26

Comparing Copy Number Variations and SNPs

17.9K
Sequencing of the human genome has opened up several best-kept secrets of the genome. Scientists have identified thousands of genome variations that exist within a population. These variations can be a single nucleotide or a larger chromosomal variation.
Copy number variations or CNVs are the structural variations that cover more than 1kb of DNA sequence. The single nucleotide polymorphism (SNP), on the other hand, is a single nucleotide change or a point mutation that is found in more than 1%...
17.9K

También podría leer

Artículos Relacionados

Artículos vinculados a este trabajo por autores compartidos, revista y gráfico de citas.

Ordenar por
Same author

Spatial co-expression and cell-cell communication inference from spatially resolved transcriptomics with CONCISE.

bioRxiv : the preprint server for biology·2026
Same author

A unified framework for selecting and evaluating cell-type-specific gene co-expressions in single-cell data.

Briefings in bioinformatics·2026
Same author

MIXPRS enables multi-population and multi-method polygenic risk scores using summary statistics.

Nature genetics·2026
Same author

Identification of multi-omic pleiotropy factors for peripheral artery disease.

Human molecular genetics·2026
Same author

Multi-ancestry transcriptome-wide association studies uncover insights into breast cancer genetics and biology.

Nature communications·2026
Same author

Loss of Cyclin G-Associated Kinase Leads to Lysosome Dysfunction and Immune Modulation in Podocytes.

Journal of the American Society of Nephrology : JASN·2026

Video Experimental Relacionado

Updated: Sep 9, 2025

Screening for Functional Non-coding Genetic Variants Using Electrophoretic Mobility Shift Assay EMSA and DNA-affinity Precipitation Assay DAPA
11:35

Screening for Functional Non-coding Genetic Variants Using Electrophoretic Mobility Shift Assay EMSA and DNA-affinity Precipitation Assay DAPA

Published on: August 21, 2016

13.1K

Modelo de lenguaje genómico de preentrenamiento con variantes para un mejor modelado de la genómica funcional

Tianyu Liu, Xiangyu Zhang, Jiecong Lin

    bioRxiv : the preprint server for biology
    |September 2, 2025
    PubMed
    Resumen

    Desarrollamos UKBioBERT, un modelo de lenguaje de ADN, para mejorar la predicción de la expresión génica mediante la integración de datos genéticos con modelos de secuencia a función. Este enfoque mejora la comprensión de la regulación genética y los efectos de las variantes genéticas.

    Más Videos Relacionados

    Identification and Classification of Position-specific GABAA Receptor Subunit Missense Variants for Their Role In Hippocampal Pyramidal Neurons
    08:04

    Identification and Classification of Position-specific GABAA Receptor Subunit Missense Variants for Their Role In Hippocampal Pyramidal Neurons

    Published on: June 6, 2025

    492
    Targeted Next-generation Sequencing and Bioinformatics Pipeline to Evaluate Genetic Determinants of Constitutional Disease
    09:34

    Targeted Next-generation Sequencing and Bioinformatics Pipeline to Evaluate Genetic Determinants of Constitutional Disease

    Published on: April 4, 2018

    34.0K

    Videos de Experimentos Relacionados

    Last Updated: Sep 9, 2025

    Screening for Functional Non-coding Genetic Variants Using Electrophoretic Mobility Shift Assay EMSA and DNA-affinity Precipitation Assay DAPA
    11:35

    Screening for Functional Non-coding Genetic Variants Using Electrophoretic Mobility Shift Assay EMSA and DNA-affinity Precipitation Assay DAPA

    Published on: August 21, 2016

    13.1K
    Identification and Classification of Position-specific GABAA Receptor Subunit Missense Variants for Their Role In Hippocampal Pyramidal Neurons
    08:04

    Identification and Classification of Position-specific GABAA Receptor Subunit Missense Variants for Their Role In Hippocampal Pyramidal Neurons

    Published on: June 6, 2025

    492
    Targeted Next-generation Sequencing and Bioinformatics Pipeline to Evaluate Genetic Determinants of Constitutional Disease
    09:34

    Targeted Next-generation Sequencing and Bioinformatics Pipeline to Evaluate Genetic Determinants of Constitutional Disease

    Published on: April 4, 2018

    34.0K

    Área de la Ciencia:

    • La genómica
    • Biología computacional
    • La bioinformática

    Sus antecedentes:

    • Los modelos de lenguaje genómico (GLM) aprenden de las secuencias de ADN para representar el contexto genómico.
    • Los modelos de secuencia a función (S2F) vinculan la información genética con la expresión génica y los fenotipos.
    • El puente entre los modelos GLM y S2F para la predicción de la expresión génica individualizada sigue siendo un desafío.

    Objetivo del estudio:

    • Desarrollar un nuevo modelo de lenguaje de ADN, UKBioBERT, utilizando datos genéticos del Reino Unido BioBank.
    • Integrar UKBioBERT con los modelos S2F existentes (Enformer, Borzoi) para crear modelos predictivos mejorados.
    • Mejorar la predicción de los niveles de expresión génica y comprender el impacto de las variantes genéticas.

    Principales métodos:

    • Formación previa de un modelo de lenguaje de ADN (UKBioBERT) sobre las variantes genéticas del BioBank del Reino Unido.
    • La generación de secuencias informativas desde UKBioBERT.
    • La combinación de las incorporaciones de UKBioBERT con las arquitecturas S2F (Enformer, Borzoi) para formar UKBioFormer y UKBioZoi.
    • Evaluación del rendimiento del modelo en la predicción de la expresión génica en diferentes cohortes.

    Principales resultados:

    • Las incorporaciones de UKBioBERT identifican eficazmente las funciones genéticas y mejoran la predicción de la expresión génica en las líneas celulares.
    • UKBioFormer y UKBioZoi demuestran un rendimiento superior en la predicción de niveles de expresión génica altamente predecibles.
    • Los modelos integrados se generalizan bien en diversas cohortes.
    • UKBioFormer captura con precisión las relaciones genotipo-fenotipo, lo que permite el análisis de mutaciones in silico.

    Conclusiones:

    • La integración de modelos de lenguaje genómico con enfoques de secuencia a función avanza significativamente en la genómica funcional.
    • UKBioBERT proporciona información valiosa para comprender la función genética y la previsibilidad de la expresión.
    • Los modelos desarrollados UKBioFormer y UKBioZoi ofrecen herramientas mejoradas para predecir la expresión génica y analizar los efectos de las variantes genéticas.