Embeddings de modelos de lenguaje como buenos aprendices para el análisis de datos de células únicas

Tianyu Liu1,2, Tianqi Chen2, Wangjie Zheng2

  • 1Interdepartmental Program in Computational Biology & Bioinformatics, Yale University, New Haven, CT 06511, USA.

Patterns (New York, N.Y.)
|February 23, 2026
PubMed
Resumen

scELMo aprovecha los modelos de lenguaje grandes (LLM) para analizar datos de células únicas, lo que permite el agrupamiento celular, la anotación y el análisis de perturbaciones sin necesidad de entrenar nuevos modelos. Este método ofrece un enfoque eficiente en recursos para la interpretación de datos de células únicas.

Videos de Conceptos Relacionados

Improving Translational Accuracy02:07

Improving Translational Accuracy

Base complementarity between the three base pairs of mRNA codon and the tRNA anticodon is not a failsafe mechanism. Inaccuracies can range from a single mismatch to no correct base pairing at all. The free energy difference between the correct and nearly correct base pairs can be as small as 3 kcal/ mol. With complementarity being the only proofreading step, the estimated error frequency would be one wrong amino acid in every 100 amino acids incorporated. However, error frequencies observed in...
Synthetic Biology02:55

Synthetic Biology

Synthetic biology is an interdisciplinary science that involves using principles from disciplines such as engineering, molecular biology, cell biology, and systems biology. It involves remodeling existing organisms from nature or constructing completely new synthetic organisms for applications such as protein or enzyme production, bioremediation, value-added macromolecule production, and the addition of desirable traits to crops, to name a few.
Golden rice
Golden rice is a genetically modified...
Improving Translational Accuracy02:07

Improving Translational Accuracy

Base complementarity between the three base pairs of mRNA codon and the tRNA anticodon is not a failsafe mechanism. Inaccuracies can range from a single mismatch to no correct base pairing at all. The free energy difference between the correct and nearly correct base pairs can be as small as 3 kcal/ mol. With complementarity being the only proofreading step, the estimated error frequency would be one wrong amino acid in every 100 amino acids incorporated. However, error frequencies observed in...