Related Experiment Video
Updated: May 15, 2025

Augmenting Large Language Models via Vector Embeddings to Improve Domain-Specific Responsiveness
Published on: December 6, 2024
Encoding of pretrained large language models mirrors the genetic architectures of human psychological traits
Bohan Xu1,2, Nick Obradovich1, Wenjie Zheng3
1Laureate Institute for Brain Research, Tulsa, Oklahoma, USA.
Abstract:
Recent advances in large language models (LLMs) have prompted a frenzy in utilizing them as universal translators for biomedical terms. However, the black box nature of LLMs has forced researchers to rely on artificially designed benchmarks without understanding what exactly LLMs encode. We demonstrate that pretrained LLMs can already explain up to 51% of the genetic correlation between items from a psychometrically-validated neuroticism questionnaire, without any fine-tuning. For psychiatric diagnoses, we found disorder names aligned better with genetic relationships than diagnostic descriptions. Our results indicate the pretrained LLMs have encodings mirroring genetic architectures. These findings highlight LLMs' potential for validating phenotypes, refining taxonomies, and integrating textual and genetic data in mental health research.
Related Concept Videos
Human Genetics
The complex relationship between genetics and psychology is observable through common biological components such...
Inheritance
Each gene exists in pairs, and the combination of these genes from both parents forms an individual's genotype. This genotype is a blueprint of potential traits. Examples of genotype...
Introduction to Personality Psychology
Early Theories of Personality
The study of...
Polygenic Traits
Behavioral Genetics and Its Designs
The primary methodologies used in behavior genetics include family studies, twin studies, and adoption studies, each providing unique...
Introduction to Biological Bases of Psychology
The nervous system, the cornerstone of...

