Related Experiment Video
Updated: Sep 12, 2025

Augmenting Large Language Models via Vector Embeddings to Improve Domain-Specific Responsiveness
Published on: December 6, 2024
Enhancing and Disaggregating Native Hawaiian and Pacific Islander (NHPI) Data Using Natural Language Processing and
Benjamin Viernes1, Qiwei Gan2, Elizabeth E Hanchrow
1Center for Pacific Islander Veteran Health. US Dept of VA. Honolulu, HI, USA.
Abstract:
Native Hawaiian and Pacific Islander (NHPI) populations are often aggregated into broad racial categories, obscuring potential disparities. This study leverages an expanded race/ethnicity lexicon and natural language processing (NLP) to identify documentation of NHPI subgroups to address gaps in electronic health records' (EHRs) recorded race. Results demonstrate the potential of NLP to classify NHPI documentation, disaggregate legacy categories, and improve health equity by incorporating more detailed subgroup data into standardized healthcare data sets.
More Related Videos
09:09Foreign Accent and Forensic Speaker Identification in Voice Lineups: The Influence of Acoustic Features Based on Prosody
Published on: September 27, 2024
09:20Cloud-Based Phrase Mining and Analysis of User-Defined Phrase-Category Association in Biomedical Publications
Published on: February 23, 2019
Related Concept Videos
RACE - Rapid Amplification of cDNA Ends
Stereotypes, Prejudice, and Discrimination
The Nativist Approach
Surveys
Ethnic Identity within a Larger Culture
How Data are Classified: Categorical Data
Data are classified based on whether they are measurable or not. Categorical data cannot be measured; instead, it can be divided into categories. For example, if Y denotes a person's party affiliation, some examples of Y include...