Related Experiment Video
Updated: Jun 14, 2025

Author Spotlight: Developing a Point-of-Care Hemoglobin Estimation Method for Anemia Management
Published on: January 19, 2024
Content Analysis of Social Determinants of Health Accelerator Plans Using Artificial Intelligence: A Use Case for
Kelli DePriest1, John Feher Iii, Kailen Gore
1Author Affiliations: Research Triangle Institute: RTI International, Research Triangle Park, North Carolina (Dr DePriest, Mr Feher III, Ms Gore, Dr Glasgow, and Mr Chew); Association of State and Territorial Health Officials (ASTHO), Arlington, Virginia (Mr Grant); National Association of County and City Health Officials (NACCHO), Washington, District of Columbia (Mr Holtgrave); Centers for Disease Control and Prevention (CDC), Atlanta, Georgia (Dr Hacker).
Context:
Public health practice involves the development of reports and plans, including funding progress reports, strategic plans, and community needs assessments. These documents are valuable data sources for program monitoring and evaluation. However, practitioners rarely have the bandwidth to thoroughly and rapidly review large amounts of primarily qualitative data to support real-time and continuous program improvement. Systematically examining and categorizing qualitative data through content analysis is labor-intensive. Large language models (LLMs), a type of generative artificial intelligence (AI) focused on language-based tasks, hold promise for expediting content analysis of public health documents, which, in turn, could facilitate continuous program improvement.
Objectives:
To explore the feasibility and potential of using LLMs to expedite content analysis of real-world public health documents. The focus was on comparing semiautomated outputs from GPT-4o with human outputs for abstracting and synthesizing information from health improvement plans.
Design:
Our study team conducted a content analysis of 4 publicly available community health improvement plans and compared the results with GPT-4o's performance on 20 data elements. We also assessed the resources required for both methods, including time spent on prompt engineering and error correction.
Main Outcome Measures:
Accuracy of data abstraction and time required.
Results:
GPT-4o demonstrated abstraction accuracy of 79% (n = 17 errors) compared to 94% accuracy by the study team for individual plans, with 8 instances of falsified data. Out of the 18 synthesis data elements, GPT-4o made 9 errors, demonstrating an accuracy of 50%. On average, GPT-4o abstraction required fewer hours than study team abstraction, but resource savings diminished when accounting for time for developing prompts and identifying/correcting errors.
Conclusions:
Public health professionals who explore the use of generative AI tools should approach the method with cautious curiosity and consider the potential tradeoffs between resource savings and data accuracy.
More Related Videos
Related Concept Videos
Issues And Trends In Healthcare Delivery System
Cost Containment
Payment for healthcare services has historically promoted adoption of costly and often unnecessary or inefficient...
Methods Of Healthcare Delivery System
Managed Care System:
The managed care system is designed to control the cost while maintaining the quality of care. The patient's care from admission to discharge is planned by the primary care provider or the case manager, also known as the gatekeeper. In a managed care system, the number of care providers is...
Primary Healthcare Services
In 1978, international leaders convened in Alma-Ata, Kazakhstan, for what would be a pivotal event in global health. The Alma-Ata Declaration was the first to call...
Preventive Healthcare Services
Models of Health Promotion and Illness Prevention I
The health belief model (HBM) attempts to predict health-related behavior in specific belief patterns. According to the HBM, a person's...
Steps in Outbreak Investigation

