Related Experiment Video
Updated: Sep 19, 2025

Eye Tracking During A Complex Aviation Task For Insights Into Information Processing
Published on: April 4, 2025
Empirically derived evaluation requirements for responsible deployments of AI in safety-critical settings
Dane A Morey1, Michael F Rayo2, David D Woods2
1The Ohio State University, Columbus, OH, USA. morey.38@osu.edu.
Abstract:
Processes to assure the safe, effective, and responsible deployment of artificial intelligence (AI) in safety-critical settings are urgently needed. Here we show a procedure to empirically evaluate the impacts of AI augmentation as a basis for responsible deployment. We evaluated three augmentative AI technologies nurses used to recognize imminent patient emergencies, including combinations of AI recommendations and explanations. The evaluation involved 450 nursing students and 12 licensed nurses assessing 10 historical patient cases. With each technology, nurses' performance was both improved and degraded when the AI algorithm was most correct and misleading, respectively. Our findings caution that AI capabilities alone do not guarantee a safe and effective joint human-AI system. We propose two minimum requirements for evaluating AI in safety-critical settings: (1) empirically measure the performance of people and AI together and (2) examine a range of challenging cases which produce a range of strong, mediocre, and poor AI performance.
More Related Videos
Related Concept Videos
Ethics in Research
Psychosurgery
Historical Development of Psychosurgery
In the 1930s, Portuguese neurologist Antonio Egas Moniz introduced a surgical procedure designed...
Types of Biopharmaceutical Studies: Controlled and Non-Controlled Approaches
Non-controlled studies, commonly employed for initial exploration, lack a control group, rendering them susceptible to biases and external influences. In contrast,...
Distribution Reliability and Automation
Current Trends in Nursing II

