Skin Tone Analysis for Representation in Educational Materials (STAR-ED) using machine learning Journal Article


Authors: Tadesse, G. A.; Cintas, C.; Varshney, K. R.; Staar, P.; Agunwa, C.; Speakman, S.; Jia, J.; Bailey, E. E.; Adelekun, A.; Lipoff, J. B.; Onyekaba, G.; Lester, J. C.; Rotemberg, V.; Zou, J.; Daneshjou, R.
Article Title: Skin Tone Analysis for Representation in Educational Materials (STAR-ED) using machine learning
Abstract: Images depicting dark skin tones are significantly underrepresented in the educational materials used to teach primary care physicians and dermatologists to recognize skin diseases. This could contribute to disparities in skin disease diagnosis across different racial groups. Previously, domain experts have manually assessed textbooks to estimate the diversity in skin images. Manual assessment does not scale to many educational materials and introduces human errors. To automate this process, we present the Skin Tone Analysis for Representation in EDucational materials (STAR-ED) framework, which assesses skin tone representation in medical education materials using machine learning. Given a document (e.g., a textbook in.pdf), STAR-ED applies content parsing to extract text, images, and table entities in a structured format. Next, it identifies images containing skin, segments the skin-containing portions of those images, and estimates the skin tone using machine learning. STAR-ED was developed using the Fitzpatrick17k dataset. We then externally tested STAR-ED on four commonly used medical textbooks. Results show strong performance in detecting skin images (0.96 ± 0.02 AUROC and 0.90 ± 0.06 F1 score) and classifying skin tones (0.87 ± 0.01 AUROC and 0.91 ± 0.00 F1 score). STAR-ED quantifies the imbalanced representation of skin tones in four medical textbooks: brown and black skin tones (Fitzpatrick V-VI) images constitute only 10.5% of all skin images. We envision this technology as a tool for medical educators, publishers, and practitioners to assess skin tone diversity in their educational materials. © 2023, Springer Nature Limited.
Keywords: automation; publication; medical education; medical imaging; diagnosis; physician; skin disease; task performance; dermatology; primary care; book; machine learning; human errors; pipeline; human; article; domain experts; machine-learning; disease diagnosis; stars; health educator; skin tone; education material; educational materials; skin images; skin tone analysis for representation in educational material; textbooks
Journal Title: npj Digital Medicine
Volume: 6
ISSN: 2398-6352
Publisher: Nature Publishing Group  
Date Published: 2023-08-18
Start Page: 151
Language: English
DOI: 10.1038/s41746-023-00881-0
PROVIDER: scopus
PMCID: PMC10439178
PUBMED: 37596324
DOI/URL:
Notes: The MSK Cancer Center Support Grant (P30 CA008748) is acknowledged in the PDF -- Source: Scopus
Altmetric
Citation Impact
BMJ Impact Analytics
MSK Authors