MétaCan
Menu
Back to cohort
Record W4414346927 · doi:10.2196/69420

Potential of the World Health Organization’s Skin NTDs App to Support and Improve the Detection of Skin-Related Neglected Tropical Diseases: Protocol for a Performance Evaluation and Feasibility Study in Senegal

2025· article· en· W4414346927 on OpenAlexvenueno aff
Dior Sall, Dominik Jockers, Pauline Dioussé, Jonas Wachinger, Gilbert Batista, José A Ruiz-Postigo, Laurène Petitfour, Charlotte Robert, Bachir Diallo, Fulgence Abdou Faye, Y Dieng, Maresa Neuerer, A Lawson, Felicitas Schwermann, Carme Carrión, Louis Hyacinthe Zoubi, Papa Mamadou Diagne, Christa Kasang, Fatou Ndiaye Oumar Sy, Mahamath Cissé, Till Bärnighausen

Bibliographic record

VenueJMIR Research Protocols · 2025
Typearticle
Languageen
FieldMedicine
TopicCutaneous Melanoma Detection and Management
Canadian institutionsnot available
FundersWorld Health Organization
KeywordsProtocol (science)Neglected tropical diseasesTropical diseasemHealthData collectionPublic healthGlobal healthDeveloping country

Abstract

fetched live from OpenAlex

Background The World Health Organization (WHO) roadmap aims to control, eliminate, or eradicate neglected tropical diseases (NTDs) by promoting innovation in prevention, diagnosis, and treatment. In this context, mobile health (mHealth) tools could play an important role in improving health care across the globe, including for skin-related NTDs. One such tool is the WHO Skin NTDs App (currently available in its beta version), which utilizes artificial intelligence (AI) algorithms to classify skin lesion images and offers diagnostic suggestions and management information to bolster early detection at primary care levels. However, to harness the full potential of this and similar mHealth tools, additional insights into their diagnostic performance and potential implementation avenues in settings with limited access to trained dermatologists are essential. Objective The objective of our mixed methods study is to test the functionality, operability, and potential of the AI-supported diagnostic component of the WHO Skin NTDs App (beta version) to support the detection of skin NTDs and common skin conditions in Senegal. Methods We are conducting a diagnostic accuracy study combined with a qualitative preimplementation usability exploration. For the quantitative component, we will collect and analyze approximately 800 skin lesion images from patients presenting to the dermatology unit at the Thiès regional hospital in Senegal. Each lesion will be independently assessed by the AI-based WHO Skin NTDs App and by a dermatologist who will provide a diagnosis serving as the reference standard. Performance metrics, including accuracy, sensitivity, specificity, precision, F1-score, and area under the receiver operating characteristic curve, will be calculated for each diagnostic category to evaluate the app’s ability to detect skin-related NTDs. In parallel, we will conduct semistructured in-depth interviews with a purposive sample of 70-80 stakeholders, including policymakers, health care workers, community leaders, dermatologists, and members of leprosy-affected communities. Interviews will explore perceptions of the app’s usability, acceptability, and potential barriers and facilitators to its adoption within Senegal’s health system. Thematic analysis will be used to interpret qualitative data. Findings will help inform the design of an app-based intervention to be piloted in future community-level studies. Results We expect the results to provide detailed insights into the feasibility and potential of the WHO Skin NTDs App to support and improve the detection of skin NTDs and common skin conditions at the community level in Senegal. We started data collection in August 2024, with the first results expected to be available in 2025. Conclusions Our study will assess the performance and potential use of the WHO Skin NTDs App to detect skin NTDs and common skin conditions in Senegal, outlining its potential role in supporting early diagnoses and enhancing public health responses. Trial Registration German Clinical Trials Register DRKS00034297; https://drks.de/search/de/trial/DRKS00034297/details International Registered Report Identifier (IRRID) DERR1-10.2196/69420

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame machine prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.

metaresearch head score (Codex)0.063
metaresearch head score (Gemma)0.056
Version: metacan-v3-hybrid-931329e0061cValidation status: machine_predicted_unvalidated
Candidate categoriesnone
Consensus categoriesnone
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Not applicable · Consensus signal: none
GenreCandidate signal: Protocol · Consensus signal: Protocol
Teacher disagreement score0.063
Threshold uncertainty score0.334

Distilled classifier scores by category (both heads)

CategoryCodexGemma
Metaresearch0.0630.056
Meta-epidemiology (narrow)0.0030.002
Meta-epidemiology (broad)0.0020.004
Bibliometrics0.0020.002
Science and technology studies0.0040.003
Scholarly communication0.0030.002
Open science0.0030.003
Research integrity0.0030.004
Insufficient payload (model declined to judge)0.0220.006

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.064
GPT teacher head0.486
Teacher spread0.422 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.

The models applied no category: nothing in the taxonomy fit this work.
Study designNot applicable
Domainnot available
GenreProtocol

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations2
Published2025
Admission routes1
Has abstractyes

Explore more

Same venueJMIR Research ProtocolsSame topicCutaneous Melanoma Detection and ManagementFrench-language works237,207