MétaCan
Menu
← Back to cohort
Record W4413814670 · doi:10.2196/75270

Artificial Intelligence–Assisted Image Extraction in Neonatal Echocardiography for Congenital Heart Disease Diagnosis in Sub-Saharan Africa: Protocol for Model Development

2025· article· en· W4413814670 on OpenAlexvenueno aff
Aminkeng Zawuo Leke, Lionel Landry Sop Deffo, Yunkavi Sabastian Wirsiy, Thomas Aldersley, Thomas G. Day, Andrew P. King, Patrick McAllister, Michel N Maboh, John Lawrenson, Cabral Tantchou, Bernhard Kainz, Frank Casey, Raymond Bond, Dewar Finlay, Ngoe Kelson Tchinda, Obale Armstrong, Frunwi Ndeh Mugri, Liesl Zühlke, Helen Dolk

Bibliographic record

VenueJMIR Research Protocols · 2025
Typearticle
Languageen
FieldMedicine
TopicCongenital Heart Disease Studies
Canadian institutionsnot available
FundersNational Heart, Lung, and Blood Institute
KeywordsPreprintProtocol (science)MedicineArtificial intelligenceComputer sciencePathologyWorld Wide WebAlternative medicine

Abstract

fetched live from OpenAlex

BACKGROUND: Sub-Saharan Africa (SSA) bears the highest global burden of under-5 mortality, with congenital heart disease (CHD) as a major contributor. Despite advancements in high-income countries, CHD-related mortality in SSA remains largely unchanged due to limited diagnostic capacity and centralized health care. While pulse oximetry aids early detection, confirmation typically relies on echocardiography, a procedure constrained by a shortage of specialized personnel. Artificial intelligence (AI) offers a promising solution to bridge this diagnostic gap. OBJECTIVE: This study aims to develop an AI-assisted echocardiography system that enables nonexpert operators, such as nurses, midwives, and medical doctors, to perform basic cardiac ultrasound sweeps on neonates suspected of CHD and extract accurate cardiac images for remote interpretation by a pediatric cardiologist. METHODS: The study will use a 2-phase approach to develop a deep learning model for real-time cardiac view detection in neonatal echocardiography, utilizing data from St. Padre Pio Hospital in Cameroon and the Red Cross War Memorial Children's Hospital in South Africa to ensure demographic diversity. In phase 1, the model will be pretrained on retrospective data from nearly 500 neonates (0-28 days old). Phase 2 will fine-tune the model using prospective data from 1000 neonates, which include background elements absent in the retrospective dataset, enabling adaptation to local clinical environments. The datasets will consist of short and continuous echocardiographic video clips covering 10 standard cardiac views, as defined by the American Society of Echocardiography. The model architecture will leverage convolutional neural networks and convolutional long short-term memory layers, inspired by the interleaved visual memory framework, which integrates fast and slow feature extractors via a shared temporal memory mechanism. Video preprocessing, annotation with predefined cardiac view codes using Labelbox, and training with TensorFlow and PyTorch will be performed. Reinforcement learning will guide the dynamic use of feature extractors during training. Iterative refinement, informed by clinical input, will ensure that the model effectively distinguishes correct from incorrect views in real time, enhancing its usability in resource-limited settings. RESULTS: Retrospective data collection for the project began in September 2024, and to date, data from 308 babies have been collected and labeled. In parallel, the initial model framework has been developed and training initiated using a subset of the labeled data. The project is currently in the intensive execution phase, with all objectives progressing in parallel and final results expected within 10 months. CONCLUSIONS: The AI-assisted echocardiography model developed in this project holds promise for improving early CHD diagnosis and care in SSA and other low-resource settings. INTERNATIONAL REGISTERED REPORT IDENTIFIER (IRRID): DERR1-10.2196/75270.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame machine prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.

metaresearch head score (Codex)0.007
metaresearch head score (Gemma)0.020
Version: metacan-v3-hybrid-931329e0061cValidation status: machine_predicted_unvalidated
Candidate categoriesnone
Consensus categoriesnone
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Simulation or modeling · Consensus signal: none
GenreCandidate signal: Protocol · Consensus signal: none
Teacher disagreement score0.031
Threshold uncertainty score0.105

Distilled classifier scores by category (both heads)

CategoryCodexGemma
Metaresearch0.0070.020
Meta-epidemiology (narrow)0.0020.001
Meta-epidemiology (broad)0.0010.002
Bibliometrics0.0010.001
Science and technology studies0.0010.001
Scholarly communication0.0010.001
Open science0.0030.002
Research integrity0.0020.003
Insufficient payload (model declined to judge)0.0310.006

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.260
GPT teacher head0.536
Teacher spread0.276 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.

The models applied no category: nothing in the taxonomy fit this work.
Study designSimulation or modeling
Domainnot available
GenreProtocol

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations3
Published2025
Admission routes1
Has abstractyes

Explore more

Same venueJMIR Research Protocols→Same topicCongenital Heart Disease Studies→French-language works237,207→