MétaCan
Menu
Back to cohort
Record W4414108032 · doi:10.2196/73040

Evaluating Diversity in Open Photoplethysmography Datasets: Protocol for a Systematic Review

2025· article· en· W4414108032 on OpenAlexvenueno aff
Vedha Penmetcha, Lekaashree Rambabu, Brandon Smith, Orla Mantle, Thomas Edmiston, Laura Hobbs, Shobhana Nagraj, Peter Charlton, Tom Bashford

Bibliographic record

VenueJMIR Research Protocols · 2025
Typearticle
Languageen
FieldEngineering
TopicNon-Invasive Vital Sign Monitoring
Canadian institutionsnot available
FundersNational Institute for Health and Care ResearchBritish Heart FoundationWellcome Trust
KeywordsProtocol (science)PhotoplethysmogramData collectionDiversity (politics)Open data

Abstract

fetched live from OpenAlex

Background: Photoplethysmography (PPG) is an optical method for measuring blood volume changes in microcirculation through noninvasive photodetection. It has become a widespread and essential clinical tool, used in pulse oximeters and wearable devices. However, technical aspects of PPG make it susceptible to intrinsic bias, with the potential to adversely affect particular patient and consumer populations. Developments in PPG technology, increasingly driven by openly accessible datasets as opposed to de novo experimentation, have the potential to help monitor an array of physiological variables. However, some populations may be underrepresented in PPG datasets. We describe a protocol for a systematic review to assess the biases within open access PPG datasets. Objective: This review aims to evaluate the underlying reporting patterns and structure of openly accessible PPG datasets. We will provide insight into the measured biosignals and demographic variables included in the datasets in the hope of shedding light on what PPG data parameters are being used to develop medical devices. Therefore, we can elucidate current gaps and areas for improvement to reduce bias in medical device development. Methods: This review will be reported in accordance with the standard PRISMA (Preferred Reporting Items for Systematic Reviews and Meta-Analyses) guidelines. We will include primary studies that mention PPG and specifically reference openly accessible datasets since 2000. The datasets must contain physiological parameters such as heart rate, blood pressure, or respiratory rate, as well as the PPG waveform data, collected from humans. Searches will be conducted in literature databases and data repositories, including MedLine OVID, IEEE Xplore, Scopus, and PhysioNet. Studies will be evaluated in accordance with the Standing Together Initiative recommendations, which are urging for health care technologies supported by representative data. Biosignal and demographic variables will be extracted from the PPG datasets, with steps taken to harmonize and store this information. Statistical analysis will be performed, including descriptive statistics and the chi-square test for comparisons. Additional statistical analyses will be performed after data extraction is completed and the level of heterogeneity is characterized. Results: We will analyze the dataset diversity and the structural basis of PPG datasets. This includes statistically analyzing the demographic and biosignal variables in the datasets. By using statistical test fit for nominal variable comparisons, we will evaluate the frequencies of characteristics like the devices used, biosignals collected, clinical parameters, demographic characteristics, and geographic information. This systematic review is expected to be completed by September 2025. The screening and review of the articles is currently being conducted. Conclusions: This review will provide insight into the potential gaps of existing open access PPG datasets. It will inform future data collection and design of openly available PPG datasets for training medical devices, including wearables, to avoid perpetuating biases, allowing for application in diverse clinical settings.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame distilled prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.

metaresearch head score (Codex)0.007
metaresearch head score (Gemma)0.001
Version: codex-gemma-dda1882f352aValidation status: machine_predicted_unvalidated
Candidate categoriesnone
Consensus categoriesnone
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Systematic review · Consensus signal: Systematic review
GenreCandidate signal: Protocol · Consensus signal: Protocol
Teacher disagreement score0.022
Threshold uncertainty score0.907

Codex and Gemma teacher scores by category

CategoryCodexGemma
Metaresearch0.0070.001
Meta-epidemiology (narrow)0.0000.000
Meta-epidemiology (broad)0.0010.000
Bibliometrics0.0010.002
Science and technology studies0.0000.000
Scholarly communication0.0000.001
Open science0.0020.002
Research integrity0.0000.001
Insufficient payload (model declined to judge)0.0000.000

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.469
GPT teacher head0.621
Teacher spread0.152 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; a candidate call from one teacher head, not a consensus.

The models applied no category: nothing in the taxonomy fit this work.
Study designSystematic review
Domainnot available
GenreProtocol

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations2
Published2025
Admission routes1
Has abstractyes

Explore more

Same venueJMIR Research ProtocolsSame topicNon-Invasive Vital Sign MonitoringFrench-language works237,207