Assessment of Patient-Reported Outcome Measures for Maternal Postpartum Depression Using the Consensus-Based Standards for the Selection of Health Measurement Instruments Guideline
Bibliographic record
Abstract
Importance: Maternal depression is frequently reported in the postpartum period, with an estimated prevalence of approximately 15% during the first postpartum year. Despite the high prevalence of postpartum depression, there is no consensus regarding which patient-reported outcome measure (PROM) should be used to screen for this complex, multidimensional construct. Objective: To evaluate psychometric measurement properties of existing PROMs of maternal postpartum depression using the Consensus-Based Standards for the Selection of Health Measurement Instruments (COSMIN) guideline and identify the best available patient-reported screening measure. Evidence Review: This systematic review followed the Preferred Reporting Items for Systematic Reviews and Meta-analyses (PRISMA) reporting guideline. PubMed, CINAHL, Embase, and Web of Science were searched on July 1, 2019, for validated PROMs of postpartum depression, and an additional search including a hand search of references from eligible studies was conducted in June 2021. Included studies evaluated 1 or more psychometric measurement properties of the identified PROMs. A risk-of-bias assessment was performed to evaluate methods of each included study. Psychometric measurement properties of each PROM were rated according to COSMIN criteria. A modified Grading of Recommendations Assessment, Development, and Evaluation approach was used to assess the level of evidence supporting each rating, and a recommendation class (A, recommended for use; B, further research required; or C, not recommended) was given based on the overall quality of each included PROM. Findings: Among 10 264 postpartum recovery studies, 27 PROMs were identified. Ten PROMs (37.0%) met the inclusion criteria and were used in 43 studies (0.4%) involving 22 095 postpartum women. At least 1 psychometric measurement property was assessed for each of the 10 validated PROMs identified. Content validity was sufficient in all PROMs. The Edinburgh Postnatal Depression Scale (EPDS) demonstrated adequate content validity and a moderate level of evidence for sufficient internal consistency (with sufficient structural validity), resulting in a recommendation of class A. The other 9 PROMs evaluated received a recommendation of class B. Conclusions and Relevance: The findings of this systematic review suggest that the EPDS is the best available patient-reported screening measure of maternal postpartum depression. Future studies should focus on evaluating the cross-cultural validity, reliability, and measurement error of the EPDS to improve understanding of its psychometric properties and utility.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.011 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.002 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.001 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.001 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".