Validation of an Anti-Müllerian Hormone Cutoff for Polycystic Ovarian Morphology in the Diagnosis of Polycystic Ovary Syndrome in the HARMONIA Study: Protocol for a Prospective, Noninterventional Study
Bibliographic record
Abstract
BACKGROUND: Polycystic ovary syndrome (PCOS) is one of the most common endocrine disorders in women and is diagnosed using the Rotterdam criteria, including diagnosis of polycystic ovarian morphology (PCOM) by transvaginal ultrasound (TVUS). Due to high cost, availability, and the impact of the operator and ultrasound equipment on the reliability of the antral follicle count (AFC) by TVUS, an unmet need exists for a diagnostic test to determine PCOM without TVUS. A strong positive correlation between elevated anti-Müllerian hormone (AMH) levels and AFCs has been demonstrated in women with PCOS. In addition, recent updates to the international evidence-based PCOS guidelines state that serum AMH can be used as an alternative to TVUS-determined AFC, in the diagnosis of PCOM. The retrospective APHRODITE study derived and validated an AMH cutoff of 3.2 ng/mL for the Elecsys AMH Plus or Elecsys AMH assays (Roche) to diagnose PCOM in patients with PCOS. OBJECTIVE: This study aims to further validate, in an independent prospective cohort, the AMH cutoff (3.2 ng/mL) for PCOM determination, which was previously derived and validated in the APHRODITE study. METHODS: This large, prospective, multicenter, population-based, noninterventional study will evaluate the previously established AMH cutoff for the determination of PCOM during the diagnosis of PCOS using the Elecsys AMH Plus immunoassay in an independent population. Participants were women born between July 1985 and December 1987 in Northern Finland; the study partially links to the Northern Finland Birth Cohort 1986. We assessed the enrolled women, determined with the 2023 PCOS Guidelines, for current PCOS status and divided them by phenotype if positive. Each participant had 1 study visit to collect serum samples, record clinical data, and undergo a gynecological examination including TVUS. All data were collected by highly trained midwives or trained gynecologists. Sensitivity, specificity, and agreement measures were used to validate the previously determined cutoff in the whole population and in subpopulations based on phenotype and relevant demographic or clinical factors. The minimum target sample size was approximately 1800 women, including approximately 10% with PCOS. RESULTS: At the time of manuscript submission, participant recruitment had concluded, and 1803 women were enrolled into the study. Data collection is complete and biostatistical analysis is planned for 2023. CONCLUSIONS: To limit variability, there were few TVUS operators and only 2 TVUS machines of the same type. Additionally, all women who were taking oral contraceptives were excluded from the primary analysis population. Selection bias was limited as this was a population-based study and participants were not seeking treatment for PCOS symptoms. Validating the AMH cutoff in a large, population-based study will provide further evidence on the utility of the Elecsys AMH Plus or Elecsys AMH assays in PCOM diagnosis as an alternative to TVUS. Measuring AMH for PCOM diagnosis could reduce delayed or missed diagnoses due to operator-dependent TVUS examinations. TRIAL REGISTRATION: ClinicalTrials.gov NCT05527353; http://tinyurl.com/2f3ffbdz. INTERNATIONAL REGISTERED REPORT IDENTIFIER (IRRID): DERR1-10.2196/48854.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.007 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.001 | 0.002 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.001 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".