Standardizing definitions and reporting guidelines for the infertility core outcome set: an international consensus development study
Bibliographic record
Abstract
Study Question Can consensus definitions for the core outcome set for infertility be identified in order to recommend a standardized approach to reporting? Summary Answer Consensus definitions for individual core outcomes, contextual statements, and a standardized reporting table have been developed. What is Known Already Different definitions exist for individual core outcomes for infertility. This variation increases the opportunities for researchers to engage with selective outcome reporting, which undermines secondary research and compromises clinical practice guideline development. Study Design, Size, Duration Potential definitions were identified by a systematic review of definition development initiatives and clinical practice guidelines and by reviewing Cochrane Gynaecology and Fertility Group guidelines. These definitions were discussed in a face-to-face consensus development meeting, which agreed consensus definitions. A standardized approach to reporting was also developed as part of the process. Participants/Materials, Setting, Methods Healthcare professionals, researchers, and people with fertility problems were brought together in an open and transparent process using formal consensus development methods. Main Results and the Role of Chance Forty-four potential definitions were inventoried across four definition development initiatives, including the Harbin Consensus Conference Workshop Group and International Committee for Monitoring Assisted Reproductive Technologies, 12 clinical practice guidelines, and Cochrane Gynaecology and Fertility Group guidelines. Twenty-seven participants, from 11 countries, contributed to the consensus development meeting. Consensus definitions were successfully developed for all core outcomes. Specific recommendations were made to improve reporting. Limitations, Reasons for Caution We used consensus development methods, which have inherent limitations. There was limited representation from low- and middle-income countries. Wider Implications of the Findings A minimum data set should assist researchers in populating protocols, case report forms, and other data collection tools. The generic reporting table should provide clear guidance to researchers and improve the reporting of their results within journal publications and conference presentations. Research funding bodies, the Standard Protocol Items: Recommendations for Interventional Trials statement, and over 80 specialty journals have committed to implementing this core outcome set. Study Funding/Competing Interest(s) This research was funded by the Catalyst Fund, Royal Society of New Zealand, Auckland Medical Research Fund, and Maurice and Phyllis Paykel Trust. Siladitya Bhattacharya reports being the Editor-in-Chief of Human Reproduction Open and an editor of the Cochrane Gynaecology and Fertility group. Hans Evers reports being the Editor Emeritus of Human Reproduction. Richard Legro reports consultancy fees from Abbvie, Bayer, Ferring, Fractyl, Insud Pharma and Kindex and research sponsorship from Guerbet and Hass Avocado Board. Ben Mol reports consultancy fees from Guerbet, iGenomix, Merck, Merck KGaA and ObsEva. Craig Niederberger reports being the Editor-in-Chief of Fertility and Sterility and Section Editor of the Journal of Urology, research sponsorship from Ferring, and a financial interest in NexHand. Ernest Ng reports research sponsorship from Merck. Annika Strandell reports consultancy fees from Guerbet. Jack Wilkinson reports being a statistical editor for the Cochrane Gynaecology and Fertility group. Andy Vail reports that he is a Statistical Editor of the Cochrane Gynaecology & Fertility Review Group and of the journal Reproduction. His employing institution has received payment from HFEA for his advice on review of research evidence to inform their ‘traffic light' system for infertility treatment ‘add-ons'. Lan Vuong reports consultancy and conference fees from Ferring, Merck and Merck Sharp and Dohme. The remaining authors declare no competing interests in relation to the work presented. All authors have completed the disclosure form. Trial Registration Number Core Outcome Measures in Effectiveness Trials Initiative: 1023.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.743 | 0.775 |
| Meta-epidemiology (narrow) | 0.002 | 0.003 |
| Meta-epidemiology (broad) | 0.006 | 0.010 |
| Bibliometrics | 0.022 | 0.018 |
| Science and technology studies | 0.006 | 0.009 |
| Scholarly communication | 0.014 | 0.018 |
| Open science | 0.012 | 0.023 |
| Research integrity | 0.006 | 0.012 |
| Insufficient payload (model declined to judge) | 0.005 | 0.002 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; the direct Gemma label and the distilled Codex classifier agree on what is shown here.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".