Thyroid function testing and management during and after pregnancy among women without thyroid disease before pregnancy
Bibliographic record
Abstract
BACKGROUND: Screening in pregnancy for subclinical hypothyroidism, often defined as thyroid-stimulating hormone (TSH) greater than 2.5 mIU/L or greater than 4.0 mIU/L, is controversial. We determined the frequency and distribution of TSH testing by gestational age, as well as TSH values associated with treatment during pregnancy and the frequency of postpartum continuation of thyroid hormone therapy. METHODS: We performed a retrospective cohort study of pregnancies in Alberta, Canada. We included women without thyroid disease who delivered between October 2014 and September 2017. We used delivery records, physician billings, and pharmacy and laboratory administrative data. Our key outcomes were characteristics of TSH testing and the initiation and continuation of thyroid hormone therapy. We calculated the proportion of pregnancies with thyroid testing and the frequency of each specific thyroid test. RESULTS: Of the 188 490 pregnancies included, 111 522 (59.2%) had at least 1 TSH measurement. The most common time for testing was at gestational week 5 to 6. Thyroid hormone therapy was initiated at a median gestational age of 7 (interquartile range 5-12) weeks. Among women with first TSH measurements of 4.01 to 9.99 mIU/L who were not immediately treated, the repeat TSH measurement was 4.00 mIU/L or below in 67.9% of pregnancies. Thyroid hormone was continued post partum for 44.6% of the women who started therapy during their pregnancy. INTERPRETATION: The findings of our study suggest that current practice patterns may contribute to overdiagnosis of hypothyroidism and overtreatment during pregnancy and post partum.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.006 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.001 | 0.000 |
| Open science | 0.001 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".