Feasibility Study on Menstrual Cycles With Fitbit Device (FEMFIT): Prospective Observational Cohort Study
Bibliographic record
Abstract
BACKGROUND: Despite its importance to women's reproductive health and its impact on women's daily lives, the menstrual cycle, its regulation, and its impact on health remain poorly understood. As conventional clinical trials rely on infrequent in-person assessments, digital studies with wearable devices enable the collection of longitudinal subjective and objective measures. OBJECTIVE: The study aims to explore the technical feasibility of collecting combined wearable and digital questionnaire data and its potential for gaining biological insights into the menstrual cycle. METHODS: This prospective observational cohort study was conducted online over 12 weeks. A total of 42 cisgender women were recruited by their local gynecologist in Berlin, Germany, and given a Fitbit Inspire 2 device and access to a study app with digital questionnaires. Statistical analysis included descriptive statistics on user behavior and retention, as well as a comparative analysis of symptoms from the digital questionnaires with metrics from the sensor devices at different phases of the menstrual cycle. RESULTS: The average time spent in the study was 63.3 (SD 33.0) days with 9 of the 42 individuals dropping out within 2 weeks of the start of the study. We collected partial data from 114 ovulatory cycles, encompassing 33 participants, and obtained complete data from a total of 50 cycles. Participants reported a total of 2468 symptoms in the daily questionnaires administered during the luteal phase and menses. Despite difficulties with data completeness, the combined questionnaire and sensor data collection was technically feasible and provided interesting biological insights. We observed an increased heart rate in the mid and end luteal phase compared with menses and participants with severe premenstrual syndrome walked substantially fewer steps (average daily steps 10,283, SD 6277) during the luteal phase and menses compared with participants with no or low premenstrual syndrome (mean 11,694, SD 6458). CONCLUSIONS: We demonstrate the feasibility of using an app-based approach to collect combined wearable device and questionnaire data on menstrual cycles. Dropouts in the early weeks of the study indicated that engagement efforts would need to be improved for larger studies. Despite the challenges of collecting wearable data on consecutive days, the data collected provided valuable biological insights, suggesting that the use of questionnaires in conjunction with wearable data may provide a more complete understanding of the menstrual cycle and its impact on daily life. The biological findings should motivate further research into understanding the relationship between the menstrual cycle and objective physiological measurements from sensor devices.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".