Predicting Clinically Significant Weight Loss in a Multimodal Commercial Digital Weight Management Program: Machine Learning Approach
Bibliographic record
Abstract
Background Generally, adherence to diet and physical activity in behavioral weight management programs predicts weight loss. Moreover, meeting attendance and platform or app use predict weight loss. Many weight-loss–prediction studies primarily use regression. However, little is known about the levels of adherence to specific web-based weight management program tools as well as behavioral and psychosocial variables that predict significant weight loss using machine-learning approaches. Objective In this paper, we aimed to examine variables that can predict clinically significant weight loss (≥5%) in a multimodal commercial digital weight management program, including levels of adherence to intervention tools and changes in behavioral and psychosocial states using machine-learning approaches. Methods We performed secondary analyses using data from a one-arm trial online WW (formerly Weight Watchers) weight-loss intervention program, lasting 6 months, that recruited US adults with a BMI range of 25-45 kg/m2. We used a WW Bluetooth scale and digital intervention tools such as mobile app for point tracking, weekly virtual workshops, weekly wellness check-ins and a Facebook group, and changes in psychosocial and behavioral variables (ie, food craving, Pittsburgh sleep quality index, diet, hunger, and physical activity). Using the receiver operating characteristics (ROC) curve, we identified the predictors of significant weight loss, as well as the associated cut points (CP) and area-under-curve (AUC) values for each variable. We further used a classification tree to confirm the importance of these predictors and assessed the out-of-sample prediction accuracies using 5-fold cross-validation. Results Participants (N=153) were 70% female and 66% White, with a mean age of 41.09 (SD 13.78) years, and had a mean BMI of 31.8 kg/m2 (SD 5.0). Approximately 51% of participants lost ≥5% weight. Using ROC curve, food tracking (CP≥9.4%; AUC=0.744; P<.001), increase in self-weighing (CP≥0.5; AUC=0.733; P<.001), wellness check-in attendance (CP≥31.3%; AUC=0.723; P<.001), and workshop attendance (CP≥35.4%; AUC=0.699; P<.001) were identified as significant predictors of achieving ≥5% weight loss. We confirmed the importance of these variables using classification tree, and together they achieved an out-of-sample prediction accuracy of 78%. Conclusions ROC curve and classification tree provided consistent results of predictors of clinically significant weight loss (ie, food tracking, increase in self-weighing, workshop attendance, and wellness check-in attendance). Our study extends the weight management literature by using machine-learning approaches to identify significant weight loss predictors and specific levels of these predictors needed to achieve clinically significant weight loss. Conflicts of Interest None declared.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.002 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.000 | 0.002 |
| Science and technology studies | 0.001 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.002 |
| Insufficient payload (model declined to judge) | 0.000 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".