What Are We Measuring When We Evaluate Digital Interventions for Improving Lifestyle? A Scoping Meta-Review
Bibliographic record
Abstract
Background: Lifestyle Medicine (LM) aims to address six main behavioral domains: diet/nutrition, substance use (SU), physical activity (PA), social relationships, stress management, and sleep. Digital Health Interventions (DHIs) have been used to improve these domains. However, there is no consensus on how to measure lifestyle and its intermediate outcomes aside from measuring each behavior separately. We aimed to describe (1) the most frequent lifestyle domains addressed by DHIs, (2) the most frequent outcomes used to measure lifestyle changes, and (3) the most frequent DHI delivery methods. Methods: We followed the Preferred Reporting Items for Systematic Reviews and Meta-Analyses (PRISMA-ScR) Extension for Scoping Reviews. A literature search was conducted using MEDLINE, Cochrane Library, EMBASE, and Web of Science for publications since 2010. We included systematic reviews and meta-analyses of clinical trials using DHI to promote health, behavioral, or lifestyle change. Results: Overall, 954 records were identified, and 72 systematic reviews were included. Of those, 35 conducted meta-analyses, 58 addressed diet/nutrition, and 60 focused on PA. Only one systematic review evaluated all six lifestyle domains simultaneously; 1 systematic review evaluated five lifestyle domains; 5 systematic reviews evaluated 4 lifestyle domains; 14 systematic reviews evaluated 3 lifestyle domains; and the remaining 52 systematic reviews evaluated only one or two domains. The most frequently evaluated domains were diet/nutrition and PA. The most frequent DHI delivery methods were smartphone apps and websites. Discussion: The concept of lifestyle is still unclear and fragmented, making it hard to evaluate the complex interconnections of unhealthy behaviors, and their impact on health. Clarifying this concept, refining its operationalization, and defining the reporting guidelines should be considered as the current research priorities. DHIs have the potential to improve lifestyle at primary, secondary, and tertiary levels of prevention—but most of them are targeting clinical populations. Although important advances have been made to evaluate DHIs, some of their characteristics, such as the rate at which they become obsolete, will require innovative research designs to evaluate long-term outcomes in health.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.014 | 0.002 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.002 | 0.001 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.004 | 0.000 |
| Scholarly communication | 0.000 | 0.001 |
| Open science | 0.001 | 0.001 |
| Research integrity | 0.000 | 0.002 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".