Informing the Physical Activity Evaluation Framework: A Scoping Review of Reviews
Bibliographic record
Abstract
OBJECTIVE: Robust program evaluations can identify effective promotion strategies. This scoping review aimed to analyze review articles (including systematic reviews, meta-analysis, meta-synthesis, scoping review, narrative review, rapid review, critical review, and integrative reviews) to systematically map and describe physical activity program evaluations published between January 2014 and July 2020 to summarize key characteristics of the published literature and suggest opportunities to strengthen current evaluations. DATA SOURCE: We conducted a systematic search of the following databases: Medline, Scopus, Sportdiscus, Eric, PsycInfo, and CINAHL. INCLUSION/EXCLUSION CRITERIA: Abstracts were screened for inclusion based on the following criteria: review article, English language, human subjects, primary prevention focus, physical activity evaluation, and evaluations conducted in North America. EXTRACTION: Our initial search yielded 3193 articles; 211 review articles met the inclusion criteria. SYNTHESIS: We describe review characteristics, evaluation measures, and "good practice characteristics" to inform evaluation strategies. RESULTS: Many reviews (72%) did not assess or describe the use of an evaluation framework or theory in the primary articles that they reviewed. Among those that did, there was significant variability in terminology making comparisons difficult. Process indicators were more common than outcome indicators (63.5% vs 46.0%). There is a lack of attention to participant characteristics with 29.4% capturing participant characteristics such as race, income, and neighborhood. Negative consequences from program participation and program efficiency were infrequently considered (9.3% and 13.7%). CONCLUSION: Contextual factors, negative outcomes, the use of evaluation frameworks, and measures of program sustainability would strengthen evaluations and provide an evidence-base for physical activity programming, policy, and funding.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.388 | 0.554 |
| Meta-epidemiology (narrow) | 0.005 | 0.005 |
| Meta-epidemiology (broad) | 0.015 | 0.012 |
| Bibliometrics | 0.067 | 0.051 |
| Science and technology studies | 0.005 | 0.007 |
| Scholarly communication | 0.019 | 0.027 |
| Open science | 0.007 | 0.011 |
| Research integrity | 0.008 | 0.007 |
| Insufficient payload (model declined to judge) | 0.005 | 0.002 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".