Examining the Effectiveness of Gamification in Mental Health Apps for Depression: Systematic Review and Meta-analysis
Bibliographic record
Abstract
BACKGROUND: Previous research showed that computerized cognitive behavioral therapy can effectively reduce depressive symptoms. Some mental health apps incorporate gamification into their app design, yet it is unclear whether features differ in their effectiveness to reduce depressive symptoms over and above mental health apps without gamification. OBJECTIVE: The aim of this study was to determine whether mental health apps with gamification elements differ in their effectiveness to reduce depressive symptoms when compared to those that lack these elements. METHODS: A meta-analysis of studies that examined the effect of app-based therapy, including cognitive behavioral therapy, acceptance and commitment therapy, and mindfulness, on depressive symptoms was performed. A total of 5597 articles were identified via five databases. After screening, 38 studies (n=8110 participants) remained for data extraction. From these studies, 50 total comparisons between postintervention mental health app intervention groups and control groups were included in the meta-analysis. RESULTS: A random effects model was performed to examine the effect of mental health apps on depressive symptoms compared to controls. The number of gamification elements within the apps was included as a moderator. Results indicated a small to moderate effect size across all mental health apps in which the mental health app intervention effectively reduced depressive symptoms compared to controls (Hedges g=-0.27, 95% CI -0.36 to -0.17; P<.001). The gamification moderator was not a significant predictor of depressive symptoms (β=-0.03, SE=0.03; P=.38), demonstrating no significant difference in effectiveness between mental health apps with and without gamification features. A separate meta-regression also did not show an effect of gamification elements on intervention adherence (β=-1.93, SE=2.28; P=.40). CONCLUSIONS: The results show that both mental health apps with and without gamification elements were effective in reducing depressive symptoms. There was no significant difference in the effectiveness of mental health apps with gamification elements on depressive symptoms or adherence. This research has important clinical implications for understanding how gamification elements influence the effectiveness of mental health apps on depressive symptoms.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.008 | 0.000 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.009 | 0.002 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".