Exploring the Use of Multiple Mental Health Apps Within a Platform: Secondary Analysis of the IntelliCare Field Trial
Bibliographic record
Abstract
BACKGROUND: IntelliCare is a mental health app platform with 14 apps that are elemental, simple and brief to use, and eclectic. Although a variety of apps may improve engagement, leading to better outcomes, they may require navigation aids such as recommender systems that can quickly direct a person to a useful app. OBJECTIVE: As the first step toward developing navigation and recommender tools, this study explored app-use patterns across the IntelliCare platform and their relationship with depression and anxiety outcomes. METHODS: This is a secondary analysis of the IntelliCare Field Trial, which recruited people with depression or anxiety. Participants of the trial received 8 weeks of coaching, primarily by text, and weekly random recommendations for apps. App-use metrics included frequency and lifetime use. Depression and anxiety, measured using the Patient Health Questionnaire-9 and Generalized Anxiety Disorder-7, respectively, were assessed at baseline and end of treatment. Cluster analysis was utilized to determine patterns of app use; ordinal logistic regression models and log-rank tests were used to determine if these use metrics alone, or in combination, predicted improvement or remission in depression or anxiety. RESULTS: The analysis included 96 people who generally followed recommendations to download and try new apps each week. Apps were clustered into 5 groups: Thinking (apps that targeted or relied on thinking), Calming (relaxation and insomnia), Checklists (apps that used checklists), Activity (behavioral activation and activity), and Other. Both overall frequency of use and lifetime use predicted response for depression and anxiety. The Thinking, Calming, and Checklist clusters were associated with improvement in depression and anxiety, and the Activity cluster was associated with improvement in Anxiety only. However, the use of clusters was less strongly associated with improvement than individual app use. CONCLUSIONS: Participants in the field trial remained engaged with a suite of apps for the full 8 weeks of the trial. App-use patterns did fall into clusters, suggesting that some knowledge about the use of one app may be useful in selecting another app that the person is more likely to use and may help suggest apps based on baseline symptomology and personal preference.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".