Behavior Change Content, Understandability, and Actionability of Chronic Condition Self-Management Apps Available in France: Systematic Search and Evaluation
Bibliographic record
Abstract
BACKGROUND: The quality of life of people living with chronic conditions is highly dependent on self-management behaviors. Mobile health (mHealth) apps could facilitate self-management and thus help improve population health. To achieve their potential, apps need to target specific behaviors with appropriate techniques that support change and do so in a way that allows users to understand and act upon the content with which they interact. OBJECTIVE: Our objective was to identify apps targeted toward the self-management of chronic conditions and that are available in France. We aimed to examine what target behaviors and behavior change techniques (BCTs) they include, their level of understandability and actionability, and the associations between these characteristics. METHODS: We extracted data from the Google Play store on apps labelled as Top in the Medicine category. We also extracted data on apps that were found through 12 popular terms (ie, keywords) for the four most common chronic condition groups-cardiovascular diseases, cancers, respiratory diseases, and diabetes-along with apps identified through a literature search. We selected and downloaded native Android apps available in French for the self-management of any chronic condition in one of the four groups and extracted background characteristics (eg, stars and number of ratings), coded the presence of target behaviors and BCTs using the BCT taxonomy, and coded the understandability and actionability of apps using the Patient Education Material Assessment Tool for audiovisual materials (PEMAT-A/V). We performed descriptive statistics and bivariate statistical tests. RESULTS: A total of 44 distinct native apps were available for download in France and in French: 39 (89%) were found via the Google Play store and 5 (11%) were found via literature search. A total of 19 (43%) apps were for diabetes, 10 for cardiovascular diseases (23%), 8 for more than one condition in the four groups (18%), 6 for respiratory diseases (14%), and 1 for cancer (2%). The median number of target behaviors per app was 2 (range 0-7) and of BCTs per app was 3 (range 0-12). The most common BCT was self-monitoring of outcome(s) of behavior (31 apps), while the most common target behavior was tracking symptoms (30 apps). The median level of understandability was 42% and of actionability was 0%. Apps with more target behaviors and more BCTs were also more understandable (ρ=.31, P=.04 and ρ=.35, P=.02, respectively), but were not significantly more actionable (ρ=.24, P=.12 and ρ=.29, P=.054, respectively). CONCLUSIONS: These apps target few behaviors and include few BCTs, limiting their potential for behavior change. While content is moderately understandable, clear instructions on when and how to act are uncommon. Developers need to work closely with health professionals, users, and behavior change experts to improve content and format so apps can better support patients in coping with chronic conditions. Developers may use these criteria for assessing content and format to guide app development and evaluation of app performance. TRIAL REGISTRATION: PROSPERO CRD42018094012; https://www.crd.york.ac.uk/prospero/display_record.php?RecordID=94012.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.024 | 0.075 |
| Meta-epidemiology (narrow) | 0.001 | 0.001 |
| Meta-epidemiology (broad) | 0.006 | 0.005 |
| Bibliometrics | 0.028 | 0.015 |
| Science and technology studies | 0.001 | 0.001 |
| Scholarly communication | 0.003 | 0.003 |
| Open science | 0.001 | 0.002 |
| Research integrity | 0.001 | 0.001 |
| Insufficient payload (model declined to judge) | 0.004 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".