Digital Phenotyping of Geriatric Depression Using a Community-Based Digital Mental Health Monitoring Platform for Socially Vulnerable Older Adults and Their Community Caregivers: 6-Week Living Lab Single-Arm Pilot Study
Bibliographic record
Abstract
BACKGROUND: Despite the increasing need for digital services to support geriatric mental health, the development and implementation of digital mental health care systems for older adults have been hindered by a lack of studies involving socially vulnerable older adult users and their caregivers in natural living environments. OBJECTIVE: This study aims to determine whether digital sensing data on heart rate variability, sleep quality, and physical activity can predict same-day or next-day depressive symptoms among socially vulnerable older adults in their everyday living environments. In addition, this study tested the feasibility of a digital mental health monitoring platform designed to inform older adult users and their community caregivers about day-to-day changes in the health status of older adults. METHODS: A single-arm, nonrandomized living lab pilot study was conducted with socially vulnerable older adults (n=25), their community caregivers (n=16), and a managerial social worker over a 6-week period during and after the COVID-19 pandemic. Depressive symptoms were assessed daily using the 9-item Patient Health Questionnaire via scripted verbal conversations with a mobile chatbot. Digital biomarkers for depression, including heart rate variability, sleep, and physical activity, were measured using a wearable sensor (Fitbit Sense) that was worn continuously, except during charging times. Daily individualized feedback, using traffic signal signs, on the health status of older adult users regarding stress, sleep, physical activity, and health emergency status was displayed on a mobile app for the users and on a web application for their community caregivers. Multilevel modeling was used to examine whether the digital biomarkers predicted same-day or next-day depressive symptoms. Study staff conducted pre- and postsurveys in person at the homes of older adult users to monitor changes in depressive symptoms, sleep quality, and system usability. RESULTS: Among the 31 older adult participants, 25 provided data for the living lab and 24 provided data for the pre-post test analysis. The multilevel modeling results showed that increases in daily sleep fragmentation (P=.003) and sleep efficiency (P=.001) compared with one's average were associated with an increased risk of daily depressive symptoms in older adults. The pre-post test results indicated improvements in depressive symptoms (P=.048) and sleep quality (P=.02), but not in the system usability (P=.18). CONCLUSIONS: The findings suggest that wearable sensors assessing sleep quality may be utilized to predict daily fluctuations in depressive symptoms among socially vulnerable older adults. The results also imply that receiving individualized health feedback and sharing it with community caregivers may help improve the mental health of older adults. However, additional in-person training may be necessary to enhance usability. TRIAL REGISTRATION: ClinicalTrials.gov NCT06270121; https://clinicaltrials.gov/study/NCT06270121.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.003 | 0.002 |
| Meta-epidemiology (narrow) | 0.001 | 0.001 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.001 | 0.000 |
| Science and technology studies | 0.002 | 0.001 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.001 | 0.001 |
| Research integrity | 0.001 | 0.001 |
| Insufficient payload (model declined to judge) | 0.002 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".