Electronic Health Diary Campaigns to Complement Longitudinal Assessments in Persons With Multiple Sclerosis: Nested Observational Study
Bibliographic record
Abstract
BACKGROUND: Electronic health diaries hold promise in complementing standardized surveys in prospective health studies but are fraught with numerous methodological challenges. OBJECTIVE: The study aimed to investigate participant characteristics and other factors associated with response to an electronic health diary campaign in persons with multiple sclerosis, identify recurrent topics in free-text diary entries, and assess the added value of structured diary entries with regard to current symptoms and medication intake when compared with survey-collected information. METHODS: Data were collected by the Swiss Multiple Sclerosis Registry during a nested electronic health diary campaign and during a regular semiannual Swiss Multiple Sclerosis Registry follow-up survey serving as comparator. The characteristics of campaign participants were descriptively compared with those of nonparticipants. Diary content was analyzed using the Linguistic Inquiry and Word Count 2015 software (Pennebaker Conglomerates, Inc) and descriptive keyword analyses. The similarities between structured diary data and follow-up survey data on health-related quality of life, symptoms, and medication intake were examined using the Jaccard index. RESULTS: Campaign participants (n=134; diary entries: n=815) were more often women, were not working full time, did not have a higher education degree, had a more advanced gait impairment, and were on average 5 years older (median age 52.5, IQR 43.25-59.75 years) than eligible nonparticipants (median age 47, IQR 38-55 years; n=524). Diary free-text entries (n=632; participants: n=100) most often contained references to the following standard Linguistic Inquiry and Word Count word categories: negative emotion (193/632, 30.5%), body parts or body functioning (191/632, 30.2%), health (94/632, 14.9%), or work (67/632, 10.6%). Analogously, the most frequently mentioned keywords (diary entries: n=526; participants: n=93) were "good," "day," and "work." Similarities between diary data and follow-up survey data, collected 14 months apart (median), were high for health-related quality of life and stable for slow-changing symptoms such as fatigue or gait disorder. Similarities were also comparatively high for drugs requiring a regular application, including interferon beta-1a (Avonex) and glatiramer acetate (Copaxone), and for modern oral therapies such as fingolimod (Gilenya) and teriflunomide (Aubagio). CONCLUSIONS: Diary campaign participation seemed dependent on time availability and symptom burden and was enhanced by reminder emails. Electronic health diaries are a meaningful complement to regular structured surveys and can provide more detailed information regarding medication use and symptoms. However, they should ideally be embedded into promotional activities or tied to concrete research study tasks to enhance regular and long-term participation.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.008 | 0.020 |
| Meta-epidemiology (narrow) | 0.000 | 0.001 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.001 | 0.002 |
| Science and technology studies | 0.001 | 0.000 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.001 | 0.001 |
| Research integrity | 0.001 | 0.001 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".