Multimodal Freezing of Gait Detection: Analyzing the Benefits and Limitations of Physiological Data
Bibliographic record
Abstract
Freezing of gait (FOG) is a debilitating symptom of Parkinson's disease (PD), characterized by an absence or reduction in forward movement of the legs despite the intention to walk. Detecting FOG during free-living conditions presents significant challenges, particularly when using only inertial measurement unit (IMU) data, as it must be distinguished from voluntary stopping events that also feature reduced forward movement. Influences from stress and anxiety, measurable through galvanic skin response (GSR) and electrocardiogram (ECG), may assist in distinguishing FOG from normal gait and stopping. However, no study has investigated the fusion of IMU, GSR, and ECG for FOG detection. Therefore, this study introduced two methods: a two-step approach that first identified reduced forward movement segments using a Transformer-based model with IMU data, followed by an XGBoost model classifying these segments as FOG or stopping using IMU, GSR, and ECG features; and an end-to-end approach employing a multi-stage temporal convolutional network to directly classify FOG and stopping segments from IMU, GSR, and ECG data. Results showed that the two-step approach with all data modalities achieved an average F1 score of 0.728 and F1@50 of 0.725, while the end-to-end approach scored 0.771 and 0.759, respectively. However, no significant difference was found compared to using only IMU data in both approaches (p-values: 0.466 to 0.887). In conclusion, adding physiological data did not provide a statistically significant benefit in distinguishing between FOG and stopping. The limitations may be specific to GSR and ECG data, and may not generalize to other physiological modalities.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".