An EEG-based machine learning framework for diagnosing acute sleep deprivation
Bibliographic record
Abstract
Study objective: Acute sleep deprivation significantly impacts cognitive function, contributes to accidents, and increases the risk of chronic illnesses, underscoring the need for reliable and objective diagnosis. Our work aims to develop a machine learning-based approach to discriminate between EEG recordings from acutely sleep-deprived individuals and those that are well-rested, facilitating the objective detection of acute sleep deprivation and enabling timely intervention to mitigate its adverse effects. Methods: Sixty-one-channel eyes-open resting-state electroencephalography (EEG) data from a publicly available dataset of 71 participants were analyzed. Following preprocessing, EEG recordings were segmented into contiguous, non-overlapping 20-second epochs. For each epoch, a comprehensive set of features was extracted, including statistical descriptors, spectral measures, functional connectivity indices, and graph-theoretic metrics. Four machine learning classifiers - Light Gradient-Boosting Machine (LightGBM), eXtreme Gradient Boosting (XGBoost), Random Forest (RF), and Support Vector Classifier (SVC) - were trained on these features using nested stratified cross-validation to ensure unbiased performance evaluation. In parallel, three deep learning models-a Convolutional Neural Network (CNN), Long Short-Term Memory network (LSTM), and Transformer-were trained directly on the raw multi-channel EEG time-series data. All models were evaluated under two conditions: (i) without subject-level separation, allowing the same participant to contribute to both training and test sets, and (ii) with subject-level separation, where models were tested exclusively on unseen participants. Model performance was assessed using accuracy, F1-score, and area under the receiver operating characteristic curve (AUC). Results: Without subject-level separation, CNN achieved the highest accuracy (95.72%), followed by XGBoost (95.42%), LightGBM (94.83%), RF (94.53%), and SVC (85.25%), with the Transformer (77.39%) and LSTM (66.75%) models achieving lower accuracies. Under subject-level separation, RF achieved the highest accuracy (68.23%), followed by XGBoost (66.36%), LightGBM (66.21%), CNN (65.35%), and SVC (65.08%), while the Transformer (63.35%) and LSTM (61.70%) models achieved the lowest accuracies. Conclusion: This study demonstrates the potential of EEG-based machine learning for detecting acute sleep deprivation, while underscoring the challenges of achieving robust subject-level generalization. Despite reduced accuracy under cross-subject evaluation, these findings support the feasibility of developing scalable, non-invasive tools for sleep deprivation detection using EEG and advanced ML techniques.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.001 | 0.001 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".