Temporally Continuous Automated Sleep-Wake Classification Using Deep Learning
Somaskandhan, P.; Korkalainen, H.; Leppänen, T.; Töyräs, J.; Melehan, K.; Ruehland, W.; Sands, S. A.; Mann, D. L.; Wilson, D. L.; Terrill, P. I.
Show abstract
IntroductionSegmenting sleep into fixed 30-second epochs remains central to current sleep scoring practice, yet it imposes rigid boundaries that may not accurately reflect the true temporal sleep dynamics. We aimed to develop a deep learning-based, high-temporal-resolution sleep-wake classifier leveraging temporally continuous manual reference scoring without fixed epoch boundaries and transfer learning techniques to facilitate progress toward a more physiologically consistent sleep assessment. MethodsThree independent datasets were utilized, of which two included sleep-wake scoring manually conducted in a temporally continuous manner. A U-Net based model was initially trained on a large dataset scored using 30-second epochs, with post hoc scoring modifications (n=2034). It was then fine-tuned via transfer learning using a subset of one of the datasets with temporally continuous scoring (n=39) and validated on both its holdout portion (n=40) and the other independent temporally continuous scoring dataset (n=20). Wakefulness and arousals were consolidated, acknowledging their shared physiological characteristics. Prediction confidence estimates were also generated. ResultsThe model achieved overall concordance of 88.96% ({kappa}=0.78) and 88.23% ({kappa}=0.76) in the holdout and second independent evaluation dataset, respectively, with temporally continuous scoring. Correlation between 1-second automatic predictions and temporally continuous manual scoring was r=0.93 (p<0.001) for total sleep time and r=0.67 (p<0.001) for sleep-to-wake transition index. ConclusionsThese findings support the utility of our model in addressing key limitations of 30-second epoch-based scoring and progressing toward more physiologically consistent sleep-wake assessment by providing a practical basis for subsequent analyses. Misclassifications generally showed lower confidences, indicating additional value for targeted review. Statement of SignificanceConventional sleep scoring remains constrained by fixed 30-second epochs, which may fail to capture the true temporal dynamics of the underlying changes between sleep and wakefulness. In this study, we used polysomnography data manually scored on a temporally continuous basis as the gold standard to develop and validate a deep learning model capable of classifying sleep and wakefulness-like states (consolidating wakefulness and arousal) at high temporal resolution without fixed 30-second epochs. The model demonstrated strong agreement with the gold standard, and as such, lays a practical foundation for deriving improved physiologically meaningful biomarkers of sleep fragmentation and continuity, with potential diagnostic and prognostic value and broad applicability toward a more precise and physiologically consistent sleep assessment.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- A foundational transformer leveraging full night, multichannel sleep study data accurately classifies sleep stages 98%
- Evaluation of Dreem headband for sleep staging and EEG spectral analysis in people living with Alzheimer’s and older adults 96%
- Corticothalamic modelling of sleep neurophysiology with applications to mobile EEG 95%
Similar papers in this journal
- Novel Digital Markers of Sleep Dynamics: A Causal Inference Approach Revealing Age and Gender Phenotypes in Obstructive Sleep Apnea 96%
- Topographical relocation of adolescent sleep spindles reveals a new maturational pattern of the human brain 95%
- Decreased electrocortical temporal complexity distinguishes sleep from wakefulness 94%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.