Did you miss me? Making the most of digital phenotyping data by imputing missingness with point process models
Leaning, I. E.; Costanzo, A.; Jagesar, R.; Knol, L.; Tjeerdsma, S.; Tyborowska, A.; Ikani, N.; Reus, L. M.; Visser, P. J.; Kas, M. J. H.; Beckmann, C. F.; Ruhe, H. G.; Marquand, A. F.
Show abstract
AO_SCPLOWBSTRACTC_SCPLOWO_ST_ABSObjectivesC_ST_ABSDigital phenotyping has broad clinical potential, providing low-burden objective measures of behaviour as individuals go about their lives. Digital phenotyping goals include the prediction of relapse in mental illness and improved symptom monitoring. However, progress in making clinical inferences from these data is severely challenged by the common occurrence of missing data. We propose a novel method to address this by using non-homogeneous Poisson point process models (PPPMs) to impute missing digital phenotyping data, accounting for their timeseries nature, where smartphone-based activities are modelled as points. MethodsWe demonstrate the use of PPPMs for imputing timeseries data and evaluate their influence on downstream analysis. We evaluate the inclusion of time-varying covariates to model diurnal variations in personalised PPPMs. We validate this model using participants from SMARD (n=26) in a ground truth evaluation, then in PRISM (n=65) and Hersenonderzoek studies (n=283), performing a replication analysis involving hidden Markov models (HMMs). ResultsIn the ground truth evaluation, PPPMs including hour of the day as a covariate, encoded using one-hot encoding, provided the best fit (highest out-of-sample likelihood). Using this imputation method, HMM properties such as daily rhythms were preserved and we successfully replicated findings from our prior work. DiscussionPersonalised PPPMs using covariates provide tailored simulations of behaviour that can be used for imputation in behavioural time series. ConclusionPPPMs using covariates are a promising imputation tool that may contribute to improved utility of digital phenotyping.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Uncovering social states in healthy and clinical populations using digital phenotyping and Hidden Markov Models 96%
- High-Resolution Digital Phenotypes from Consumer Wearables Enhance Prediction of Cardiometabolic Risk Markers 92%
- Tracking private WhatsApp discourse about COVID-19: A longitudinal infodemiology study in Singapore 92%
Similar papers in this journal
- Listening to mental health crisis needs at scale: using Natural Language Processing to understand and evaluate a mental health crisis text messaging service 94%
- Remote digital measurement of visual and auditory markers of Major Depressive Disorder severity and treatment response. 91%
- CharMark: A Markov Approach to Linguistic Biomarkers in Dementia 90%
Similar papers in this journal
- Estimating the trend of COVID-19 in Norway by combining multiple surveillance indicators 93%
- Discrimination of sleep and wake periods from a hip-worn raw acceleration sensor using recurrent neural networks 92%
- Compressive Big Data Analytics: An Ensemble Meta-Algorithm for High-dimensional Multisource Datasets 92%
Similar papers in this journal
- Development and validation of risk scores for all-cause mortality for the purposes of a smartphone-based ‘general health score’ application: a prospective cohort study using the UK Biobank 93%
- Comparison of accelerometry-based measures of physical activity 92%
- A Neural Network Based Algorithm for Dynamically Adjusting Activity Targets to Sustain Exercise Engagement Among People Using Activity Trackers 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.