Elder-Sim: A Psychometrically Validated Platform for Personality-Stable Elderly Digital Twins
Wang, J.; Yang, Z.; Zhu, Z.; Zhu, X.; Huang, Z.; Wang, H.; Tian, L.; Cao, Y.; Qu, X.; Qi, X.; Wu, B.
Show abstract
Background: LLMs enable patient-facing conversational agents, creating a pathway toward digital twins that capture older adults' lived experiences and behavioral responses across time. A central barrier is personality drift---inconsistent trait expression across repeated interactions---which undermines reliability of generated trajectories and intervention-response simulation in geriatric care. Objective: To develop ELDER-SIM, a multi-role elderly-care conversational platform for building personality-stable digital twin agents, and to propose a psychometric validation framework for quantifying personality consistency in LLM-based agents. Methods: ELDER-SIM was implemented via n8n workflow orchestration with local LLM inference (Ollama/vLLM), integrating (1) Big Five (OCEAN) trait specifications, (2) a Cognitive Conceptualization Diagram (CCD) grounded in Beck's CBT framework, and (3) a MySQL-based long-term memory module. Ablation studies across four conditions---Baseline, +Memory, +CCD, and +LoRA (fine-tuned on 19,717 instruction pairs from CHARLS)---were evaluated via Cronbach's $\alpha$, ICC, and role discrimination accuracy. Results: Personality measurement reliability was acceptable to excellent across conditions (Cronbach's : 0.70-0.94), with consistently high test-retest stability (ICC: 0.85- 2 0.96). Role discrimination improved stepwise from 83.3% (Baseline) to 88.9% (+Memory), 94.4% (+CCD), and 97.2% (+LoRA). CCD produced the largest gain in internal consistency (mean 0.702[->]0.892), while LoRA achieved the highest overall internal consistency ( 0.940) and ICC (0.958). Conclusions: ELDER-SIM provides a psychometrically validated approach for constructing personality-consistent elderly digital twin agents. Structured cognitive modeling and domain adaptation reduce personality drift, supporting reliable longitudinal simulation for elderly mental health care and reproducible in silico evaluation before clinical deployment.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Validating a Clinical Decision Support System for Palliative Care using healthcare professionals’ insights 93%
- User Experience Evaluation of Cogscreen for Screening Mild Cognitive Impairment: Formative and Summative Evaluation 93%
- A digital self-care intervention for Ugandan patients with heart failure and their clinicians: User-centred design and usability study 93%
Similar papers in this journal
- Development of the AD F ICE_IT clinical decision support system to assist deprescribing of fall-risk increasing drugs: A user-centered design approach 94%
- Evaluation of a novel community-based COVID-19 'Test-to-Care model' for low-income populations 91%
- Feasibility and reliability of online vs in-person cognitive testing in healthy older people 91%
Similar papers in this journal
- Developing and validating an explainable digital mortality prediction tool for extremely preterm infants 92%
- Feasibility characteristics of wrist-worn fitness trackers in health status monitoring for post-COVID patients in remote and rural areas 92%
- Collaborative intelligence in AI: Evaluating the performance of a council of AIs on the USMLE 91%
Similar papers in this journal
- Caregivers’ burden of care during emergency department care transitions among older adults: a mixed methods cohort study 91%
- Protocol for a randomised, double-blind trial of a chronotherapeutic mHealth behaviour change intervention to optimise light exposure among older adults aged >=60 years in Singapore (LightSPAN) 91%
- Patterns of Physical Activity Over Time in Older Patients Rehabilitating after Hip Fracture Surgery: A Preliminary Study 90%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.