Spine Reviews: Crowdsourcing Global Spine Expert Knowledge via Digital Ledger Technology
Challier, V.; Diebo, B.; Lafage, V.; Dehouche, N.; Lonjon, G.; Cristini, J.; SpineDAO,
Show abstract
Study DesignProspective observational study using a novel digital ledger technology (DLT)-based crowdsourcing platform. ObjectiveTo develop and evaluate "Spine Reviews," a blockchain-based platform for aggregating spine treatment recommendations from an international specialist panel, and to validate the clinical coherence of the resulting dataset. Summary of Background DataPredictive models for low back pain treatment are limited by small, homogeneous datasets that fail to capture inter-clinician variability. Traditional multi-center data collection is expensive, slow, and geographically constrained. DLT-based crowdsourcing with cryptographic credentialing may overcome these barriers. MethodsFive hundred synthetic patient vignettes (digital twins) were generated; 463 retained after quality control. A review platform was built on the Solana blockchain using non-transferable Soulbound Tokens (SBTs) for credentialing and smart-contract compensation. Fifty-two specialists from 7 countries provided [≥]4 reviews per vignette across four treatment tiers, without access to imaging or physical examination. Mixed-effects regression with reviewer random intercepts partitioned decision variability. ResultsThe platform collected 2,066 completed reviews (97.7%) over 37 days at $0.97/review. Variance decomposition revealed that 36.7% of treatment tier variability was attributable to patient presentation, 19.2% to reviewer practice style, and 44.1% to their interaction. Neurological deficits ({beta}=0.39), symptom duration ({beta}=0.12), and pain ({beta}=0.09) independently predicted treatment escalation (all p<0.001). Gwets AC1 was almost perfect for emergency (0.92) and substantial for conservative decisions (0.67). Reviewer confidence in treatment recommendations decreased with escalating tier severity (conservative 4.59/5 vs surgical 4.05/5), suggesting appropriate uncertainty calibration. ConclusionsDLT with SBT credentialing enables rapid, global, cost-effective aggregation of clinically coherent expert judgment. The three-component variance structure quantifies clinical equipoise in spine care and establishes that predictive models require diverse, multi-reviewer training data.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Bridging the Literacy Gap for Surgical Consents: An AI-Human Expert Collaborative Approach 93%
- The clinician-AI interface: intended use and explainability in FDA-cleared AI devices for medical image interpretation 93%
- International Electronic Health Record-Derived COVID-19 Clinical Course Profiles: The 4CE Consortium 92%
Similar papers in this journal
- Diversity and inclusion: A hidden additional benefit of Open Data 93%
- Predictability and Stability Testing to Assess Clinical Decision Instrument Performance for Children After Blunt Torso Trauma 93%
- Artificial Intelligence's Contribution to Biomedical Literature Search: Revolutionizing or Complicating? 92%
Similar papers in this journal
- Clinical And Cost-Effectiveness Of A Personalised Guided Consultation Versus Usual Physiotherapy Care In People Presenting With Shoulder Pain: A Protocol For The Panda-S Cluster Randomised Controlled Trial And Process Evaluation 93%
- Doctors of chiropractic working with or within integrated health care delivery systems: a scoping review protocol 93%
- Financial Conflicts of Interest among U.S. Physician Authors of 2020 Clinical Practice Guidelines: A Cross-Sectional Study 93%
Similar papers in this journal
- Cracking the Code: A Scoping Review to Unite Disciplines in Tackling Legal Issues in Health Artificial Intelligence 91%
- The performance of national COVID-19 ‘Symptom Checkers’: A comparative case simulation study 91%
- Impact of the Federated Data Platform's digital surgery scheduling system on elective theatre utilisation at an NHS Trust: an interrupted time series analysis 90%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.