Informed Injury Prediction in Elite Football: Decision Theory meets Machine Learning
Huth, M.; Canal-Simon, B.; Ferrer, E.; Rodas, G.; Yanguas, X.; Hasenauer, J.; Gonzalez, J. R.
Show abstract
Injuries in elite sports disrupt team performance, shorten careers, and incur significant financial costs, highlighting the critical need for accurate predictions to inform optimal decisions that effectively prevent injuries. Existing approaches to injury prediction fail to account for cumulative risk, overlook injury severity, lack reliable probability calibration, and omit statistically guided decision thresholds. Here, we present a novel injury prediction framework integrating risk accumulation via survival analysis with machine learning, probability beta calibration, and statistical decision theory. Using a unique dataset spanning four seasons from FC Barcelonas womens team, we demonstrate that our framework outperforms standard classifiers, yielding superior discrimination ability. Our framework identifies fatigue-related measures as key injury predictors and incorporates flexible thresholds based on match importance and decision-maker certainty, improving player availability. Scalable and transferable to other sports, this framework bridges academic research and practical deployment, empowering sports organizations to optimize player performance and long-term outcomes.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Machine Learning for Real-Time Aggregated Prediction of Hospital Admission for Emergency Patients 92%
- Continuous-Time and Dynamic Suicide Attempt Risk Prediction with Neural Ordinary Differential Equations 92%
- Modeling trajectories of routine blood tests as dynamic biomarkers for outcome in spinal cord injury 91%
Similar papers in this journal
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.