Symptoms and signs of lung cancer prior to diagnosis: Comparative study using electronic health records
Thompson, M. J.; Prado, M. G.; Kessler, L. G.; Au, M. A.; Burkhardt, H. A.; Suchsland, M. Z.; Kowalski, L.; Stephens, K. A.; Yetisgen, M.; Walter, F. M.; Neal, R. D.; Lybarger, K.; Thompson, C. A.; Achkar, M. A.; Sarma, E. A.; Turner, G.; Farjah, F.
Show abstract
BackgroundLung cancer is the most common cause of cancer-related death in the United States (US), with most patients diagnosed at later stages (3 or 4). While most patients are diagnosed following symptomatic presentation, no studies have compared symptoms and physical examination signs at or prior to diagnosis from electronic health records (EHR) in the United States (US). ObjectiveTo identify symptoms and signs in patients prior to lung cancer diagnosis in EHR data. Study DesignCase-control study. MethodsWe studied 698 primary lung cancer cases in adults diagnosed between January 1, 2012 and December 31, 2019, and 6,841 controls matched by age, sex, smoking status, and type of clinic. Coded and free-text data from the EHR were extracted from 2 years prior to diagnosis date for cases and index date for controls. Univariate and multivariate conditional logistic regression were used to identify symptoms and signs associated with lung cancer. Analyses were repeated excluding symptom data from 1, 3, 6, and 12 months before the diagnosis/index dates. ResultsEleven symptoms and signs recorded during the study period were associated with a significantly higher chance of being a lung cancer case in multivariate analyses. Of these, seven were significantly associated with lung cancer six months prior to diagnosis: hemoptysis (OR 3.2, 95%CI 1.9-5.3), cough (OR 3.1, 95%CI 2.4-4.0), chest crackles or wheeze (OR 3.1, 95%CI 2.3-4.1), bone pain (OR 2.7, 95%CI 2.1-3.6), back pain (OR 2.5, 95%CI 1.9-3.2), weight loss (OR 2.1, 95%CI 1.5-2.8) and fatigue (OR 1.6, 95%CI 1.3-2.1). ConclusionsPatients diagnosed with lung cancer appear to have symptoms and signs recorded in the EHR that distinguish them from similar matched patients in ambulatory care, often six months or more before their diagnosis. These findings suggest opportunities to improve the diagnostic process for lung cancer in the US.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Post-COVID assessment in a specialist clinical service: a 12-month, single-centre analysis of symptoms and healthcare needs in 1325 individuals 93%
- Clinical characteristics and outcomes of adult patients admitted with COVID-19 in East London: a retrospective cohort analysis 92%
- Investigating a structured diagnostic approach for chronic breathlessness in primary care: a mixed-methods feasibility cluster Randomised Controlled Trial. 92%
Similar papers in this journal
Similar papers in this journal
- Disparities in outcomes among patients diagnosed with cancer associated with emergency department visits 94%
- Association of Chronic Acid Suppression and Social Determinants of Health with COVID-19 Infection 93%
- Pan-cancer analyses of the associations between 109 pre-existing conditions and cancer treatment patterns across 19 adult cancers 93%
Similar papers in this journal
- Remote Covid Assessment in Primary Care (RECAP) risk prediction tool: derivation and real-world validation studies 91%
- COVID-19 collateral: Indirect acute effects of the pandemic on physical and mental health in the UK 90%
- An external validation of the QCovid risk prediction algorithm for risk of mortality from COVID-19 in adults: national validation cohort study in England 89%
Similar papers in this journal
- Development and validation of a parent proxy bronchiectasis child quality of life instrument: The BC-QoL 91%
- Symptoms persisting after hospitalization for COVID-19: 12 months interim results of the COFLOW study 91%
- COVID-PCD – a participatory research study on the impact of COVID-19 in people with Primary Ciliary Dyskinesia 91%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.