A Light-weight Text Summarizer for Fast Access to Medical Evidence
Sarker, A.; Yang, Y.-C.; Al-Garadi, M. A.
Show abstract
The performances of current medical text summarization systems rely on resource-heavy domain-specific knowledge sources, and preprocessing methods (e.g., classification or deep learning) for deriving semantic information. Consequently, these systems are often difficult to customize, extend or deploy in low-resource settings, and are operationally slow. We propose a fast summarization system that can aid practitioners at point-of-care, and, thus, improve evidence-based healthcare. At runtime, our system utilizes similarity measurements derived from pre-trained domain-specific word embeddings in addition to simple features, rather than clunky knowledge bases and resource-heavy preprocessing. Automatic evaluation on a public dataset for evidence-based medicine shows that our systems performance, despite the simple implementation, is statistically comparable with the state-of-the-art.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Large language models improve transferability of electronic health record-based predictions across countries and coding systems 95%
- Interpretable Fine-tuned Large Language Models Facilitate Making Genetic Test Decisions for Rare Diseases 94%
- A Framework to Assess Clinical Safety and Hallucination Rates of LLMs for Medical Text Summarisation 94%
Similar papers in this journal
- Analysis of Eligibility Criteria Clusters Based on Large Language Models for Clinical Trial Design 94%
- Trialstreamer: a living, automatically updated database of clinical trial reports 93%
- Annotation-preserving machine translation of English corpora to validate Dutch clinical concept extraction tools 92%
Similar papers in this journal
- Extracting social determinants of health from electronic health records: development and comparison of rule-based and large language models-based methods 93%
- Federated Learning of Electronic Health Records Improves Mortality Prediction in Patients Hospitalized with COVID-19 91%
- An Intrinsic and Extrinsic Evaluation of Learned COVID-19 Concepts using Open-Source Word Embedding Sources 91%
Similar papers in this journal
- Temporal Relationship of Computed and Structured Diagnoses in Electronic Health Record Data 92%
- ARDSFlag: An NLP/Machine Learning Algorithm to Visualize and Detect High-Probability ARDS Admissions Independent of Provider Recognition and Billing Codes 91%
- MelAnalyze: Fact-Checking Melatonin claims using Large Language Models and Natural Language Inference 91%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.