Accelerating Insight Discovery in Large Biomedical Text with Scalable Processing Framework
Kim, D.; Hauptman, M.; Patrick, M. T.
Show abstract
Large language models are increasingly being used by dermatology professionals to support diagnostic investigation, patient education, and medical research. While these models can help manage information overload and improve efficiency, concerns persist regarding their accuracy and potential reliance on dubious sources. We introduce Quanta, a hybrid system that combines large language models with established evaluation metrics, such as cosine similarity, to enable efficient summarization and interpretation of curated research corpora. This methodology ensures that synthesized insights remain domain-specific and contextually relevant, thereby supporting clinicians and researchers in navigating the expanding digital landscape of dermatology literature. Deployed within an interactive chatbot, the tool delivers direct answers to user queries, provides cross-publication insights, and can suggest new directions for research. Comparative evaluations on benchmark datasets demonstrate improvements in accuracy, efficiency, and computational cost, with the curated document approach enhancing reliability and reducing misinformation risk.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- Building a Best-in-Class De-identification Tool for Electronic Medical Records Through Ensemble Learning 94%
- KG-COVID-19: a framework to produce customized knowledge graphs for COVID-19 response 93%
- Inferring global-scale temporal latent topics from news reports to predict public health interventions for COVID-19 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.