Exploring the role of Large Language Models in Melanoma: a Systemic Review
Zarfati, M.; Nadkarni, G.; Glicksberg, B. S.; Harats, M.; Greenberger, S.; Klang, E.; Soffer, S.
Show abstract
BackgroundLarge language models (LLMs) are gaining recognition across various medical fields; however, their specific role in dermatology, particularly in melanoma care, is not well- defined. This systematic review evaluates the current applications, advantages, and challenges associated with the use of LLMs in melanoma care. MethodsWe conducted a systematic search of PubMed and Scopus databases for studies published up to July 23, 2024, focusing on the application of LLMs in melanoma. Identified studies were categorized into three subgroups: patient education, diagnosis and clinical management. The review process adhered to PRISMA guidelines, and the risk of bias was assessed using the modified QUADAS-2 tool. ResultsNine studies met the inclusion criteria. Five studies compared various LLM models, while four focused on ChatGPT. Three studies specifically examined multi-modal LLMs. In the realm of patient education, ChatGPT demonstrated high accuracy, though it often surpassed the recommended readability levels for patient comprehension. In diagnosis applications, multi- modal LLMs like GPT-4V showed capabilities in distinguishing melanoma from benign lesions. However, the diagnostic accuracy varied considerably, influenced by factors such as the quality and diversity of training data, image resolution, and the models ability to integrate clinical context. Regarding management advice, one study found that ChatGPT provided more reliable management advice compared to other LLMs, yet all models lacked depth and specificity for individualized decision-making. ConclusionsLLMs, particularly multimodal models, show potential in improving melanoma care through patient education, diagnosis, and management advice. However, current LLM applications require further refinement and validation to confirm their clinical utility. Future studies should explore fine-tuning these models on large dermatological databases and incorporate expert knowledge.
Matching journals
The top 11 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Deep neural frameworks improve the accuracy of general practitioners in the classification of pigmented skin lesions 92%
- The NILS study protocol - a retrospective validation study of a preoperative decision-making tool for non-invasive lymph node staging in women with primary breast cancer [ISRCTN14341750] 90%
- A Machine Learning Ensemble Based on Radiomics to Predict BI-RADS Category and Reduce the Biopsy Rate of Ultrasound-Detected Suspicious Breast Masses 88%
Similar papers in this journal
- Pilot Feasibility Clinical Trial of Virtual Reality for Pain Management During Repeated Pediatric Laser Procedures: Study Protocol for a Randomized Clinical Trial 91%
- Scientific hypothesis generation process in clinical research: a secondary data analytic tool versus experience study protocol 90%
- Strategies and Tools for electronic health records and physician workflow alignment: A scoping review protocol 87%
Similar papers in this journal
- Computer vision detects inflammatory arthritis in standardized smartphone photographs in an Indian patient cohort 89%
- Risk Factors, Clinical Outcomes and Prognostic Factors of Bacterial Keratitis: The Nottingham Infectious Keratitis Study 89%
- Biologic Drug Survival in Psoriasis: A Systematic Review & Comparative Meta-Analysis 89%
Similar papers in this journal
- Reconstructed human pigmented skin/epidermis models achieve epidermal pigmentation through melanocore transfer. 87%
- Diversity, Equity, and Inclusion in the Melanoma Research Community 87%
- Melanocortin-1 receptor (MC1R) genotypes do not correlate with size in two cohorts of medium-to-giant congenital melanocytic nevi 86%
Similar papers in this journal
- Use of Skin Cancer Procedures, Medicare Reimbursement, and Overall Expenditures, 2012-2017 91%
- Characterizing Potential Conflicts of Interest Among UpToDate and DynaMed Content Contributors 89%
- Evaluation of an artificial intelligence model for detection of pneumothorax and tension pneumothorax on chest radiograph 88%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.