Algorithmic implementation of pancreatic cancer staging guidelines: comparison with a retrieval-augmented large language model
Komaba, A.; Amakawa, A.; Tozuka, R.; Sato, J.; Fujihara, K.; Emoto, M.; Sawada, S.; Kasai, S.; Sakamoto, K.; Shimura, K.; Johno, Y.; Nakamoto, K.; Ichikawa, S.; Johno, H.
Show abstract
Purpose: To implement a comprehensive knowledge-based algorithm (KBA) for pancreatic cancer staging based on the current Japanese guidelines and to evaluate its performance as a clinical decision support system in comparison with a retrieval-augmented large language model (LLM) system. Materials and methods: A KBA covering TNM classification, stage classification, and resectability classification was implemented as a web application. The correctness of the system outputs was exhaustively verified for all possible inputs. Subsequently, six non-board-certified radiologists performed pancreatic cancer staging for 12 simulated cases with imaging findings under three conditions: unassisted, LLM-assisted, and KBA-assisted. Staging accuracy and staging time were compared among the three conditions using pairwise proportion z-tests and Welch's t-tests, respectively. Results: In the comparative experiment, staging accuracy was 81.9%, 80.6%, and 98.6% in the unassisted, LLM-assisted, and KBA-assisted conditions, respectively. Mean staging time was 229.2, 401.9, and 196.2 s, respectively. The KBA-assisted condition showed higher accuracy than both the unassisted and LLM-assisted conditions (both p<0.001). Staging time was longer in the LLM-assisted condition than in the other two conditions (both p<0.001). Conclusion: A comprehensive KBA for pancreatic cancer staging based on the current Japanese guidelines was implemented and exhaustively verified. In a preliminary comparative experiment, KBA assistance improved staging accuracy without increasing staging time, whereas LLM assistance increased staging time without improving staging accuracy. These findings suggest that verified KBA systems may be feasible and useful for clinical tasks governed by explicit guideline-based rules.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Assessing GPT-4 Multimodal Performance in Radiological Image Analysis 95%
- Predicting EGFR mutation status in lung adenocarcinoma presenting as ground-glass opacity: utilizing radiomics model in clinical translation 94%
- Evaluating Large Language Model-Generated Brain MRI Protocols: Performance of GPT4o, o3-mini, DeepSeek-R1 and Qwen2.5-72B 92%
Similar papers in this journal
- Content-based image retrieval assists radiologists in diagnosing eye and orbital mass lesions in MRI 95%
- Segmentation of Pancreatic Ductal Adenocarcinoma (PDAC) and surrounding vessels in CT images using deep convolutional neural networks and Texture Descriptors 94%
- On evaluation metrics for medical applications of artificial intelligence 94%
Similar papers in this journal
- Demarcation line determination for diagnosis of gastric cancer disease range using unsupervised machine learning in magnifying narrow-band imaging 95%
- Auto-detection of motion artifacts on CT pulmonary angiograms with a physician-trained AI algorithm 93%
- Detection, Isolation and Quantification of Myocardial Infarct with Four Different Histological Staining Techniques 92%
Similar papers in this journal
- Quantification of abdominal fat from computed tomography using deep learning and its association with electronic health records in an academic biobank 92%
- Automated stratification of trauma injury severity across multiple body regions using multi-modal, multi-class machine learning models 92%
- Usability of a Machine-Learning Clinical Order Recommender System Interface for Clinical Decision Support and Physician Workflow 91%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.