Back

The SMART-AI trial: Real-time cholangioscopy artificial intelligence for the classification of biliary strictures

Marya, N.; Powers, P.; Marcello, M.; Rau, P.; Nasser-Ghodsi, N.; Marshall, C.; Zivny, J.; AbiMansour, J.; Chandrasekhara, V.

2025-12-27 gastroenterology
10.64898/2025.12.17.25342271 medRxiv
Show abstract

BackgroundSampling techniques have poor accuracy for classifying biliary strictures as benign or malignant. Previously, a cholangioscopy artificial intelligence (AI) outperfromed sampling techniques based solely on analysis of previously recorded cholangioscopy footage. The aim of this trial was to compare the performance of a real-time cholangioscopy AI to both sampling techniques and human observers for the task of biliary stricture classification. MethodsA cholangioscopy AI computer connected directly to a cholangioscope console. The computer analyzed the cholangioscopy video stream during procedures for suspected biliary strictures. The primary outcome of the study was comparison of the performance of cholangioscopy AI to sampling techniques - brush cytology and transpapillary forceps biopsy - for biliary stricture classification. Secondary outcomes included comparison of the AI classification performance to that of human observers (separated into junior-level and experienced-level cohorts) who reviewed the cholangioscopy footage. ResultsA total of 41 patients were enrolled in the trial and had biliary strictures analyzed by cholangioscopy AI. For the classification of strictures, the AI had greater classification accuracy than standard sampling techniques (87.8% versus 67.4%; p = 0.043). Additionally, the cholangioscopy AI was significantly more accurate for biliary stricture classification than both junior-level (87.8% versus 61.5%; p = 0.001) and experienced endoscopists (87.8% versus 63.15%; p = 0.011). ConclusionsThis trial demonstrates that sampling techniques and human assessment of biliary strictures are flawed and there may be a benefit to the use of a cholangioscopy AI system to aid in biliary stricture classification.

Matching journals

The top 5 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.