Back

Medical Students' Use of Large Language Models: A National Survey

Barr, A. A.; Rozman, R. C.; Liu, K.; Pham, M.; Klarenbach, Z.; Chinna-Meyyappan, A.; Hassan, A. Y.; Zarychta, M.; El Ferri, O.; Al-Khaz'Aly, A.; Datt, P.; Herik, A. I.; Sadek, K.; Paget, M.; Holodinsky, J. K.

2026-01-29 medical education
10.64898/2026.01.26.26344898 medRxiv
Show abstract

BackgroundLarge language models (LLMs) are increasingly embedded in medical education and clinical care settings, yet limited empirical data describe medical students in Canadas use and perceptions of these tools. We aimed to characterize student engagement including LLMs used, frequency, purposes, trust, accuracy, perceived impacts, and attitudes toward educational and clinical integration. MethodsWe conducted a national survey of medical students in Canada distributed between November and December 2025. We summarized responses using descriptive statistics and compared results between students in preclerkship versus clerkship using Fishers exact test. ResultsAmong 286 respondents from 10 medical schools, 96.50% reported using at least one LLM. The most commonly used LLMs were ChatGPT (93.36%) and OpenEvidence (57.69%). Daily/weekly use was most frequent for coursework assistance (60.22%) and clinical questions (57.14%). Most respondents reported positive impacts on efficiency (81.62%), learning (77.01%), and academic performance (59.49%). Students commonly reported encountering inaccurate information (90.18%). Formal instruction on LLM use was uncommon (10.95%), though 67.67% of students agreed medical schools should integrate formal instruction on LLMs. Only 21.43% of respondents felt adequately educated on data privacy regulations applicable to these tools. ConclusionsLLM use among surveyed medical students in Canada was nearly universal and perceived favourably. However, students reported exposure to inaccurate outputs and substantial gaps in formal training and privacy literacy. These findings support the development of structured curricular guidance on appropriate application of these tools, including information verification practices and ethical, privacy-aware engagement.

Published in International Journal of Medical Informatics (predicted rank #5) · training set

Matching journals

The top 2 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.