The Silent Author in Urology: Quantifying Large Language Model Influence at the Corpus Level
Arezki, A.; Tsui, J.; Kassouf, W.
Show abstract
IntroductionThe recent increasing use of large language models (LLMs) such as ChatGPT in scientific writing raises concerns about authenticity and accuracy in biomedical literature. This study aims to quantify the prevalence of Artificial Intelligence (AI)-generated text in urology abstracts from 2010 to 2024 and assess its variation across journal categories and impact factor quartiles. MethodsA retrospective analysis was conducted on 64,444 unique abstracts from 38 journals in the field of urology published between 2010-2024 and retrieved via the Entrez API from PubMed. Abstracts were then categorized by subspecialty and stratified by impact factor (IF). A synthetic reference corpus of 10,000 abstracts was generated using GPT-3.5-turbo. A mixture model estimated the proportion of AI-generated text [Formula] annually, using maximum likelihood estimation with Laplace smoothing. Calibration and specificity of the estimator were assessed using pre-LLM abstracts and controlled mixtures of real and synthetic text. Statistical analyses were performed using Python 3.13. ResultsThe proportion of AI-like text was negligible from 2010 to 2019, rising to 1.8% in 2020 and 5.3% in 2024. In 2024, Mens Health journals showed the highest AI-like text, while Oncology journals had the lowest. Journals in the highest and lowest IF quartiles showed a higher proportion of AI-like text than mid-quartiles. Type-Token Ratio (TTR) remained stable across the study period. Our validation calibration set showed an area under the curve of 0.5326. ConclusionTextual patterns similar to AI-generated language has risen sharply in urology abstracts since ChatGPTs release in late 2022, and this trend varies by journal type and IF. Patient summaryWe looked at whether large language models (LLM) such as ChatGPT are influencing the way urology research is written. We found that signs of LLM-like text were rare in abstracts before 2020 but increased in recent years, especially after the release of ChatGPT. Some journal types use these tools more than others. These findings help raise awareness about the growing role of artificial intelligence in scientific communication.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Performance of o1 pro and GPT-4 in self-assessment questions for nephrology board renewal 91%
- Emerging Applications of NLP and Large Language Models in Gastroenterology and Hepatology: A Systematic Review 90%
- Semantic and geographical analysis of Covid-19 trials reveals a fragmented clinical research landscape likely to impair informativeness. 87%
Similar papers in this journal
- Diversity and inclusion: A hidden additional benefit of Open Data 93%
- Artificial Intelligence's Contribution to Biomedical Literature Search: Revolutionizing or Complicating? 93%
- Inferring Gender from First Names: Comparing the Accuracy of Genderize, Gender API, and the gender R Package on Authors of Diverse Nationality 93%
Similar papers in this journal
- Introducing the EMPIRE Index: A novel, value-based metric framework to measure the impact of medical publications 94%
- The Rise of Open Data Practices Among Bioscientists at the University of Edinburgh 92%
- COVID-19-related research data availability and quality according to the FAIR principles: A meta-research study 92%
Similar papers in this journal
- The effect of digital-enabled multidisciplinary therapy conferences on efficiency and quality of the decision making in prostate-cancer care 92%
- Network Graph Representation of COVID-19 Scientific Publications to Aid Knowledge Discovery 91%
- Impact of the Federated Data Platform's digital surgery scheduling system on elective theatre utilisation at an NHS Trust: an interrupted time series analysis 89%
Similar papers in this journal
- GPT for RCTs?: Using AI to measure adherence to reporting guidelines 93%
- Publishing at any cost: a cross-sectional study of the amount that medical researchers spend on open-access publishing each year 92%
- Prescribing patterns for medical treatment of suspected prostatic obstruction: A spatiotemporal statistical analysis of Scottish open access data 91%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.