Inter-Rater Agreement for the Annotation of Neurologic Concepts in Electronic Health Records
Oommen, C.; Howlett-Prieto, Q.; Carrithers, M. D.; Hier, D. B.
Show abstract
The extraction of patient signs and symptoms recorded as free text in electronic health records is critical for precision medicine. Once extracted, signs and symptoms can be made computable by mapping to clinical concepts in an ontology. Extracting clinical concepts from free text is tedious and time-consuming. Prior studies have suggested that inter-rater agreement for clinical concept extraction is low. We have examined inter-rater agreement for annotating neurologic concepts in clinical notes from electronic health records. After training on the annotation process, the annotation tool, and the supporting neuro-ontology, three raters annotated 15 clinical notes in three rounds. Inter-rater agreement between the three annotators was high for text span and category label. A machine annotator based on a convolutional neural network had a high level of agreement with the human annotators, but one that was lower than human inter-rater agreement. We conclude that high levels of agreement between human annotators are possible with appropriate training and annotation tools. Furthermore, more training examples combined with improvements in neural networks and natural language processing should make machine annotators capable of high throughput automated clinical concept extraction with high levels of agreement with human annotators.
Matching journals
The top 2 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Development and Validation of Phenotype Classifiers across Multiple Sites in the Observational Health Sciences and Informatics (OHDSI) Network 93%
- Large Language Models Facilitate the Generation of Electronic Health Record Phenotyping Algorithms 93%
- Annotation-preserving machine translation of English corpora to validate Dutch clinical concept extraction tools 93%
Similar papers in this journal
- Transforming Estonian health data to the Observational Medical Outcomes Partnership (OMOP) Common Data Model: lessons learned 93%
- Natural Language Processing for Automated Annotation of Medication Mentions in Primary Care Visit Conversations 93%
- Automating Evaluation of LLM-generated Responses to Patient Questions about Rare Diseases 92%
Similar papers in this journal
- A method for rapid machine learning development for data mining with Doctor-In-The-Loop 94%
- tbiExtractor: A framework for Extracting Traumatic Brain Injury Common Data Elements from Radiology Reports 93%
- CohortDiagnostics: phenotype evaluation across a network of observational data sources using population-level characterization 93%
Similar papers in this journal
- Development of a Post-Acute Sequelae of COVID-19 (PASC) Symptom Lexicon Using Electronic Health Record Clinical Notes 94%
- ConceptWAS: a high-throughput method for early identification of COVID-19 presenting symptoms 94%
- Developing A Deep Learning Natural Language Processing Algorithm For Automated Reporting Of Adverse Drug Reactions 94%
Similar papers in this journal
- Evaluating Semantic Similarity Methods for Comparison of Text-derived Phenotype Profiles 92%
- Automated abstraction of clinical parameters of multiple myeloma from real-world clinical notes using large language models 92%
- Temporal Relationship of Computed and Structured Diagnoses in Electronic Health Record Data 91%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.