A corpus for differential diagnosis: an eye diseases use case
Jimeno Yepes, A.; Martinez Iraola, D.; Barnard, P.; Joy, T.
Show abstract
We have created a corpus for the extraction of information related to diagnosis from scientific literature focused on eye diseases. It was shown that the annotation of entities has a relatively large agreement among annotators, which translates into strong performance of the trained methods, mostly BioBERT. Furthermore it was observed that relation annotation in this domain has challenges, which might require additional exploration. When using the trained models on MEDLINE, we could identify confirmed knowledge about the diagnosis of eye diseases and relevant new information, which supports the developments in this work. The corpus that we have developed is publicly available, thus the scientific community is able to reproduce our work and reuse the corpus in their work.
Matching journals
The top 9 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Automated recognition of functional compound-protein relationships in literature 95%
- Comparison of local large language models for extraction of signs and symptoms data from electronic health records 93%
- Understanding signaling and metabolic paths using semantified and harmonized information about biological interactions 93%
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- Evaluating Semantic Similarity Methods for Comparison of Text-derived Phenotype Profiles 95%
- Ontology-based expansion of virtual gene panels to improve diagnostic efficiency for rare genetic diseases 95%
- Towards semantic interoperability: finding and repairing hidden contradictions in biomedical ontologies 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.