Back

Komenti: A semantic text mining framework

Slater, L. T.; Bradlow, W.; Hoehndorf, R.; Motti, D. F.; Ball, S.; Gkoutos, G. V.

2020-08-04 bioinformatics
10.1101/2020.08.04.233049 bioRxiv
Show abstract

SummaryKomenti is a reasoner-enabled semantic query and information extraction framework. It is the only text mining tool that enables querying inferred knowledge from biomedical ontologies. It also contains multiple novel components for vocabulary construction and context disambiguation, which can improve the power of text mining and ontology-based analysis tasks, with a view towards making full use of the semantic provision of biomedical ontologies for text characterisation and analysis. Here, we describe Komenti and its features, and present a use case wherein we automate a clinical audit, extracting medications for hypertrophic cardiomyopathy patients from text, revealing a high precision, and identifying a sub-cohort of patients with atrial fibrillation who are not anti-coagulated, and are therefore at a higher risk of stroke. Availability and ImplementationKomenti is freely available under an open source licence at http://github.com/reality/komenti. More information concerning the use-case is available in supplementary data.

Matching journals

The top 6 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.