WikiGOA: Gene set enrichment analysis based on Wikipedia and the Gene Ontology
Lubiana, T.; Dias, T. L.; Peixe, D. G.; Nakaya, H. T. I.
Show abstract
O_LIGene sets curated to Gene Ontology terms are widely used by the transcriptomics community C_LIO_LIPresence in Wikipedia is a common proxy for the relevance of a concept. C_LIO_LIIn this work, we describe the use of Wikidata to generate a dataset comprising only gene sets with a corresponding Wikipedia page. C_LIO_LIWe refer to the dataset as "WikiGOA", standing for "Wikipedia Gene Ontology Annotations" C_LIO_LIWe use the dataset to analyze gene expression data and show that it provides readily understandable results. C_LIO_LIWe envision WikiGOA to be useful for exploring complex biological datasets both in academic research and educational contexts. C_LI NoteThis report was written in a non-standard, experimental format, where assertions are expressed in bullet points. This was done to clarify statements and assumptions, simplify reading and pave the way for conversion to structured formats (e.g., nanopublications). [1]
Matching journals
The top 8 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- ChatGPT-Enhanced ROC Analysis (CERA): A Shiny Web Tool for Finding Optimal Cutoff in Biomarker Analysis 93%
- Cluster analysis on high dimensional RNA-seq data with applications to cancer research- An evaluation study 93%
- Automated recognition of functional compound-protein relationships in literature 93%
Similar papers in this journal
- PathWalks: Identifying pathway communities using a disease-related map of integrated information 94%
- Differential co-expression network analysis with DCoNA reveals isomiR targeting aberrations in prostate cancer 92%
- Flame (v2.0): advanced integration and interpretation of functional enrichment results from multiple sources 92%
Similar papers in this journal
- Analysis of Pan-Omics Data in Human Interactome Network (APODHIN) 95%
- Universal nature of drug treatment responses in drug-tissue-wide model-animal experiments using tensordecomposition-based unsupervised featureextraction 93%
- Tensor decomposition-Based Unsupervised Feature Extraction Applied to Single-Cell Gene Expression Analysis 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.