Back

SPARClink: an interactive tool to visualize the impact of the SPARC program

Soundarajan, S.; Kuruppu, S.; Singh, A.; Kim, J.; Achalla, M.

2021-10-25 bioinformatics
10.1101/2021.10.22.465507 bioRxiv
Show abstract

The NIH SPARC program seeks to accelerate the development of therapeutic devices that modulate electrical activity in nerves to improve organ function. SPARC-funded researchers are generating rich datasets from neuromodulation research that are curated and shared according to FAIR (Findable, Accessible, Interoperable, and Reusable) guidelines and are accessible to the public on the SPARC data portal. Keeping track of the utilization of these datasets within the larger research community is a feature that will benefit data generating researchers in showcasing the impact of their SPARC outcomes. This will also allow the SPARC program to display the impact of the FAIR data curation and sharing practices that have been implemented. This manuscript provides the methods and outcomes of SPARClink, our web tool for visualizing the impact of SPARC, which won the 2nd prize at the 2021 SPARC FAIR Codeathon. With SPARClink, we built a system that automatically and continuously finds new published SPARC scientific outputs (datasets, publications, protocols) and the external resources referring to them. SPARC datasets and protocols are queried using publicly accessible REST APIs (provided by Pennsieve and Protocols.io) and stored in a publicly accessible database. Citation information for these resources is retrieved using the NIH reporter API and NCBI Entrez system. A novel knowledge-graph-based structure was created to visualize the results of these queries and showcase the impact that the FAIR data principles can have on the research landscape when they are adopted by a consortium.

Matching journals

The top 3 journals account for 50% of the predicted probability mass.

1
Database
61 papers in training set
Top 0.1%
40.3%
2
Nucleic Acids Research
1281 papers in training set
Top 3%
7.4%
3
Bioinformatics
1204 papers in training set
Top 3%
6.9%
50% of probability mass above
4
PLOS ONE
5266 papers in training set
Top 27%
5.7%
5
Bioinformatics Advances
203 papers in training set
Top 1%
4.4%
6
BMC Bioinformatics
457 papers in training set
Top 3%
3.3%
7
PLOS Digital Health
106 papers in training set
Top 2%
2.8%
8
GigaScience
212 papers in training set
Top 1%
2.7%
9
PLOS Computational Biology
1863 papers in training set
Top 13%
1.9%
10
NAR Genomics and Bioinformatics
242 papers in training set
Top 3%
1.7%
11
F1000Research
88 papers in training set
Top 2%
1.5%
12
BMC Research Notes
33 papers in training set
Top 0.5%
1.4%
13
Computational and Structural Biotechnology Journal
242 papers in training set
Top 4%
1.4%
14
PeerJ
308 papers in training set
Top 10%
1.0%
15
Journal of Molecular Biology
232 papers in training set
Top 3%
0.9%
16
BioData Mining
22 papers in training set
Top 0.7%
0.9%
17
Briefings in Bioinformatics
354 papers in training set
Top 7%
0.9%
18
International Journal of Medical Informatics
26 papers in training set
Top 1%
0.9%
19
Biology
45 papers in training set
Top 1%
0.6%
20
BMC Biology
265 papers in training set
Top 6%
0.6%
21
Computer Methods and Programs in Biomedicine
28 papers in training set
Top 1%
0.6%
22
PROTEOMICS
43 papers in training set
Top 0.9%
0.6%
23
Frontiers in Medicine
120 papers in training set
Top 5%
0.6%
24
BMC Genomics
406 papers in training set
Top 9%
0.6%