Back

XVCF: Exquisite Visualization of VCF Data from Genomic Experiments

Almuneef, G.; Aljouie, A.; Bokhari, Y.; Almazroa, A.; Rashid, M.

2025-05-06 bioinformatics
10.1101/2025.04.30.651450 bioRxiv
Show abstract

BackgroundHigh-throughput genomic analyses of germline and cancer genomes facilitate the identification of causal and actionable genetic variants. The recent advances in next-generation sequencing technology generated large-scale genomic and/or multi-omics datasets from different disease types or models. Owing to the huge volume of data coming out of genomic experiments, scientists are facing challenges in handling, manipulating, visualizing, and interpreting the data. Currently, available tools to visualize genetic variants from VCF files are not very user-friendly as most of them require knowledge of command line tools or scripts to install and run those software. Moreover, biologists or clinicians lack this knowledge of computer programming. Therefore, graphical user interface (GUI) based tools or software are needed to effectively summarize and visualize the huge volume of genomic data such as VCF data. MethodsWe have developed a Shiny App, interactive tool using the R programming language that utilizes other R packages like "vcfR" and "maftools" to visualize and generate quality control metrics for genetic data effectively and exquisitely. A key improvement is the addition of a user-friendly interface, providing researchers with an interactive way to explore VCF or Cancer genomics data. Our tool is powered by Shiny, making it even easier for researchers to analyze and visualize genomic variation data using a GUI. Researchers can upload their datasets and customize the analysis parameters to suit their specific research needs. ResultsA user-friendly interactive tool has been developed for the summarization and visualization of data related to genomic variation research. The application features an easy and friendly interface, allowing users to perform various functions such as data loading, summarization, and visualization interactively. XVCF offers an easy-to-use GUI platform to read genetic variant data (annotated or unannotated) and extract useful information such as read depth, mapping quality, genotype, quality control summary, and allele frequency from unannotated data. In the second module of XVCF, the cancer genomic data (annotated, so far supported by ANNOVAR) is analyzed using "maftools" to produce oncoplot, comparison of mutational load across different TCGA datasets, gene summary, etc. XVCF is available for free download from https://github.com/rashidma/XVCF. ConclusionXVCF Shiny web application can serve as a robust visualization and quality control GUI platform for the germline and cancer genomics dataset. We expect this tool will be immensely useful for researchers with less computational or technical knowledge. Being a shiny R package, XVCF can be installed across different operating systems and utilize different computer hardware configurations.

Matching journals

The top 3 journals account for 50% of the predicted probability mass.

1
BMC Bioinformatics
457 papers in training set
Top 0.1%
39.6%
2
Bioinformatics
1204 papers in training set
Top 3%
9.8%
3
GigaScience
212 papers in training set
Top 0.5%
5.5%
50% of probability mass above
4
PLOS ONE
5266 papers in training set
Top 30%
5.2%
5
Bioinformatics Advances
203 papers in training set
Top 2%
3.2%
6
Database
61 papers in training set
Top 0.3%
2.8%
7
Computational and Structural Biotechnology Journal
242 papers in training set
Top 3%
2.1%
8
Briefings in Bioinformatics
354 papers in training set
Top 4%
2.1%
9
PLOS Computational Biology
1863 papers in training set
Top 13%
2.1%
10
Nucleic Acids Research
1281 papers in training set
Top 8%
1.9%
11
Gigabyte
62 papers in training set
Top 0.6%
1.7%
12
NAR Genomics and Bioinformatics
242 papers in training set
Top 3%
1.7%
13
F1000Research
88 papers in training set
Top 2%
1.7%
14
Informatics in Medicine Unlocked
22 papers in training set
Top 0.6%
1.7%
15
Journal of Bioinformatics and Systems Biology
15 papers in training set
Top 0.1%
1.5%
16
SoftwareX
15 papers in training set
Top 0.2%
1.3%
17
BMC Genomics
406 papers in training set
Top 6%
1.1%
18
BMC Medical Genomics
50 papers in training set
Top 0.8%
1.1%
19
Scientific Reports
3612 papers in training set
Top 65%
1.1%
20
Journal of Open Source Software
25 papers in training set
Top 0.3%
1.0%
21
Genes
144 papers in training set
Top 4%
0.8%
22
BioData Mining
22 papers in training set
Top 1%
0.6%
23
Frontiers in Genetics
230 papers in training set
Top 7%
0.6%