Back

Performance characterization of PCR-free whole genome sequencing for clinical diagnosis

Zhang, N.; Zhou, M.; Zeng, F.; Wang, X.; Liu, F.; Qiao, Z.; Fan, C.; Wang, Y.; Fang, Z.; Dai, W.; Xiang, J.; Sun, J.; Peng, Z.; Song, L.; Sun, Y.

2020-06-20 genetics
10.1101/2020.06.19.160739 bioRxiv
Show abstract

PurposeTo evaluate the performance of PCR-free whole genome sequencing (WGS) for clinical diagnosis, and thereby revealing how experimental parameters affect variant detection. MethodsAll the 5 NA12878 samples were sequenced using MGISEQ-2000. NA12878 samples underwent WGS with differing DNA input and library preparation protocol (PCR-based versus PCR-free protocols for library preparation). The DP (depth of coverage) and GQ (genotype quality) of each sample were compared. We developed a systematic WGS pipeline for the analysis of down-sampling samples of the 5 NA12878 samples. The performance of each sample was measured for sensitivity, coverage of depth and breadth of coverage of disease-related genes and CNVs. ResultsIn general, NA12878-2 (PCR-free WGS) showed better DP and GQ distribution than NA12878-1 (PCR-based WGS). With a mean depth of ~40X, the sensitivity of homozyous and heterozygous SNPs of NA12878-2 showed higher sensitivity (>99.77% and > 99.82%) than NA12878-1, and positive predictive value (PPV) exceeded 99.98% and 99.07%. The sensitivity and PPV of homozygous and heterozygous indels for NA12878-2 (PCR-free WGS) showed great improvement than NA128878-1. The breadths of coverage for disease-related genes and CNVs are slightly better for samples with PCR-free library preparation protocol than the sample with PCR-based library preparation protocol. DNA input also influences the performance of variant detection in samples with PCR-free WGS. ConclusionDifferent experimental parameters may affect variant detection for clinical WGS. Clinical scientists should know the range of sensitivity of variants for different methods of WGS, which would be useful when interpreting and delivering clinical reports.

Matching journals

The top 5 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.