A systematic mapping of the genomic and proteomic variation associated with monogenic diabetes
Kuznetsova, K.; Vasicek, J.; Skiadopoulou, D.; Molnes, J.; Udler, M.; Johansson, S.; Njolstad, P. R.; Manning, A. K.; Vaudel, M.
Show abstract
AimsMonogenic diabetes is characterized as a group of diseases caused by rare variants in single genes. Multiple genes have been described to be responsible for monogenic diabetes, but the information on the variants is not unified among different resources. In this work, we aimed to develop an automated pipeline that collects all the genetic variants associated with monogenic diabetes from different resources, unify the data and translate the genetic sequences to the proteins. MethodsThe pipeline developed in this work is written in Python with the use of Jupyter notebook. It consists of 6 modules that can be implemented separately. The translation step is performed using the ProVar tool also written in Python. All the code along with the intermediate and final results is available for public access and reuse. ResultsThe resulting database had 2701 genomic variants in total and was divided into two levels: the variants reported to have an association with monogenic diabetes and the variants that have evidence of pathogenicity. Of them, 2565 variants were found in the ClinVar database and the rest 136 were found in the literature showing that the overlap between resources is not absolute. ConclusionsWe have developed an automated pipeline for collecting and harmonizing data on genetic variants associated with monogenic diabetes. Furthermore, we have translated variant genetic sequences into protein sequences accounting for all protein isoforms and their variants. This allows researchers to consolidate information on variant genes and proteins associated with monogenic diabetes and facilitates their study using proteomics or structural biology. Our open and flexible implementation using Jupyter notebooks enables tailoring and modifying the pipeline and its application to other rare diseases. Research in contextO_LIMonogenic diabetes is a group of Mendelian diseases with an autosomal-dominant pattern of inheritance. C_LIO_LIMonogenic diabetes is mainly caused by rare genetic variants that are usually evaluated manually. C_LIO_LIThe data on the variants are stored in several resources and are not unified in terms of the genomic coordinates, alleles, and variant annotation. C_LIO_LIWhat can be done for the systematic evaluation of the variants and their protein consequences? C_LIO_LIIn this work, we have created an automated Jupyter notebook-based pipeline for the collection and unification of the variants associated with monogenic diabetes. C_LIO_LIThe database of the genetic variants was created and translated to all possible variant protein sequences. C_LIO_LIThese results will be used for the analysis of proteomics data and protein structure modeling. C_LI
Matching journals
The top 10 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- \"Mitochondrial GWAS and Association of Nuclear - Mitochondrial Epistasis with BMI in T1DM Patients\" 90%
- GeneTerpret: a customizable multilayer approach to genomic variant prioritization and interpretation 90%
- Genome-wide survey of tandem repeats by nanopore sequencing shows that disease-associated repeats are more polymorphic in the general population 89%
Similar papers in this journal
- REVEL is better at predicting pathogenicity of loss-of-function than gain-of-function variants 92%
- Splicing impact of deep exonic missense variants in CAPN3 explored systematically by minigene functional assay 91%
- Using single molecule Molecular Inversion Probes as a cost-effective, high-throughput sequencing approach to target all genes and loci associated with macular diseases 91%
Similar papers in this journal
- Hypusinated eIF5A is expressed in pancreas and spleen of individuals with type 1 and type 2 diabetes 92%
- Towards development of a statistical framework to evaluate myotonic dystrophy type 1 mRNA biomarkers in the context of a clinical trial 91%
- Replacing murine insulin 1 with human insulin protects NOD mice from diabetes 91%
Similar papers in this journal
- Young onset diabetes in Asian Indians is associated with lower measured and genetically determined beta-cell function: an INSPIRED study 94%
- Subgroups of young type 2 diabetes in India reveal insulin deficiency as a major driver 93%
- Development and validation of a Trans-Ancestry polygenic risk score for Type 1 Diabetes 93%
Similar papers in this journal
- Common variation in a long non-coding RNA gene modulates variation of circulating TGF- β 2 levels in metastatic colorectal cancer patients (Alliance) 89%
- Regulatory Modules of Human Thermogenic Adipocytes: Functional Genomics of Meta-Analyses Derived Marker-Genes 89%
- Expanding the Potential Genes of Inborn Errors of Immunity through Protein Interactions 89%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.