Enhanced Data Pre-processing for the Identification of Alzheimer's Disease-Associated SNPs
Alves, J. F.; Cerri, R.; Costa, E.; Brito, L. F.; Xavier, A.
Show abstract
Alzheimers Disease (AD) is a complex neurodegenerative disorder that has gained significant attention in scientific research, particularly since the Human Genome Project. Based on twin studies that utilize the resemblance of Alzheimers disease risk between pairs of twins, it has been found that the overall heritability of the disease is estimated at 0.58. When shared environmental factors are taken into account, the maximum heritability reaches 0.79. This suggests that approximately 58-79% of the susceptibility to late-onset Alzheimers disease can be attributed to genetic factors [4]. In 2022, it is estimated that AD will affect over 50 million people worldwide, and its economic burden exceeds a trillion US dollars per year. One promising approach is Genome-Wide Association Studies (GWAS), which allow the identification of genetic variants associated with AD susceptibility. Of particular interest are Single Nucleotide Polymorphisms (SNPs), which represent variations in a single nucleotide base in the DNA sequence. In this study, we investigated the association between SNPs and AD susceptibility by applying various quality control (QC) parameters during data pre-processing and rank the SNP associations through mixed linear models-based GWAS implemented in BLUPF90. Our findings indicate that the identified SNPs are located in regions already associated with Alzheimers Disease, including non-coding regions. We also investigated the impact of incorporating demographic data into our models. However, the results indicated that the inclusion of such data did not yield any benefits for the model. This study highlights the importance of GWAS in identifying potential genetic risk factors for AD and underscores the need for further research to gain a better understanding of the complex genetic mechanisms underlying this debilitating disease.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- ChatGPT-Enhanced ROC Analysis (CERA): A Shiny Web Tool for Finding Optimal Cutoff in Biomarker Analysis 95%
- Identification of functionally connected multi-omic biomarkers for Alzheimer’s Disease using modularity-constrained Lasso 94%
- Testing gene-environment interactions for rare and/or common variants in sequencing association studies 94%
Similar papers in this journal
- Extensive In Silico Analysis of the Functional and Structural Consequences of SNPs in Human ARX Gene 94%
- An Inexpensive Smartphone-Based Device and Predictive Models for Rapid, Non-Invasive, and Point-of-Care Monitoring of Ocular and Cardiovascular Complications Related to Diabetes 92%
- Finding Consensus miRNAs Silencing KLF1 Expression as A Promising Therapeutic Option of Sickle Cell Anemia 92%
Similar papers in this journal
- Analysis of Pan-Omics Data in Human Interactome Network (APODHIN) 93%
- Module analysis using single-patient differential expression signatures improve the power of association study for Alzheimer's disease 93%
- GMQN: A reference-based method for correcting batch effects as well as probes bias in HumanMethylation BeadChip 92%
Similar papers in this journal
- Integrated Analysis of Tissue-specific Gene Expression in Diabetes by Tensor Decomposition Can Identify Possible Associated Diseases. 95%
- Integrating Bioinformatics and Artificial Intelligence Methods to identify disruptive STAT1 variants impacting Protein Stability and Function 93%
- DLO Hi-C Tool for Digestion-Ligation-Only Hi-C Chromosome Conformation Capture Data Analysis 92%
Similar papers in this journal
- AITeQ: A machine learning framework for Alzheimer's prediction using a distinctive 5-gene signature 96%
- LDBlockShow: a fast and convenient tool for visualizing linkage disequilibrium and haplotype blocks based on variant call format files 95%
- Blood-based transcriptomic signature panel identification for cancer diagnosis: Benchmarking of feature extraction methods 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.