Genomic Dimensionality Bounds Mixed-Model Association Power and Fine-Mapping Resolution
Jiang, J.
Show abstract
Mixed-model genome-wide association studies (GWAS) behave differently in livestock than in humans, yet a unified explanation is lacking. Analyses using the full genomic relationship matrix (full-GRM; from genome-wide SNPs) yield only a few significant peaks even with hundreds of thousands of animals, whereas leave-one-chromosome-out (LOCO), numerator-relationship-matrix, and sparse-GRM approaches report many broad associations over similar data. Here we develop a framework that traces these behaviors to the low effective genomic dimensionality, Me, of small-Ne populations. Starting from the mixed-model association statistic, we derive the per-SNP non-centrality parameter under full-GRM testing and show that its sample-size dependence is fully captured by a sigmoid sum S(N) over LD-matrix eigenmodes. S(N) grows concavely with N toward a practical ceiling Me, from which the framework predicts a full-GRM detection floor qmin {approx} 30h2/Me on per-SNP proportion of phenotypic variance explained at 50% power (e.g., [~]0.09% for cattle at h2 = 0.3), and a fine-mapping resolution limit through both Me and 4Ned-scaled LD decay. LOCO bypasses the full-GRM ceiling but detects LD-aggregated block-level signals rather than SNP-level excess effects, explaining its inflation in livestock and agreement with full-GRM in humans. The framework is supported by analyses of real livestock chip panels, coalescent eigenvalue spectra, and phenotype simulations. The same sigmoid sum, normalized as{varphi} (N) {approx} S(N)/Me, recovers the in-sample average GBLUP reliability, unifying why genomic prediction is comparatively easy in livestock while SNP-level mapping and fine-mapping remain difficult. For livestock GWAS aimed at SNP-level interpretation (e.g., candidate-gene prioritization, fine-mapping, or molecular-QTL colocalization), the framework supports full-GRM approaches as the appropriate default. Article SummaryGenome-wide association methods that largely agree in humans can give strikingly different results in livestock, complicating their use and interpretation for animal breeding. This study develops a framework, derived from the association test statistic, that traces this divergence to one cause. The small effective population size of livestock explains why some methods detect few signals even in huge datasets, why others report many broad associations, why livestock differ from humans, how precisely causal variants can be mapped, and why prediction stays comparatively easy. The framework predicts these outcomes quantitatively, with explicit formulas. It guides method choice and sets realistic expectations.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Mendelian randomization accounting for correlated and uncorrelated pleiotropic effects using genome-wide summary statistics. 95%
- Fast and flexible joint fine-mapping of multiple traits via the Sum of Single Effects model 95%
- LDAK-KVIK performs fast and powerful mixed-model association analysis of quantitative and binary phenotypes 95%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.