Back

Recent positive selection implicates IP6K3 and MAPT as metabolically relevant loci in South Asians

Pennarun, E.; Banfalvi, B.; Li, Y.; Bui, V.; Hodgson, S.; Bigossi, M.; Arnab, S.; Naimah, T.; Rison, S.; Stow, D.; Baskar, V.; Saravanan, J.; Radha, V.; G&H research team, ; MDRF research team, ; Mohan, V.; DeGiorgio, M.; Mohan Anjana, R.; Finer, S.; Fumagalli, M.; Siddiqui, M. K.

2026-03-13 genetic and genomic medicine
10.64898/2026.03.12.26348237 medRxiv
Show abstract

South Asians constitute 25% of the global population yet account for 33% of individuals living with type 2 diabetes (T2D) (1). Compared with individuals of European ancestry, South Asians develop metabolic disease at lower body mass index and progress more rapidly to cardiometabolic complication (2,3). Despite disproportionate disease burden, South Asians represented less than 1% of participants in genome-wide association studies (4). This underrepresentation limits biological insight and hampers genomic discovery. Here we integrate genome-wide signatures of recent positive selection with cross-trait genetic association data across 13 South Asian populations (5,6) (n = 676 whole genomes). We identify 1,797 genes putatively under selection, prioritise 65 shared across the subcontinent, and refine four loci--RBM6, IP6K3, MAPT and PEPD--using colocalisation (from publicly available multi-ancestry summary statistics) and multi-trait fine-mapping (7,8) in two independent south Asian cohorts (9,10). Integrating signatures of recent positive selection with genetic association data prioritises metabolically relevant loci for T2D in South Asians and establishes a scalable framework for locus discovery in underrepresented populations, without reliance on ancestry-matched molecular quantitative trait loci (QTL) resources.

Matching journals

The top 5 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.