Back

Gene tree discordance, rapid diversification and convergence impact phylogenetic inference in Zygophylloideae

van der Merwe, P. D. W.; Pirie, M. D.; Bellstedt, D. U.

2026-01-17 evolutionary biology
10.64898/2026.01.16.699981 bioRxiv
Show abstract

The Zygophyllaceae subfamily Zygophylloideae includes seven genera and around 180 species distributed worldwide particularly in arid areas where they show various adaptations to drought including C4 and CAM photosynthesis. Phylogenetic relationships of Zygophylloideae have been analysed in several studies using Sanger sequenced markers, but deeper relationships between clades have remained recalcitrant. Here, we use ion semi-conductor sequencing of chloroplast and nuclear ribosomal regions to analyse these relationships and assess why they have proved so challenging to resolve. We analyse new data for 39 chloroplast genes and nuclear ribosomal 18S, 5.8S, and 26S gene and internal transcribed spacers for key species of Zygophyllum sensu stricto, Roepera, Tetraena, Fagonia, Zygophyllum stapffii and monotypic Augea capensis; and Tribulus of the subfamily Tribuloideae as outgroup, compared this to published chloroplast genomes, summarising phylogenetic results separately from non-coding introns and intergenic spacers, photosynthetic genes and non-photosynthetic genes. We combined this data with sequences of three highly variable non-coding chloroplast regions and ITS used in previous studies of greater numbers of taxa and assessed the effect on phylogenetic inference. We used a chloroplast sequence matrix to infer a dated phylogeny. Adding data did not always improve phylogenetic resolution which was impacted by non-informative variation in chloroplast photosynthesis related genes as well as by conflicting chloroplast and nuclear signals. This provides evidence for both selection and introgressive hybridization and/or incomplete lineage sorting during an ancient rapid divergence that will require multiple independent, informative, and neutral gene trees to resolve.

Matching journals

The top 5 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.