Back

Animal retrozymes are nonautonomous sequences of Penelope-like elements

Li, Z.; Clavereau, I.; Pollet, N.

2026-07-24 genomics
10.64898/2026.07.22.740005 bioRxiv
Show abstract

Hammerhead ribozymes (HHRs) are small catalytic RNAs found across diverse life forms. In animal genomes, they can be encoded by genes organised in dispersed copies or in tandem genomic arrangements. These tandemly organised forms, known as Non-LTR retrozymes, were recently identified as a distinct group of non-autonomous retrotransposons, likely mobilised via a rolling-circle transposition mechanism and potentially involved in host transcriptome regulation. However, their evolutionary origins remain poorly understood. Here, we investigate the presence, genomic distribution, and possible origins of Non-LTR retrozymes across a broad range of vertebrate species. We find that these elements display a patchy phylogenetic distribution, notably absent from the Aves and Mammalia lineages. In species where they are present, retrozyme copy number, consensus length, and monomer proportion vary widely across species and retrozyme families, suggesting diverse amplification dynamics. Genomic mapping reveals a significant enrichment of Non-LTR retrozymes in intergenic regions and their exclusion from introns and exons, indicating selective pressure against genic insertion. Strikingly, the phylogenetic distribution of Non-LTR retrozymes coincides with that of Penelope-like elements (PLEs). Phylogenetic analysis further shows that the pLTR region of PLEs is closely related to Non-LTR retrozymes, supporting the hypothesis that Non-LTR retrozymes are non-autonomous derivatives of PLEs. Together, our findings shed new light on the evolutionary origin and genomic behaviour of Non-LTR retrozymes and underscore their potential regulatory roles in vertebrate genomes.

Matching journals

The top 2 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.