Systematic identification of cargo-mobilizing genetic elements reveals new dimensions of eukaryotic diversity
Cargo-mobilizing mobile elements (CMEs) are genetic entities that faithfully transpose diverse protein coding sequences. Although common in bacteria, we know little about eukaryotic CMEs because no appropriate tools exist for their annotation. For example, Starships are giant fungal CMEs whose funct...
Gespeichert in:
Veröffentlicht in: | Nucleic acids research 2024-06, Vol.52 (10), p.5496-5513 |
---|---|
Hauptverfasser: | , |
Format: | Artikel |
Sprache: | eng |
Schlagworte: | |
Online-Zugang: | Volltext |
Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Zusammenfassung: | Cargo-mobilizing mobile elements (CMEs) are genetic entities that faithfully transpose diverse protein coding sequences. Although common in bacteria, we know little about eukaryotic CMEs because no appropriate tools exist for their annotation. For example, Starships are giant fungal CMEs whose functions are largely unknown because they require time-intensive manual curation. To address this knowledge gap, we developed starfish, a computational workflow for high-throughput eukaryotic CME annotation. We applied starfish to 2 899 genomes of 1 649 fungal species and found that starfish recovers known Starships with 95% combined precision and recall while expanding the number of annotated elements ten-fold. Extant Starship diversity is partitioned into 11 families that differ in their enrichment patterns across fungal classes. Starship cargo changes rapidly such that elements from the same family differ substantially in their functional repertoires, which are predicted to contribute to diverse biological processes such as metabolism. Many elements have convergently evolved to insert into 5S rDNA and AT-rich sequence while others integrate into random locations, revealing both specialist and generalist strategies for persistence. Our work establishes a framework for advancing mobile element biology and provides the means to investigate an emerging dimension of eukaryotic genetic diversity, that of genomes within genomes. |
---|---|
ISSN: | 0305-1048 1362-4962 1362-4962 |
DOI: | 10.1093/nar/gkae327 |