skip to main content


Title: Comprehensive database and evolutionary dynamics of U12-type introns
Abstract During nuclear maturation of most eukaryotic pre-messenger RNAs and long non-coding RNAs, introns are removed through the process of RNA splicing. Different classes of introns are excised by the U2-type or the U12-type spliceosomes, large complexes of small nuclear ribonucleoprotein particles and associated proteins. We created intronIC, a program for assigning intron class to all introns in a given genome, and used it on 24 eukaryotic genomes to create the Intron Annotation and Orthology Database (IAOD). We then used the data in the IAOD to revisit several hypotheses concerning the evolution of the two classes of spliceosomal introns, finding support for the class conversion model explaining the low abundance of U12-type introns in modern genomes.  more » « less
Award ID(s):
1616878
NSF-PAR ID:
10331436
Author(s) / Creator(s):
; ; ; ;
Date Published:
Journal Name:
Nucleic Acids Research
ISSN:
0305-1048
Format(s):
Medium: X
Sponsoring Org:
National Science Foundation
More Like this
  1. Ouangraoua, Aida (Ed.)
    Abstract Previous evolutionary reconstructions have concluded that early eukaryotic ancestors including both the last common ancestor of eukaryotes and of all fungi had intron-rich genomes. By contrast, some extant eukaryotes have few introns, underscoring the complex histories of intron–exon structures, and raising the question as to why these few introns are retained. Here, we have used recently available fungal genomes to address a variety of questions related to intron evolution. Evolutionary reconstruction of intron presence and absence using 263 diverse fungal species supports the idea that massive intron reduction through intron loss has occurred in multiple clades. The intron densities estimated in various fungal ancestors differ from zero to 7.6 introns per 1 kb of protein-coding sequence. Massive intron loss has occurred not only in microsporidian parasites and saccharomycetous yeasts, but also in diverse smuts and allies. To investigate the roles of the remaining introns in highly-reduced species, we have searched for their special characteristics in eight intron-poor fungi. Notably, the introns of ribosome-associated genes RPL7 and NOG2 have conserved positions; both intron-containing genes encoding snoRNAs. Furthermore, both the proteins and snoRNAs are involved in ribosome biogenesis, suggesting that the expression of the protein-coding genes and noncoding snoRNAs may be functionally coordinated. Indeed, these introns are also conserved in three-quarters of fungi species. Our study shows that fungal introns have a complex evolutionary history and underappreciated roles in gene expression. 
    more » « less
  2. Abstract

    Spliceosomal introns are gene segments removed from RNA transcripts by ribonucleoprotein machineries called spliceosomes. In some eukaryotes a second ‘minor’ spliceosome is responsible for processing a tiny minority of introns. Despite its seemingly modest role, minor splicing has persisted for roughly 1.5 billion years of eukaryotic evolution. Identifying minor introns in over 3000 eukaryotic genomes, we report diverse evolutionary histories including surprisingly high numbers in some fungi and green algae, repeated loss, as well as general biases in their positional and genic distributions. We estimate that ancestral minor intron densities were comparable to those of vertebrates, suggesting a trend of long-term stasis. Finally, three findings suggest a major role for neutral processes in minor intron evolution. First, highly similar patterns of minor and major intron evolution contrast with both functionalist and deleterious model predictions. Second, observed functional biases among minor intron-containing genes are largely explained by these genes’ greater ages. Third, no association of intron splicing with cell proliferation in a minor intron-rich fungus suggests that regulatory roles are lineage-specific and thus cannot offer a general explanation for minor splicing’s persistence. These data constitute the most comprehensive view of minor introns and their evolutionary history to date, and provide a foundation for future studies of these remarkable genetic elements.

     
    more » « less
  3. Abstract

    The oceanic igneous crust is a vast reservoir for microbial life, dominated by diverse and active bacteria, archaea, and fungi. Archaeal and bacterial viruses were previously detected in oceanic crustal fluids at the Juan de Fuca Ridge (JdFR). Here we report the discovery of two eukaryotic Nucleocytoviricota genomes from the same crustal fluids by sorting and sequencing single virions. Both genomes have a tRNATyr gene with an intron (20 bps) at the canonical position between nucleotide 37 and 38, a common feature in eukaryotic and archaeal tRNA genes with short introns (<100 bps), and fungal genes acquired through horizontal gene transfer (HGT) events. The dominance of Ascomycota fungi as the main eukaryotes in crustal fluids and the evidence for HGT point to these fungi as the putative hosts, making these the first putative fungi-Nucleocytoviricota specific association. Our study suggests active host-viral dynamics for the only eukaryotic group found in the subsurface oceanic crust and raises important questions about the impact of viral infection on the productivity and biogeochemical cycling in this ecosystem.

     
    more » « less
  4. Circular RNAs (circRNAs) are a recently discovered class of RNAs derived from protein-coding genes that have important biological and pathological roles. They are formed through backsplicing during co-transcriptional alternative splicing; however, the unified mechanism that accounts for backsplicing decisions remains unclear. Factors that regulate the transcriptional timing and spatial organization of pre-mRNA, including RNAPII kinetics, the availability of splicing factors, and features of gene architecture, have been shown to influence backsplicing decisions. Poly (ADP-ribose) polymerase I (PARP1) regulates alternative splicing through both its presence on chromatin as well as its PARylation activity. However, no studies have investigated PARP1’s possible role in regulating circRNA biogenesis. Here, we hypothesized that PARP1’s role in splicing extends to circRNA biogenesis. Our results identify many unique circRNAs in PARP1 depletion and PARylation-inhibited conditions compared to the wild type. We found that while all genes producing circRNAs share gene architecture features common to circRNA host genes, genes producing circRNAs in PARP1 knockdown conditions had longer upstream introns than downstream introns, whereas flanking introns in wild type host genes were symmetrical. Interestingly, we found that the behavior of PARP1 in regulating RNAPII pausing is distinct between these two classes of host genes. We conclude that the PARP1 pausing of RNAPII works within the context of gene architecture to regulate transcriptional kinetics, and therefore circRNA biogenesis. Furthermore, this regulation of PARP1 within host genes acts to fine tune their transcriptional output with implications in gene function. 
    more » « less
  5. Introduction

    Eukaryotic life depends on the functional elements encoded by both the nuclear genome and organellar genomes, such as those contained within the mitochondria. The content, size, and structure of the mitochondrial genome varies across organisms with potentially large implications for phenotypic variance and resulting evolutionary trajectories. Among yeasts in the subphylum Saccharomycotina, extensive differences have been observed in various species relative to the model yeastSaccharomyces cerevisiae, but mitochondrial genome sampling across many groups has been scarce, even as hundreds of nuclear genomes have become available.

    Methods

    By extracting mitochondrial assemblies from existing short-read genome sequence datasets, we have greatly expanded both the number of available genomes and the coverage across sparsely sampled clades.

    Results

    Comparison of 353 yeast mitochondrial genomes revealed that, while size and GC content were fairly consistent across species, those in the generaMetschnikowiaandSaccharomycestrended larger, while several species in the order Saccharomycetales, which includesS. cerevisiae, exhibited lower GC content. Extreme examples for both size and GC content were scattered throughout the subphylum. All mitochondrial genomes shared a core set of protein-coding genes for Complexes III, IV, and V, but they varied in the presence or absence of mitochondrially-encoded canonical Complex I genes. We traced the loss of Complex I genes to a major event in the ancestor of the orders Saccharomycetales and Saccharomycodales, but we also observed several independent losses in the orders Phaffomycetales, Pichiales, and Dipodascales. In contrast to prior hypotheses based on smaller-scale datasets, comparison of evolutionary rates in protein-coding genes showed no bias towards elevated rates among aerobically fermenting (Crabtree/Warburg-positive) yeasts. Mitochondrial introns were widely distributed, but they were highly enriched in some groups. The majority of mitochondrial introns were poorly conserved within groups, but several were shared within groups, between groups, and even across taxonomic orders, which is consistent with horizontal gene transfer, likely involving homing endonucleases acting as selfish elements.

    Discussion

    As the number of available fungal nuclear genomes continues to expand, the methods described here to retrieve mitochondrial genome sequences from these datasets will prove invaluable to ensuring that studies of fungal mitochondrial genomes keep pace with their nuclear counterparts.

     
    more » « less