Despite insertions and deletions being the most common structural variants (SVs) found across genomes, not much is known about how much these SVs vary within populations and between closely related species, nor their significance in evolution. To address these questions, we characterized the evolution of indel SVs using genome assemblies of three closely related Heliconius butterfly species. Over the relatively short evolutionary timescales investigated, up to 18.0% of the genome was composed of indels between two haplotypes of an individual Heliconius charithonia butterfly and up to 62.7% included lineage-specific SVs between the genomes of the most distant species (11 Mya). Lineage-specific sequences were mostly characterized as transposable elements (TEs) inserted at random throughout the genome and their overall distribution was similarly affected by linked selection as single nucleotide substitutions. Using chromatin accessibility profiles (i.e., ATAC-seq) of head tissue in caterpillars to identify sequences with potential cis -regulatory function, we found that out of the 31,066 identified differences in chromatin accessibility between species, 30.4% were within lineage-specific SVs and 9.4% were characterized as TE insertions. These TE insertions were localized closer to gene transcription start sites than expected at random and were enriched for sites with significant resemblance to several transcription factor binding sites with known function in neuron development in Drosophila . We also identified 24 TE insertions with head-specific chromatin accessibility. Our results show high rates of structural genome evolution that were previously overlooked in comparative genomic studies and suggest a high potential for structural variation to serve as raw material for adaptive evolution.
more »
« less
Evolutionary dynamics of genome size and content during the adaptive radiation of Heliconiini butterflies
Abstract Heliconius butterflies, a speciose genus of Müllerian mimics, represent a classic example of an adaptive radiation that includes a range of derived dietary, life history, physiological and neural traits. However, key lineages within the genus, and across the broader Heliconiini tribe, lack genomic resources, limiting our understanding of how adaptive and neutral processes shaped genome evolution during their radiation. Here, we generate highly contiguous genome assemblies for nine Heliconiini, 29 additional reference-assembled genomes, and improve 10 existing assemblies. Altogether, we provide a dataset of annotated genomes for a total of 63 species, including 58 species within the Heliconiini tribe. We use this extensive dataset to generate a robust and dated heliconiine phylogeny, describe major patterns of introgression, explore the evolution of genome architecture, and the genomic basis of key innovations in this enigmatic group, including an assessment of the evolution of putative regulatory regions at the Heliconius stem. Our work illustrates how the increased resolution provided by such dense genomic sampling improves our power to generate and test gene-phenotype hypotheses, and precisely characterize how genomes evolve.
more »
« less
- PAR ID:
- 10463993
- Date Published:
- Journal Name:
- Nature Communications
- Volume:
- 14
- Issue:
- 1
- ISSN:
- 2041-1723
- Format(s):
- Medium: X
- Sponsoring Org:
- National Science Foundation
More Like this
-
-
Despite insertions and deletions being the most common structural variants (SVs) found across genomes, not much is known about how much these SVs vary within populations and between closely related species, nor their significance in evolution. To address these questions, we characterized the evolution of indel SVs using genome assemblies of three closely related Heliconius butterfly species. Over the relatively short evolutionary timescales investigated, up to 18.0% of the genome was composed of indels between two haplotypes of an individual H. charithonia butterfly and up to 62.7% included lineage-specific SVs between the genomes of the most distant species (11 Mya). Lineage-specific sequences were mostly characterized as transposable elements (TEs) inserted at random throughout the genome and their overall distribution was similarly affected by linked selection as single nucleotide substitutions. Using chromatin accessibility profiles (i.e., ATAC-seq) of head tissue in caterpillars to identify sequences with potential cis-regulatory function, we found that out of the 31,066 identified differences in chromatin accessibility between species, 30.4% were within lineage-specific SVs and 9.4% were characterized as TE insertions. These TE insertions were localized closer to gene transcription start sites than expected at random and were enriched for several transcription factor binding site candidates with known function in neuron development in Drosophila. We also identified 24 TE insertions with head-specific chromatin accessibility. Our results show high rates of structural genome evolution that were previously overlooked in comparative genomic studies and suggest a high potential for structural variation to serve as raw material for adaptive evolution.more » « less
-
Abstract The plant genus Bidens (Asteraceae or Compositae; Coreopsidae) is a species-rich and circumglobally distributed taxon. The 19 hexaploid species endemic to the Hawaiian Islands are considered an iconic example of adaptive radiation, of which many are imperiled and of high conservation concern. Until now, no genomic resources were available for this genus, which may serve as a model system for understanding the evolutionary genomics of explosive plant diversification. Here, we present a high-quality reference genome for the Hawaiʻi Island endemic species B. hawaiensis A. Gray reconstructed from long-read, high-fidelity sequences generated on a Pacific Biosciences Sequel II System. The haplotype-aware, draft genome assembly consisted of ~6.67 Giga bases (Gb), close to the holoploid genome size estimate of 7.56 Gb (±0.44 SD) determined by flow cytometry. After removal of alternate haplotigs and contaminant filtering, the consensus haploid reference genome was comprised of 15 904 contigs containing ~3.48 Gb, with a contig N50 value of 422 594. The high interspersed repeat content of the genome, approximately 74%, along with hexaploid status, contributed to assembly fragmentation. Both the haplotype-aware and consensus haploid assemblies recovered >96% of Benchmarking Universal Single-Copy Orthologs. Yet, the removal of alternate haplotigs did not substantially reduce the proportion of duplicated benchmarking genes (~79% vs. ~68%). This reference genome will support future work on the speciation process during adaptive radiation, including resolving evolutionary relationships, determining the genomic basis of trait evolution, and supporting ongoing conservation efforts.more » « less
-
null (Ed.)Abstract Background Heliconius butterflies are widely distributed across the Neotropics and have evolved a stunning array of wing color patterns that mediate Müllerian mimicry and mating behavior. Their rapid radiation has been strongly influenced by hybridization, which has created new species and allowed sharing of color patterning alleles between mimetic species pairs. While these processes have frequently been observed in widespread species with contiguous distributions, many Heliconius species inhabit patchy or rare habitats that may strongly influence the origin and spread of species and color patterns. Here, we assess the effects of historical population fragmentation and unique biology on the origins, genetic health, and color pattern evolution of two rare and sparsely distributed Brazilian butterflies, Heliconius hermathena and Heliconius nattereri . Results We assembled genomes and re-sequenced whole genomes of eight H. nattereri and 71 H. hermathena individuals. These species harbor little genetic diversity, skewed site frequency spectra, and high deleterious mutation loads consistent with recent population bottlenecks. Heliconius hermathena consists of discrete, strongly isolated populations that likely arose from a single population that dispersed after the last glacial maximum. Despite having a unique color pattern combination that suggested a hybrid origin, we found no genome-wide evidence that H. hermathena is a hybrid species. However, H. hermathena mimicry evolved via introgression, from co-mimetic Heliconius erato , of a small genomic region upstream of the color patterning gene cortex . Conclusions Heliconius hermathena and H. nattereri population fragmentation, potentially driven by historical climate change and recent deforestation, has significantly reduced the genetic health of these rare species. Our results contribute to a growing body of evidence that introgression of color patterning alleles between co-mimetic species appears to be a general feature of Heliconius evolution.more » « less
-
Wheat, Christopher (Ed.)Abstract Paper wasps are a model system for the study of social evolution due to a high degree of inter- and intraspecific variation in cooperation, aggression, and visual signals of social status. Increasing the taxonomic coverage of genomic resources for this diverse clade will aid comparative genomic approaches for testing predictions about the molecular basis of social evolution. Here, we provide draft genome assemblies for two well-studied species of paper wasps, Polistes exclamans and Mischocyttarus mexicanus. The P. exclamans genome assembly is 221.5 Mb in length with a scaffold N50 of 4.11 Mb. The M. mexicanus genome assembly is 227 Mb in length with a scaffold N50 of 1.1 Mb. Genomes have low repeat content (9.54–10.75%) and low GC content (32.06–32.4%), typical of other social hymenopteran genomes. The DNA methyltransferase gene, Dnmt3 , was lost early in the evolution of Polistinae. We identified a second independent loss of Dnmt3 within hornets (genus: Vespa).more » « less
An official website of the United States government

