skip to main content

Title: High-Quality Genome Assembly and Comprehensive Transcriptome of the Painted Lady Butterfly Vanessa cardui
Abstract The painted lady butterfly, Vanessa cardui, has the longest migration routes, the widest hostplant diversity, and one of the most complex wing patterns of any insect. Due to minimal culturing requirements, easily characterized wing pattern elements, and technical feasibility of CRISPR/Cas9 genome editing, V. cardui is emerging as a functional genomics model for diverse research programs. Here, we report a high-quality, annotated genome assembly of the V. cardui genome, generated using 84× coverage of PacBio long-read data, which we assembled into 205 contigs with a total length of 425.4 Mb (N50 = 10.3 Mb). The genome was very complete (single-copy complete Benchmarking Universal Single-Copy Orthologs [BUSCO] 97%), with contigs assembled into presumptive chromosomes using synteny analyses. Our annotation used embryonic, larval, and pupal transcriptomes, and 20 transcriptomes across five different wing developmental stages. Gene annotations showed a high level of accuracy and completeness, with 14,437 predicted protein-coding genes. This annotated genome assembly constitutes an important resource for diverse functional genomic studies ranging from the developmental genetic basis of butterfly color pattern, to coevolution with diverse hostplants.  more » « less
Award ID(s):
Author(s) / Creator(s):
; ; ;
Lavrov, Dennis
Date Published:
Journal Name:
Genome Biology and Evolution
Medium: X
Sponsoring Org:
National Science Foundation
More Like this
  1. Abstract Comparisons of high-quality, reference butterfly, and moth genomes have been instrumental to advancing our understanding of how hybridization, and natural selection drive genomic change during the origin of new species and novel traits. Here, we present a genome assembly of the Southern Dogface butterfly, Zerene cesonia (Pieridae) whose brilliant wing colorations have been implicated in developmental plasticity, hybridization, sexual selection, and speciation. We assembled 266,407,278 bp of the Z. cesonia genome, which accounts for 98.3% of the estimated 271 Mb genome size. Using a hybrid approach involving Chicago libraries with Hi-Rise assembly and a diploid Meraculous assembly, the final haploid genome was assembled. In the final assembly, nearly all autosomes and the Z chromosome were assembled into single scaffolds. The largest 29 scaffolds accounted for 91.4% of the genome assembly, with the remaining ∼8% distributed among another 247 scaffolds and overall N50 of 9.2 Mb. Tissue-specific RNA-seq informed annotations identified 16,442 protein-coding genes, which included 93.2% of the arthropod Benchmarking Universal Single-Copy Orthologs (BUSCO). The Z. cesonia genome assembly had ∼9% identified as repetitive elements, with a transposable element landscape rich in helitrons. Similar to other Lepidoptera genomes, Z. cesonia showed a high conservation of chromosomal synteny. The Z. cesonia assembly provides a high-quality reference for studies of chromosomal arrangements in the Pierid family, as well as for population, phylo, and functional genomic studies of adaptation and speciation. 
    more » « less
  2. Abstract

    Rosalia funebris (RFUNE; Cerambycidae), the banded alder borer, is a longhorn beetle whose larvae feed on the wood of various economically and ecologically significant trees in western North America. Adults are short-lived and not known to consume plant material substantially. We sequenced, assembled, and annotated the RFUNE genome using HiFi and RNASeq data. We documented genome architecture and gene content, focusing on genes putatively involved in plant feeding (phytophagy). Comparisons were made to the well-studied genome of the Asian longhorned beetle (AGLAB; Anoplophora glabripennis) and other Cerambycidae. The 814 Mb RFUNE genome assembly was distributed across 42 contigs, with an N50 of 30.18 Mb. Repetitive sequences comprised 60.27% of the genome, and 99.0% of expected single-copy orthologous genes were fully assembled. We identified 12,657 genes, fewer than in the four other species studied, and 46.4% fewer than for Aromia moschata (same subfamily as RFUNE). Of the 7,258 orthogroups shared between RFUNE and AGLAB, 1,461 had more copies in AGLAB and 1,023 had more copies in RFUNE. We identified 240 genes in RFUNE that putatively arose via horizontal transfer events. The RFUNE genome encoded substantially fewer putative plant cell wall degrading enzymes than AGLAB, which may relate to the longer-lived plant-feeding adults of the latter species. The RFUNE genome provides new insights into cerambycid genome architecture and gene content and provides a new vantage point from which to study the evolution and genomic basis of phytophagy in beetles.

    more » « less
  3. Abstract

    Automeris moths are a morphologically diverse group with 145 described species that have a geographic range that spans from the New World temperate zone to the Neotropics. Many Automeris have elaborate hindwing eyespots that are thought to deter or disrupt the attack of potential predators, allowing the moth time to escape. The Io moth (Automeris io), known for its striking eyespots, is a well-studied species within the genus and is an emerging model system to study the evolution of deimatism. Existing research on the eyespot pattern development will be augmented by genomic resources that allow experimental manipulation of this emerging model. Here, we present a high-quality, PacBio HiFi genome assembly for Io moth to aid existing research on the molecular development of eyespots and future research on other deimatic traits. This 490 Mb assembly is highly contiguous (N50 = 15.78 mbs) and complete (benchmarking universal single-copy orthologs = 98.4%). Additionally, we were able to recover orthologs of genes previously identified as being involved in wing pattern formation and movement.

    more » « less
  4. null (Ed.)
    Abstract Background The western flower thrips, Frankliniella occidentalis (Pergande), is a globally invasive pest and plant virus vector on a wide array of food, fiber, and ornamental crops. The underlying genetic mechanisms of the processes governing thrips pest and vector biology, feeding behaviors, ecology, and insecticide resistance are largely unknown. To address this gap, we present the F. occidentalis draft genome assembly and official gene set. Results We report on the first genome sequence for any member of the insect order Thysanoptera. Benchmarking Universal Single-Copy Ortholog (BUSCO) assessments of the genome assembly (size = 415.8 Mb, scaffold N50 = 948.9 kb) revealed a relatively complete and well-annotated assembly in comparison to other insect genomes. The genome is unusually GC-rich (50%) compared to other insect genomes to date. The official gene set (OGS v1.0) contains 16,859 genes, of which ~ 10% were manually verified and corrected by our consortium. We focused on manual annotation, phylogenetic, and expression evidence analyses for gene sets centered on primary themes in the life histories and activities of plant-colonizing insects. Highlights include the following: (1) divergent clades and large expansions in genes associated with environmental sensing (chemosensory receptors) and detoxification (CYP4, CYP6, and CCE enzymes) of substances encountered in agricultural environments; (2) a comprehensive set of salivary gland genes supported by enriched expression; (3) apparent absence of members of the IMD innate immune defense pathway; and (4) developmental- and sex-specific expression analyses of genes associated with progression from larvae to adulthood through neometaboly, a distinct form of maturation differing from either incomplete or complete metamorphosis in the Insecta. Conclusions Analysis of the F. occidentalis genome offers insights into the polyphagous behavior of this insect pest that finds, colonizes, and survives on a widely diverse array of plants. The genomic resources presented here enable a more complete analysis of insect evolution and biology, providing a missing taxon for contemporary insect genomics-based analyses. Our study also offers a genomic benchmark for molecular and evolutionary investigations of other Thysanoptera species. 
    more » « less
  5. Genome assembly can be challenging for species that are characterized by high amounts of polymorphism, heterozygosity, and large effective population sizes. High levels of heterozygosity can result in genome mis-assemblies and a larger than expected genome size due to the haplotig versions of a single locus being assembled as separate loci. Here, we describe the first chromosome-level genome for the eastern oyster, Crassostrea virginica. Publicly released and annotated in 2017, the assembly has a scaffold N50 of 54 mb and is over 97.3% complete based on BUSCO analysis. The genome assembly for the eastern oyster is a critical resource for foundational research into molluscan adaptation to a changing environment and for selective breeding for the aquaculture industry. Subsequent resequencing data suggested the presence of haplotigs in the original assembly, and we developed a post hoc method to break up chimeric contigs and mask haplotigs in published heterozygous genomes and evaluated improvements to the accuracy of downstream analysis. Masking haplotigs had a large impact on SNP discovery and estimates of nucleotide diversity and had more subtle and nuanced effects on estimates of heterozygosity, population structure analysis, and outlier detection. We show that haplotig masking can be a powerful tool for improving genomic inference, and we present an open, reproducible resource for the masking of haplotigs in any published genome. 
    more » « less