Abstract We present the first long-read de novo assembly and annotation of the luna moth (Actias luna) and provide the full characterization of heavy chain fibroin (h-fibroin), a long and highly repetitive gene (>20 kb) essential in silk fiber production. There are >160,000 described species of moths and butterflies (Lepidoptera), but only within the last 5 years have we begun to recover high-quality annotated whole genomes across the order that capture h-fibroin. Using PacBio HiFi reads, we produce the first high-quality long-read reference genome for this species. The assembled genome has a length of 532 Mb, a contig N50 of 16.8 Mb, an L50 of 14 contigs, and 99.4% completeness (BUSCO). Our annotation using Bombyx mori protein and A. luna RNAseq evidence captured a total of 20,866 genes at 98.9% completeness with 10,267 functionally annotated proteins and a full-length h-fibroin annotation of 2,679 amino acid residues.
more »
« less
Genome Report: Whole Genome Sequence and Annotation of the Parasitoid Jewel Wasp Nasonia giraulti Laboratory Strain RV2X[u]
Jewel wasps in the genus of Nasonia are parasitoids with haplodiploidy sex determination, rapid development and are easy to culture in the laboratory. They are excellent models for insect genetics, genomics, epigenetics, development, and evolution. Nasonia vitripennis ( Nv ) and N. giraulti ( Ng ) are closely-related species that can be intercrossed, particularly after removal of the intracellular bacterium Wolbachia , which serve as a powerful tool to map and positionally clone morphological, behavioral, expression and methylation phenotypes. The Nv reference genome was assembled using Sanger, PacBio and Nanopore approaches and annotated with extensive RNA-seq data. In contrast, Ng genome is only available through low coverage resequencing. Therefore, de novo Ng assembly is in urgent need to advance this system. In this study, we report a high-quality Ng assembly using 10X Genomics linked-reads with 670X sequencing depth. The current assembly has a genome size of 259,040,977 bp in 3,160 scaffolds with 38.05% G-C and a 98.6% BUSCO completeness score. 97% of the RNA reads are perfectly aligned to the genome, indicating high quality in contiguity and completeness. A total of 14,777 genes are annotated in the Ng genome, and 72% of the annotated genes have a one-to-one ortholog in the Nv genome. We reported 5 million Ng-Nv SNPs which will facility mapping and population genomic studies in Nasonia . In addition, 42 Ng -specific genes were identified by comparing with Nv genome and annotation. This is the first de novo assembly for this important species in the Nasonia model system, providing a useful new genomic toolkit.
more »
« less
- Award ID(s):
- 1928770
- PAR ID:
- 10198591
- Date Published:
- Journal Name:
- G3: Genes|Genomes|Genetics
- Volume:
- 10
- Issue:
- 8
- ISSN:
- 2160-1836
- Page Range / eLocation ID:
- 2565 to 2572
- Format(s):
- Medium: X
- Sponsoring Org:
- National Science Foundation
More Like this
-
-
Eyre-Walker, Adam (Ed.)The coppery titi monkey (Plecturocebus cupreus) is an emerging nonhuman primate model system for behavioral and neurobiological research. At the same time, the almost entire absence of genomic resources for the species has hampered insights into the genetic underpinnings of the phenotypic traits of interest. To facilitate future genotype-to-phenotype studies, we here present a high-quality, fully annotated de novo genome assembly for the species with chromosome-length scaffolds spanning the autosomes and chromosome X (scaffold N50 = 130.8 Mb), constructed using data obtained from several orthologous short- and long-read sequencing and scaffolding techniques. With a base-level accuracy of ∼99.99% in chromosome-length scaffolds as well as benchmarking universal single-copy ortholog and k-mer completeness scores of >99.0% and 95.1% at the genome level, this assembly represents one of the most complete Pitheciidae genomes to date, making it an invaluable resource for comparative evolutionary genomics research to improve our understanding of lineage-specific changes underlying adaptive traits as well as deleterious mutations associated with disease.more » « less
-
Abstract Acarospora socialis, the bright cobblestone lichen, is commonly found in southwestern North America. This charismatic yellow lichen is a species of key ecological significance as it is often a pioneer species in new environments. Despite their ecological importance virtually no research has been conducted on the genomics of A. socialis. To address this, we used long-read sequencing to generate the first high-quality draft genome of A. socialis. Lichen thallus tissue was collected from Pinkham Canyon in Joshua Tree National Park, California and deposited in the UC Riverside herbarium under accession #295874. The de novo assembly of the mycobiont partner of the lichen was generated from Pacific Biosciences HiFi long reads and Dovetail Omni-C chromatin capture data. After removing algal and bacterial contigs, the fungal genome was approximately 31.2 Mb consisting of 38 scaffolds with contig and scaffold N50 of 2.4 Mb. The BUSCO completeness score of the assembled genome was 97.5% using the Ascomycota gene set. Information on the genome of A. socialis is important for California conservation purposes given that this lichen is threatened in some places locally by wildfires due to climate change. This reference genome will be used for understanding the genetic diversity, population genomics, and comparative genomics of A. socialis species. Genomic resources for this species will support population and landscape genomics investigations, exploring the use of A. socialis as a bioindicator species for climate change, and in studies of adaptation by comparing populations that occur across aridity gradients in California.more » « less
-
Abstract The aye-aye (Daubentonia madagascariensis) is the only extant member of the Daubentoniidae primate family. Although several reference genomes exist for this endangered strepsirrhine primate, the predominant usage of short-read sequencing has resulted in limited assembly contiguity and completeness, and no protein-coding gene annotations have yet been released. Here, we present a novel, fully annotated, chromosome-level hybrid de novo assembly for the species based on a combination of Oxford Nanopore Technologies long reads and Illumina short reads and scaffolded using genome-wide chromatin interaction data—a community resource that will improve future conservation efforts as well as primate comparative analyses.more » « less
-
null (Ed.)Abstract The Andean bear is the only extant member of the Tremarctine subfamily and the only extant ursid species to inhabit South America. Here, we present an annotated de novo assembly of a nuclear genome from a captive-born female Andean bear, Mischief, generated using a combination of short and long DNA and RNA reads. Our final assembly has a length of 2.23 Gb, and a scaffold N50 of 21.12 Mb, contig N50 of 23.5 kb, and BUSCO score of 88%. The Andean bear genome will be a useful resource for exploring the complex phylogenetic history of extinct and extant bear species and for future population genetics studies of Andean bears.more » « less
An official website of the United States government

