De novo genome assembly and functional annotation for Fusarium langsethiae

Mapping Intimacies ◽

10.1101/2021.04.13.439621 ◽

2021 ◽

Author(s):

Ya Zuo ◽

Carol Verheecke-Vaessen ◽

Corentin Molitor ◽

Angel Medina ◽

Naresh Magan ◽

...

Keyword(s):

Genome Assembly ◽

De Novo ◽

Animal Health ◽

Data Availability ◽

Effective Control ◽

Local Alignment ◽

De Novo Genome Assembly ◽

Genome Database ◽

A Genome ◽

Fusarium Langsethiae

AbstractMotivationFusarium langsethiae is a T-2 and HT-2 mycotoxins producing Fusarium species firstly characterised in 2004. It is commonly isolated from oats in Northern Europe. T-2 and HT-2 mycotoxins exhibit immunological and haemotological effects in animal health mainly through inhibition of protein, RNA and DNA synthesis. The development of a high-quality and comprehensively annotated assembly for this species is therefore essential in providing the molecular understanding and the mechanism of T-2 and HT-2 biosynthesis in F. langsethiae to help develop effective control strategies.ResultsThe F. langsethiae assembly was produced using PacBio long reads, which were then assembled independently using Canu, SMARTdenovo and Flye; producing a genome assembly total length of 59Mb and N50 of 3.51Mb. A total of 19,336 coding genes were identified using RNA-Seq informed ab-initio gene prediction. Finally, predicting genes were annotated using the basic local alignment search tool (BLAST) against the NCBI non-redundant (NR) genome database and protein hits were annotated using InterProScan. Genes with blast hits were functionally annotated with Gene [email protected] availabilityRaw sequence reads and assembled genome can be downloaded from: GenBank under the accession JAFFKB000000000

Download Full-text

A high-quality de novo genome assembly from a single parasitoid wasp

10.1101/2020.07.13.200725 ◽

2020 ◽

Cited By ~ 1

Author(s):

Xinhai Ye ◽

Yi Yang ◽

Zhaoyang Tian ◽

Le Xu ◽

Kaili Yu ◽

...

Keyword(s):

Genome Assembly ◽

De Novo ◽

Parasitoid Wasp ◽

Genome Comparison ◽

Single Individual ◽

High Quality ◽

De Novo Genome Assembly ◽

Low Input ◽

A Genome ◽

High Level

AbstractSequencing and assembling a genome with a single individual have several advantages, such as lower heterozygosity and easier sample preparation. However, the amount of genomic DNA of some small sized organisms might not meet the standard DNA input requirement for current sequencing pipelines. Although few studies sequenced a single small insect with about 100 ng DNA as input, it may still be challenging for many small organisms to obtain such amount of DNA from a single individual. Here, we use 20 ng DNA as input, and present a high-quality genome assembly for a single haploid male parasitoid wasp (Habrobracon hebetor) using Nanopore and Illumina. Because of the low input DNA, a whole genome amplification (WGA) method is used before sequencing. The assembled genome size is 131.6 Mb with a contig N50 of 1.63 Mb. A total of 99% Benchmarking Universal Single-Copy Orthologs are detected, suggesting the high level of completeness of the genome assembly. Genome comparison between H. hebetor and its relative Bracon brevicornis shows a high-level genome synteny, indicating the genome of H. hebetor is highly accurate and contiguous. Our study provides an example for de novo assembling a genome from ultra-low input DNA, and will be used for sequencing projects of small sized species and rare samples, haploid genomics as well as population genetics of small sized species.

Download Full-text

De novo genome assembly of the land snail Candidula unifasciata (Mollusca: Gastropoda)

G3 Genes|Genome|Genetics ◽

10.1093/g3journal/jkab180 ◽

2021 ◽

Author(s):

Luis J Chueca ◽

Tilman Schell ◽

Markus Pfenninger

Keyword(s):

Genome Assembly ◽

De Novo ◽

Draft Genome ◽

Land Snail ◽

Land Snails ◽

Agricultural Pests ◽

De Novo Genome Assembly ◽

Draft Genome Assembly ◽

Western Palearctic ◽

A Genome

Abstract Among all molluscs, land snails are a scientifically and economically interesting group comprising edible species, alien species and agricultural pests. Yet, despite their high diversity, the number of genome drafts publicly available is still scarce. Here, we present the draft genome assembly of the land snail Candidula unifasciata, a widely distributed species along central Europe, belonging to the Geomitridae family, a highly diversified taxon in the Western-Palearctic region. We performed whole genome sequencing, assembly and annotation of an adult specimen based on PacBio and Oxford Nanopore long read sequences as well as Illumina data. A genome draft of about 1.29 Gb was generated with a N50 length of 246 kb. More than 60% of the assembled genome was identified as repetitive elements. 22,464 protein-coding genes were identified in the genome, of which 62.27% were functionally annotated. This is the first assembled and annotated genome for a geometrid snail and will serve as reference for further evolutionary, genomic and population genetic studies of this important and interesting group.

Download Full-text

De novo genome assembly of the land snail Candidula unifasciata (Mollusca: Gastropoda)

10.1101/2021.01.23.427926 ◽

2021 ◽

Author(s):

Luis J. Chueca ◽

Tilman Schell ◽

Markus Pfenninger

Keyword(s):

Genome Assembly ◽

De Novo ◽

Draft Genome ◽

Land Snail ◽

Land Snails ◽

Agricultural Pests ◽

De Novo Genome Assembly ◽

Draft Genome Assembly ◽

Western Palearctic ◽

A Genome

AbstractAmong all molluscs, land snails are an economically and scientifically interesting group comprising edible species, alien species and agricultural pests. Yet, despite its high diversity, the number of whole genomes publicly available is still scarce. Here, we present the draft genome assembly of the land snail Candidula unifasciata, a widely distributed species along central Europe, which belongs to Geomitridae family, a group highly diversified in the Western-Palearctic region. We performed a whole genome sequencing, assembly and annotation of an adult specimen based on PacBio and Oxford Nanopore long read sequences as well as Illumina data. A genome of about 1.29 Gb was generated with a N50 length of 246 kb. More than 60% of the assembled genome was identified as repetitive elements, and 22,464 protein-coding genes were identified in the genome, where the 62.27% were functionally annotated. This is the first assembled and annotated genome for a geometrid snail and will serve as reference for further evolutionary, genomic and population genetic studies of this important and interesting group.

Download Full-text

Chromatin Architectures Are Associated with Response to Dark Treatment in the Oil Crop Sesamum indicum, Based on a High-Quality Genome Assembly

Plant and Cell Physiology ◽

10.1093/pcp/pcaa026 ◽

2020 ◽

Vol 61 (5) ◽

pp. 978-987

Author(s):

Chaoqiong Li ◽

Xiaoli Li ◽

Hongzhan Liu ◽

Xueqin Wang ◽

Weifeng Li ◽

...

Keyword(s):

Gene Transcription ◽

Genome Assembly ◽

De Novo ◽

Hierarchical Structures ◽

Normal Growth ◽

High Quality ◽

Dark Treatment ◽

De Novo Genome Assembly ◽

Chromosome Conformation ◽

A Genome

Abstract Eukaryotic chromatin is tightly packed into hierarchical structures, allowing appropriate gene transcription in response to environmental and developmental cues. Here, we provide a chromosome-scale de novo genome assembly of sesame with a total length of 292.3 Mb and a scaffold N50 of 20.5 Mb, containing estimated 28,406 coding genes using Pacific Biosciences long reads combined with a genome-wide chromosome conformation capture (Hi-C) approach. Based on this high-quality reference genome, we detected changes in chromatin architectures between normal growth and dark-treated sesame seedlings. Gene expression level was significantly higher in ‘A’ compartment and topologically associated domain (TAD) boundary regions than in ‘B’ compartment and TAD interior regions, which is coincident with the enrichment of H4K3me3 modification in these regions. Moreover, differentially expressed genes (DEGs) induced by dark treated were enriched in the changed TAD-related regions and genomic differential contact regions. Gene Ontology (GO) enrichment analysis of DEGs showed that genes related to ‘response to stress’ and ‘photosynthesis’ functional categories were enriched, which corresponds to dark treatment. These results suggested that chromatin organization is associated with gene transcription in response to dark treatment in sesame. Our results will facilitate the understanding of regulatory mechanisms in response to environmental cues in plants.

Download Full-text

De Novo Genome Assembly of Limpet Bathyacmaea lactea (Gastropoda: Pectinodontidae): The First Reference Genome of a Deep-Sea Gastropod Endemic to Cold Seeps

Genome Biology and Evolution ◽

10.1093/gbe/evaa100 ◽

2020 ◽

Vol 12 (6) ◽

pp. 905-910 ◽

Cited By ~ 2

Author(s):

Ruoyu Liu ◽

Kun Wang ◽

Jun Liu ◽

Wenjie Xu ◽

Yang Zhou ◽

...

Keyword(s):

Deep Sea ◽

Metal Ion ◽

De Novo ◽

Demographic History ◽

Gene Families ◽

Phylogenetic Position ◽

Cold Seeps ◽

Nitrogen And Phosphorus ◽

De Novo Genome Assembly ◽

A Genome

Abstract Cold seeps, characterized by the methane, hydrogen sulfide, and other hydrocarbon chemicals, foster one of the most widespread chemosynthetic ecosystems in deep sea that are densely populated by specialized benthos. However, scarce genomic resources severely limit our knowledge about the origin and adaptation of life in this unique ecosystem. Here, we present a genome of a deep-sea limpet Bathyacmaea lactea, a common species associated with the dominant mussel beds in cold seeps. We yielded 54.6 gigabases (Gb) of Nanopore reads and 77.9-Gb BGI-seq raw reads, respectively. Assembly harvested a 754.3-Mb genome for B. lactea, with 3,720 contigs and a contig N50 of 1.57 Mb, covering 94.3% of metazoan Benchmarking Universal Single-Copy Orthologs. In total, 23,574 protein-coding genes and 463.4 Mb of repetitive elements were identified. We analyzed the phylogenetic position, substitution rate, demographic history, and TE activity of B. lactea. We also identified 80 expanded gene families and 87 rapidly evolving Gene Ontology categories in the B. lactea genome. Many of these genes were associated with heterocyclic compound metabolism, membrane-bounded organelle, metal ion binding, and nitrogen and phosphorus metabolism. The high-quality assembly and in-depth characterization suggest the B. lactea genome will serve as an essential resource for understanding the origin and adaptation of life in the cold seeps.

Download Full-text

Improved hybrid de novo genome assembly of domesticated apple (Malus x domestica)

GigaScience ◽

10.1186/s13742-016-0139-0 ◽

2016 ◽

Vol 5 (1) ◽

Cited By ~ 28

Author(s):

Xuewei Li ◽

Ling Kui ◽

Jing Zhang ◽

Yinpeng Xie ◽

Liping Wang ◽

...

Keyword(s):

Genome Assembly ◽

De Novo ◽

Malus X Domestica ◽

De Novo Genome Assembly

Download Full-text

Meraculous: De Novo Genome Assembly with Short Paired-End Reads

PLoS ONE ◽

10.1371/journal.pone.0023501 ◽

2011 ◽

Vol 6 (8) ◽

pp. e23501 ◽

Cited By ~ 107

Author(s):

Jarrod A. Chapman ◽

Isaac Ho ◽

Sirisha Sunkara ◽

Shujun Luo ◽

Gary P. Schroth ◽

...

Keyword(s):

Genome Assembly ◽

De Novo ◽

De Novo Genome Assembly

Download Full-text

Ultra Efficient Acceleration for De Novo Genome Assembly via Near-Memory Computing

10.1109/pact52795.2021.00022 ◽

2021 ◽

Author(s):

Minxuan Zhou ◽

Lingxi Wu ◽

Muzhou Li ◽

Niema Moshiri ◽

Kevin Skadron ◽

...

Keyword(s):

Genome Assembly ◽

De Novo ◽

De Novo Genome Assembly

Download Full-text

De novo Genome Assembly from Next-Generation Sequencing (NGS) Reads

Next-Generation Sequencing Data Analysis ◽

10.1201/b19532-11 ◽

2016 ◽

pp. 144-155

Keyword(s):

Next Generation Sequencing ◽

Genome Assembly ◽

De Novo ◽

Next Generation ◽

De Novo Genome Assembly ◽

Next Generation Sequencing Ngs ◽

Generation Sequencing

Download Full-text

Optimizing de novo genome assembly from PCR-amplified metagenomes

PeerJ ◽

10.7717/peerj.6902 ◽

2019 ◽

Vol 7 ◽

pp. e6902 ◽

Cited By ~ 9

Author(s):

Simon Roux ◽

Gareth Trubl ◽

Danielle Goudeau ◽

Nandita Nath ◽

Estelle Couradeau ◽

...

Keyword(s):

Genome Assembly ◽

De Novo ◽

Pcr Amplification ◽

Error Rates ◽

De Novo Genome Assembly ◽

Low Input ◽

Assembly Algorithm ◽

Coverage Bias ◽

Size Number ◽

Assembly Pipeline

Background Metagenomics has transformed our understanding of microbial diversity across ecosystems, with recent advances enabling de novo assembly of genomes from metagenomes. These metagenome-assembled genomes are critical to provide ecological, evolutionary, and metabolic context for all the microbes and viruses yet to be cultivated. Metagenomes can now be generated from nanogram to subnanogram amounts of DNA. However, these libraries require several rounds of PCR amplification before sequencing, and recent data suggest these typically yield smaller and more fragmented assemblies than regular metagenomes. Methods Here we evaluate de novo assembly methods of 169 PCR-amplified metagenomes, including 25 for which an unamplified counterpart is available, to optimize specific assembly approaches for PCR-amplified libraries. We first evaluated coverage bias by mapping reads from PCR-amplified metagenomes onto reference contigs obtained from unamplified metagenomes of the same samples. Then, we compared different assembly pipelines in terms of assembly size (number of bp in contigs ≥ 10 kb) and error rates to evaluate which are the best suited for PCR-amplified metagenomes. Results Read mapping analyses revealed that the depth of coverage within individual genomes is significantly more uneven in PCR-amplified datasets versus unamplified metagenomes, with regions of high depth of coverage enriched in short inserts. This enrichment scales with the number of PCR cycles performed, and is presumably due to preferential amplification of short inserts. Standard assembly pipelines are confounded by this type of coverage unevenness, so we evaluated other assembly options to mitigate these issues. We found that a pipeline combining read deduplication and an assembly algorithm originally designed to recover genomes from libraries generated after whole genome amplification (single-cell SPAdes) frequently improved assembly of contigs ≥10 kb by 10 to 100-fold for low input metagenomes. Conclusions PCR-amplified metagenomes have enabled scientists to explore communities traditionally challenging to describe, including some with extremely low biomass or from which DNA is particularly difficult to extract. Here we show that a modified assembly pipeline can lead to an improved de novo genome assembly from PCR-amplified datasets, and enables a better genome recovery from low input metagenomes.

Download Full-text