High-quality Archives - Page 13 of 30

April 21, 2020

Genomic analysis of Marinobacter sp. NP-4 and NP-6 isolated from the deep-sea oceanic crust on the western flank of the Mid-Atlantic Ridge

Two Marinobacter sp. NP-4 and NP-6 were isolated from a deep oceanic basaltic crust at North Pond, located at the western flank of the Mid-Atlantic Ridge. These two strains are capable of using multiple carbon sources such as acetate, succinate, glucose and sucrose while take oxygen as a primary electron acceptor. The strain NP-4 is also able to grow anaerobically under 20?MPa, with nitrate as the electron acceptor, thus represents a piezotolerant. To explore the metabolic potentials of Marinobacter sp. NP-4 and NP-6, the complete genome of NP-4 and close-to-complete genome of NP-6 were sequenced. The genome of NP-4 contains one chromosome and two plasmids with the size of 4.6?Mb in total, and with average GC content of 57.0%. The genome of NP-6 is 4.5?Mb and consists of 6 scaffolds, with an average GC content of 57.1%. Complete glycolysis, citrate cycle and aromatics compounds degradation pathways are identified in genomes of these two strains, suggesting that they possess a heterotrophic life style. Additionally, one plasmid of NP-4 contains genes for alkane degradation, phosphonate ABC transporter and cation efflux system, enabling NP-4 extra surviving abilities. In total, genomic information of these two strains provide insights into the physiological features and adaptation strategies of Marinobacter spp. in the deep oceanic crust biosphere.

April 21, 2020

Variant Phasing and Haplotypic Expression from Single-molecule Long-read Sequencing in Maize

Haplotype phasing of genetic variants is important for interpretation of the maize genome, population genetic analysis, and functional genomic analysis of allelic activity. Accordingly, accurate methods for phasing full-length isoforms are essential for functional genomics study. In this study, we performed an isoform-level phasing study in maize, using two inbred lines and their reciprocal crosses, based on single-molecule full-length cDNA sequencing. To phase and analyze full-length transcripts between hybrids and parents, we developed a tool called IsoPhase. Using this tool, we validated the majority of SNPs called against matching short read data and identified cases of allele-specific, gene-level, and isoform-level expression. Our results revealed that maize parental and hybrid lines exhibit different splicing activities. After phasing 6,847 genes in two reciprocal hybrids using embryo, endosperm and root tissues, we annotated the SNPs and identified large-effect genes. In addition, based on single-molecule sequencing, we identified parent-of-origin isoforms in maize hybrids, different novel isoforms between maize parent and hybrid lines, and imprinted genes from different tissues. Finally, we characterized variation in cis- and trans-regulatory effects. Our study provides measures of haplotypic expression that could increase power and accuracy in studies of allelic expression.

April 21, 2020

Genome sequence resource for Ilyonectria mors-panacis, causing rusty root rot of Panax notoginseng.

Ilyonectria mors-panacis is a serious disease hampering the production of Panax notoginseng, an important Chinese medicinal herb, widely used for its anti-inflammatory, anti-fatigue, hepato-protective, and coronary heart disease prevention effects. Here, we report the first Illumina-Pacbio hybrid sequenced draft genome assembly of I. mors-panacis strain G3B and its annotation. The availability of this genome sequence not only represents an important tool toward understanding the genetics behind the infection mechanism of I. mors-panacis strain G3B but also will help illuminate the complexities of the taxonomy of this species.

April 21, 2020

Complete genome sequences of pooled genomic DNA from 10 marine bacteria using PacBio long-read sequencing.

High-quality, completed genomes are important to understand the functions of marine bacteria. PacBio sequencing technology provides a powerful way to obtain high-quality completed genomes. However individual library production is currently still costly, limiting the utility of the PacBio system for high-throughput genomics. Here we investigate how to generate high-quality genomes from pooled marine bacterial genomes.Pooled genomic DNA from 10 marine bacteria were subjected to a single library production and sequenced with eight SMRT cells on the PacBio RS II sequencing platform. In total, 7.35 Gbp of long-read data was generated, which is equivalent to an approximate 168× average coverage for the input genomes. Genome assembly showed that eight genomes with average nucleotide identities (ANI) lower than 91.4% can be assembled with high-quality and completion using standard assembly algorithms (e.g. HGAP or Canu). A reference-based reads phasing step was developed and incorporated to assemble the complete genomes of the remaining two marine bacteria that had an ANI?>?97% and whose initial assemblies were highly fragmented.Ten complete high-quality genomes of marine bacteria were generated. The findings and developments made here, including the reference-based read phasing approach for the assembly of highly similar genomes, can be used in the future to design strategies to sequence pooled genomes using long-read sequencing.Copyright © 2019. Published by Elsevier B.V.

April 21, 2020

Genome Sequence Resource of a Puccinia striiformis Isolate infecting wheatgrass.

Stripe rust caused by Puccinia striiformis is a disastrous disease of cereal crops and various grasses. To date, fourteen stripe rust genomes are publicly available, including thirteen P. striiformis f. sp. tritici and one P. striiformis f. sp. hordei. In this study, one isolate (11-281) of P. striiformis collected from wheatgrass (Agropyron cristatum), which is avirulent to most of standard differential genotypes of wheat and barley, was sequenced, assembled, and annotated. The sequences were assembled to a draft genome of 84.75 Mb, which is comparable to previously sequenced P. striiformis f. sp. tritici and P. striiformis f. sp. hordei isolates. The draft genome comprised 381 scaffolds and contained 1,829 predicted secreted proteins. The high quality draft genome of the isolate is a valuable resource in shedding light on the evolution and pathogenicity of P. striiformis.

April 21, 2020

Genome data of Fusarium oxysporum f. sp. cubense race 1 and tropical race 4 isolates using long-read sequencing.

Fusarium wilt of banana is caused by the soil-borne fungal pathogen Fusarium oxysporum f. sp. cubense (Foc). We generated two chromosome-level assemblies of Foc race 1 and tropical race 4 strains using single-molecule real-time sequencing. The Foc1 and FocTR4 assemblies had 35 and 29 contigs with contig N50 lengths of 2.08 Mb and 4.28 Mb, respectively. These two new references genomes represent a greater than 100-fold improvement over the contig N50 statistics of the previous short read-based Foc assemblies. The two high-quality assemblies reported here will be a valuable resource for the comparative analysis of Foc races at the pathogenic levels.

April 21, 2020

Complete genome sequence provides insights into the quorum sensing-related spoilage potential of Shewanella baltica 128 isolated from spoiled shrimp.

Shewanella baltica 128 is a specific spoilage organism (SSO) isolated from the refrigerated shrimp that results in shrimp spoilage. This study reported the complete genome sequencing of this strain, with the primary annotations associated with amino acid transport and metabolism (8.66%), indicating that S. baltica 128 has good potential for degrading proteins. In vitro experiments revealed Shewanella baltica 128 could adapt to the stress conditions by regulating its growth and biofilm formation. Genes that related to the spoilage-related metabolic pathways, including trimethylamine metabolism (torT), sulfur metabolism (cysM), putrescine metabolism (speC), biofilm formation (rpoS) and serine protease production (degS), were identified. Genes (LuxS, pfs, LuxR and qseC) that related to the specific QS system were also identified. Complete genome sequence of S. baltica 128 provide insights into the QS-related spoilage potential, which might provide novel information for the development of new approaches for spoilage detection and prevention based on QS target.Copyright © 2019. Published by Elsevier Inc.

April 21, 2020

Insect genomes: progress and challenges.

In the wake of constant improvements in sequencing technologies, numerous insect genomes have been sequenced. Currently, 1219 insect genome-sequencing projects have been registered with the National Center for Biotechnology Information, including 401 that have genome assemblies and 155 with an official gene set of annotated protein-coding genes. Comparative genomics analysis showed that the expansion or contraction of gene families was associated with well-studied physiological traits such as immune system, metabolic detoxification, parasitism and polyphagy in insects. Here, we summarize the progress of insect genome sequencing, with an emphasis on how this impacts research on pest control. We begin with a brief introduction to the basic concepts of genome assembly, annotation and metrics for evaluating the quality of draft assemblies. We then provide an overview of genome information for numerous insect species, highlighting examples from prominent model organisms, agricultural pests and disease vectors. We also introduce the major insect genome databases. The increasing availability of insect genomic resources is beneficial for developing alternative pest control methods. However, many opportunities remain for developing data-mining tools that make maximal use of the available insect genome resources. Although rapid progress has been achieved, many challenges remain in the field of insect genomics. © 2019 The Royal Entomological Society.

April 21, 2020

Complete genome of Pseudomonas sp. DMSP-1 isolated from the Arctic seawater of Kongsfjorden, Svalbard

The genus Pseudomonas is highly metabolically diverse and has colonized a wide range of ecological niches. The strain Pseudomonas sp. DMSP-1 was isolated from Arctic seawater (Kongsfjorden, Svalbard) using dimethylsulfoniopropionate (DMSP) as the sole carbon source. To better understand its role in the Arctic coastal ecosystem, the genome of Pseudomonas sp. strain DMSP-1 was completely sequenced. The genome contained a circular chromosome of 6,282,445?bp with an average GC content of 60.01?mol%. A total of 5510 protein coding genes, 70 tRNA genes and 19 rRNA genes were obtained. However, no genes encoding known enzymes associated with DMSP catabolism were identified in the genome, suggesting that novel DMSP degradation genes might exist in Pseudomonas sp. strain DMSP-1.

April 21, 2020

deSALT: fast and accurate long transcriptomic read alignment with de Bruijn graph-based index

Long-read RNA sequencing (RNA-seq) is promising to transcriptomics studies, however, the alignment of the reads is still a fundamental but non-trivial task due to the sequencing errors and complicated gene structures. We propose deSALT, a tailored two-pass long RNA-seq read alignment approach, which constructs graph-based alignment skeletons to sensitively infer exons, and use them to generate spliced reference sequence to produce refined alignments. deSALT addresses several difficult issues, such as small exons, serious sequencing errors and consensus spliced alignment. Benchmarks demonstrate that this approach has a better ability to produce high-quality full-length alignments, which has enormous potentials to transcriptomics studies.

April 21, 2020

Extended haplotype phasing of de novo genome assemblies with FALCON-Phase

Haplotype-resolved genome assemblies are important for understanding how combinations of variants impact phenotypes. These assemblies can be created in various ways, such as use of tissues that contain single-haplotype (haploid) genomes, or by co-sequencing of parental genomes, but these approaches can be impractical in many situations. We present FALCON-Phase, which integrates long-read sequencing data and ultra-long-range Hi-C chromatin interaction data of a diploid individual to create high-quality, phased diploid genome assemblies. The method was evaluated by application to three datasets, including human, cattle, and zebra finch, for which high-quality, fully haplotype resolved assemblies were available for benchmarking. Phasing algorithm accuracy was affected by heterozygosity of the individual sequenced, with higher accuracy for cattle and zebra finch (>97%) compared to human (82%). In addition, scaffolding with the same Hi-C chromatin contact data resulted in phased chromosome-scale scaffolds.

April 21, 2020

Full-length mRNA sequencing and gene expression profiling reveal broad involvement of natural antisense transcript gene pairs in pepper development and response to stresses.

Pepper is an important vegetable with great economic value and unique biological features. In the past few years, significant development has been made towards understanding the huge complex pepper genome; however, pepper functional genomics has not been well studied. To better understand the pepper gene structure and pepper gene regulation, we conducted full-length mRNA sequencing by PacBio sequencing and obtained 57862 high-quality full-length mRNA sequences derived from 18362 previously annotated and 5769 newly detected genes. New gene models were built that combined the full-length mRNA sequences and corrected approximately 500 fragmented gene models from previous annotations. Based on the full-length mRNA, we identified 4114 and 5880 pepper genes forming natural antisense transcript (NAT) genes in-cis and in-trans, respectively. Most of these genes accumulate small RNAs in their overlapping regions. By analyzing these NAT gene expression patterns in our transcriptome data, we identified many NAT pairs responsive to a variety of biological processes in pepper. Pepper formate dehydrogenase 1 (FDH1), which is required for R-gene-mediated disease resistance, may be regulated by nat-siRNAs and participate in a positive feedback loop in salicylic acid biosynthesis during resistance responses. Several cis-NAT pairs and subgroups of trans-NAT genes were responsive to pepper pericarp and placenta development, which may play roles in capsanthin and capsaicin biosynthesis. Using a comparative genomics approach, the evolutionary mechanisms of cis-NATs were investigated, and we found that an increase in intergenic sequences accounted for the loss of most cis-NATs, while transposon insertion contributed to the formation of most new cis-NATs. This article is protected by copyright. All rights reserved.This article is protected by copyright. All rights reserved.

April 21, 2020

Chromosome-level assembly of the common lizard (Zootoca vivipara) genome

Squamate reptiles exhibit high variation in their traits and geographical distribution and are therefore fascinating taxa for evolutionary and ecological research. However, high-quality genomic recourses are very limited for this group of species, which inhibits some research efforts. To address this gap, we assembled a high-quality genome of the common lizard Zootoca vivipara (Lacertidae) using a combination of high coverage Illumina (shotgun and mate-pair) and PacBio sequence data, with RNAseq data and genetic linkage maps. The 1.46 Gbp genome assembly has scaffold N50 of 11.52 Mbp with N50 contig size of 220.4 Kbp and only 2.96% gaps. A BUSCO analysis indicates that 97.7% of the single-copy Tetrapoda orthologs were recovered in the assembly. In total 19,829 gene models were annotated in the genome using a combination of three ab initio and homology-based methods. To improve the chromosome-level assembly, we generated a high-density linkage map from wild-caught families and developed a novel analytical pipeline to accommodate multiple paternity and unknown father genotypes. We successfully anchored and oriented almost 90% of the genome on 19 linkage groups. This annotated and oriented chromosome-level reference genome represents a valuable resource to facilitate evolutionary studies in squamate reptiles.

April 21, 2020

Investigating the role of exudates in recruiting Streptomyces bacteria to the Arabidopsis thaliana root microbiome

Arabidopsis thaliana has a diverse but consistent root microbiome, recruited in part by the release of fixed carbon in root exudates. Here we focussed on the recruitment of Streptomyces bacteria, which are well established plant-growth-promoting rhizobacteria and which have been proposed to be recruited to A. thaliana roots by the release of salicylic acid. We generated high quality genome sequences for eight Streptomyces endophyte strains and showed that although some strains do enhance plant growth, they are not attracted to, and do not feed on, salicyclic acid. We used 13CO2 DNA-stable isotope probing to determine which bacteria are fed by the plants in the rhizo- and endosphere and found that streptomycetes did not feed on root exudates in vivo, despite the fact that they can use exudate as sole carbon and nitrogen sources in vitro. We confirmed increased root colonisation by streptomycetes in plants that constitutively produce salicylic acid, but these plants exhibited a pleiotropic phenotype of early senescence and weak growth. We propose that streptomycetes are attracted to the rhizosphere by root exudates but can be outcompeted for this food source by more abundant proteobacteria and most likely feed off unlabelled complex organic matter.

April 21, 2020

Fast and accurate long-read assembly with wtdbg2

Existing long-read assemblers require tens of thousands of CPU hours to assemble a human genome and are being outpaced by sequencing technologies in terms of both throughput and cost. We developed a novel long-read assembler wtdbg2 that, for human data, is tens of times faster than published tools while achieving comparable contiguity and accuracy. It represents a significant algorithmic advance and paves the way for population-scale long-read assembly in future.

Auto Tag: High-quality

Genomic analysis of Marinobacter sp. NP-4 and NP-6 isolated from the deep-sea oceanic crust on the western flank of the Mid-Atlantic Ridge

Variant Phasing and Haplotypic Expression from Single-molecule Long-read Sequencing in Maize

Genome sequence resource for Ilyonectria mors-panacis, causing rusty root rot of Panax notoginseng.

Complete genome sequences of pooled genomic DNA from 10 marine bacteria using PacBio long-read sequencing.

Genome Sequence Resource of a Puccinia striiformis Isolate infecting wheatgrass.

Genome data of Fusarium oxysporum f. sp. cubense race 1 and tropical race 4 isolates using long-read sequencing.

Complete genome sequence provides insights into the quorum sensing-related spoilage potential of Shewanella baltica 128 isolated from spoiled shrimp.

Complete genome of Pseudomonas sp. DMSP-1 isolated from the Arctic seawater of Kongsfjorden, Svalbard

deSALT: fast and accurate long transcriptomic read alignment with de Bruijn graph-based index

Extended haplotype phasing of de novo genome assemblies with FALCON-Phase

Full-length mRNA sequencing and gene expression profiling reveal broad involvement of natural antisense transcript gene pairs in pepper development and response to stresses.

Chromosome-level assembly of the common lizard (Zootoca vivipara) genome

Investigating the role of exudates in recruiting Streptomyces bacteria to the Arabidopsis thaliana root microbiome

Fast and accurate long-read assembly with wtdbg2

Subscribe for blog updates:

Filter by topic

Talk with an expert

Antimicrobial resistance research

Subscribe for blog updates:

Filter by topic

Talk with an expert