Menu
April 21, 2020

Human contamination in bacterial genomes has created thousands of spurious proteins.

Contaminant sequences that appear in published genomes can cause numerous problems for downstream analyses, particularly for evolutionary studies and metagenomics projects. Our large-scale scan of complete and draft bacterial and archaeal genomes in the NCBI RefSeq database reveals that 2250 genomes are contaminated by human sequence. The contaminant sequences derive primarily from high-copy human repeat regions, which themselves are not adequately represented in the current human reference genome, GRCh38. The absence of the sequences from the human assembly offers a likely explanation for their presence in bacterial assemblies. In some cases, the contaminating contigs have been erroneously annotated as containing protein-coding sequences, which over time have propagated to create spurious protein “families” across multiple prokaryotic and eukaryotic genomes. As a result, 3437 spurious protein entries are currently present in the widely used nr and TrEMBL protein databases. We report here an extensive list of contaminant sequences in bacterial genome assemblies and the proteins associated with them. We found that nearly all contaminants occurred in small contigs in draft genomes, which suggests that filtering out small contigs from draft genome assemblies may mitigate the issue of contamination while still keeping nearly all of the genuine genomic sequences. © 2019 Breitwieser et al.; Published by Cold Spring Harbor Laboratory Press.


April 21, 2020

Emergence of a ST2570 Klebsiella pneumoniae isolate carrying mcr-1 and blaCTX-M-14 recovered from a bloodstream infection in China.

The worldwide emergence of the plasmid-borne colistin resistance mediated by mcr-1 gene not only extended our knowledge on colistin resistance, but also poses a serious threat to clinical and public health [1, 2]. Since its first discovery, mcr-1-carrying Enterobacteriaceae from human, animal, food, and environmental origins have been widely identified, but few mcr-1-positive clinical strains of Klebsiella pneumoniae have been reported so far, especially when associated with community-acquired infections [3, 4]. Here, we report the emergence of a colistin-resistant K. pneumoniae isolate, which belonged to a rare sporadic clone, co-carrying mcr-1 and blaCTX-M-14 genes simultaneous recovered from a community-acquired bloodstream infection in China. Whole-genome sequencing and microbiological analysis were performed to elucidate its antimicrobial resistance mechanisms.


April 21, 2020

Comprehensive evaluation of non-hybrid genome assembly tools for third-generation PacBio long-read sequence data.

Long reads obtained from third-generation sequencing platforms can help overcome the long-standing challenge of the de novo assembly of sequences for the genomic analysis of non-model eukaryotic organisms. Numerous long-read-aided de novo assemblies have been published recently, which exhibited superior quality of the assembled genomes in comparison with those achieved using earlier second-generation sequencing technologies. Evaluating assemblies is important in guiding the appropriate choice for specific research needs. In this study, we evaluated 10 long-read assemblers using a variety of metrics on Pacific Biosciences (PacBio) data sets from different taxonomic categories with considerable differences in genome size. The results allowed us to narrow down the list to a few assemblers that can be effectively applied to eukaryotic assembly projects. Moreover, we highlight how best to use limited genomic resources for effectively evaluating the genome assemblies of non-model organisms. © The Author 2017. Published by Oxford University Press.


April 21, 2020

Expedited assessment of terrestrial arthropod diversity by coupling Malaise traps with DNA barcoding 1.

Monitoring changes in terrestrial arthropod communities over space and time requires a dramatic increase in the speed and accuracy of processing samples that cannot be achieved with morphological approaches. The combination of DNA barcoding and Malaise traps allows expedited, comprehensive inventories of species abundance whose cost will rapidly decline as high-throughput sequencing technologies advance. Aside from detailing protocols from specimen sorting to data release, this paper describes their use in a survey of arthropod diversity in a national park that examined 21?194 specimens representing 2255 species. These protocols can support arthropod monitoring programs at regional, national, and continental scales.


April 21, 2020

Transmission of ciprofloxacin resistance in Salmonella mediated by a novel type of conjugative helper plasmids.

Ciprofloxacin resistance in Salmonella has been increasingly reported due to the emergence and dissemination of multiple Plasmid-Mediated Quinolone Resistance (PMQR) determinants, which are mainly located in non-conjugative plasmids or chromosome. In this study, we aimed to depict the molecular mechanisms underlying the rare phenomenon of horizontal transfer of ciprofloxacin resistance phenotype in Salmonella by conjugation experiments, S1-PFGE and complete plasmid sequencing. Two types of non-conjugative plasmids, namely an IncX1 type carrying a qnrS1 gene, and an IncH1 plasmid carrying the oqxAB-qnrS gene, both ciprofloxacin resistance determinants in Salmonella, were recovered from two Salmonella strains. Importantly, these non-conjugative plasmids could be fused with a novel Incl1 type conjugative helper plasmid, which could target insertion sequence (IS) elements located in the non-conjugative, ciprofloxacin-resistance-encoding plasmid through replicative transcription, eventually forming a hybrid conjugative plasmid transmissible among members of Enterobacteriaceae. Since our data showed that such conjugative helper plasmids are commonly detectable among clinical Salmonella strains, particularly S. Typhimurium, fusion events leading to generation and enhanced dissemination of conjugative ciprofloxacin resistance-encoding plasmids in Salmonella are expected to result in a sharp increase in the incidence of resistance to fluoroquinolone, the key choice for treating life-threatening Salmonella infections, thereby posing a serious public health threat.


April 21, 2020

The genome of the medicinal plant Andrographis paniculata provides insight into the biosynthesis of the bioactive diterpenoid neoandrographolide.

Andrographis paniculata is a herbaceous dicot plant widely used for its anti-inflammatory and anti-viral properties across its distribution in China, India and other Southeast Asian countries. A. paniculata was used as a crucial therapeutic treatment during the influenza epidemic of 1919 in India, and is still used for the treatment of infectious disease in China. A. paniculata produces large quantities of the anti-inflammatory diterpenoid lactones andrographolide and neoandrographolide, and their analogs, which are touted to be the next generation of natural anti-inflammatory medicines for lung diseases, hepatitis, neurodegenerative disorders, autoimmune disorders and inflammatory skin diseases. Here, we report a chromosome-scale A. paniculata genome sequence of 269 Mb that was assembled by Illumina short reads, PacBio long reads and high-confidence (Hi-C) data. Gene annotation predicted 25 428 protein-coding genes. In order to decipher the genetic underpinning of diterpenoid biosynthesis, transcriptome data from seedlings elicited with methyl jasmonate were also obtained, which enabled the identification of genes encoding diterpenoid synthases, cytochrome P450 monooxygenases, 2-oxoglutarate-dependent dioxygenases and UDP-dependent glycosyltransferases potentially involved in diterpenoid lactone biosynthesis. We further carried out functional characterization of pairs of class-I and -II diterpene synthases, revealing the ability to produce diversified labdane-related diterpene scaffolds. In addition, a glycosyltransferase able to catalyze O-linked glucosylation of andrograpanin, yielding the major active product neoandrographolide, was also identified. Thus, our results demonstrate the utility of the combined genomic and transcriptomic data set generated here for the investigation of the production of the bioactive diterpenoid lactone constituents of the important medicinal herb A. paniculata. © 2018 The Authors The Plant Journal © 2018 John Wiley & Sons Ltd.


April 21, 2020

A coupled role for CsMYB75 and CsGSTF1 in anthocyanin hyperaccumulation in purple tea.

Cultivars of purple tea (Camellia sinensis) that accumulate anthocyanins in place of catechins are currently attracting global interest in their use as functional health beverages. RNA-seq of normal (LJ43) and purple Zijuan (ZJ) cultivars identified the transcription factor CsMYB75 and phi (F) class glutathione transferase CsGSTF1 as being associated with anthocyanin hyperaccumulation. Both genes mapped as a quantitative trait locus (QTL) to the purple bud leaf color (BLC) trait in F1 populations, with CsMYB75 promoting the expression of CsGSTF1 in transgenic tobacco (Nicotiana tabacum). Although CsMYB75 elevates the biosynthesis of both catechins and anthocyanins, only anthocyanins accumulate in purple tea, indicating selective downstream regulation. As glutathione transferases in other plants are known to act as transporters (ligandins) of flavonoids, directing them for vacuolar deposition, the role of CsGSTF1 in selective anthocyanin accumulation was investigated. In tea, anthocyanins accumulate in multiple vesicles, with the expression of CsGSTF1 correlated with BLC, but not with catechin content, in diverse germplasm. Complementation of the Arabidopsis tt19-8 mutant, which is unable to express the orthologous ligandin AtGSTF12, restored anthocyanin accumulation, but did not rescue the transparent testa phenotype, confirming that CsGSTF1 did not function in catechin accumulation. Consistent with a ligandin function, transient expression of CsGSTF1 in Nicotiana occurred in the nucleus, cytoplasm and membrane. Furthermore, RNA-Seq of the complemented mutants exposed to 2% sucrose as a stress treatment showed unexpected roles for anthocyanin accumulation in affecting the expression of genes involved in redox responses, phosphate homeostasis and the biogenesis of photosynthetic components, as compared with non-complemented plants. © 2018 The Authors The Plant Journal © 2018 John Wiley & Sons Ltd.


April 21, 2020

Morphotypes of the common beadlet anemone Actinia equina (L.) are genetically distinct

Anemones of the genus Actinia are ecologically important and familiar organisms on many rocky shores. However, this genus is taxonomically problematical and prior evidence suggests that the North Atlantic beadlet anemone, Actinia equina, may actually consist of a number of cryptic species. Previous genetic work has been largely limited to allozyme electrophoresis and there remains a dearth of genetic resources with which to study this genus. Mitochondrial DNA sequencing may help to clarify the taxonomy of Actinia. Here, the complete mitochondrial genome of the beadlet anemone Actinia equina (Cnidaria: Anthozoa: Actinaria: Actiniidae) is shown to be 20,690?bp in length and to contain the standard complement of Cnidarian features including 13 protein coding genes, two rRNA genes, two tRNAs and two Group I introns, one with an in-frame truncated homing endonuclease gene open reading frame. However, amplification and sequencing of the standard mtDNA barcoding region of the cytochrome oxidase I gene revealed only two haplotypes, differing by a single base pair, in widely geographically separated A. equina and its congener A. prasina. COI barcoding shows that whilst A. equina and A. prasina share the common mtDNA haplotype, haplotype frequency differed significantly between A. equina with red/orange pedal discs and those with green pedal discs, consistent with the hypothesis that these morphotypes represent incipient species.


April 21, 2020

µLAS technology for DNA isolation coupled to Cas9-assisted targeting for sequencing and assembly of a 30 kb region in plant genome.

Cas9-assisted targeting of DNA fragments in complex genomes is viewed as an essential strategy to obtain high-quality and continuous sequence data. However, the purity of target loci selected by pulsed-field gel electrophoresis (PFGE) has so far been insufficient to assemble the sequence in one contig. Here, we describe the µLAS technology to capture and purify high molecular weight DNA. First, the technology is optimized to perform high sensitivity DNA profiling with a limit of detection of 20 fg/µl for 50 kb fragments and an analytical time of 50 min. Then, µLAS is operated to isolate a 31.5 kb locus cleaved by Cas9 in the genome of the plant Medicago truncatula. Target purification is validated on a Bacterial Artificial Chromosome plasmid, and subsequently carried out in whole genome with µLAS, PFGE or by combining these techniques. PacBio sequencing shows an enrichment factor of the target sequence of 84 with PFGE alone versus 892 by association of PFGE with µLAS. These performances allow us to sequence and assemble one contig of 29 441 bp with 99% sequence identity to the reference sequence. © The Author(s) 2019. Published by Oxford University Press on behalf of Nucleic Acids Research.


April 21, 2020

De novo genome assembly of the stress tolerant forest species Casuarina equisetifolia provides insight into secondary growth.

Casuarina equisetifolia (C. equisetifolia), a conifer-like angiosperm with resistance to typhoon and stress tolerance, is mainly cultivated in the coastal areas of Australasia. C. equisetifolia, making it a valuable model to study secondary growth associated genes and stress-tolerance traits. However, the genome sequence is unavailable and therefore wood-associated growth rate and stress resistance at the molecular level is largely unexplored. We therefore constructed a high-quality draft genome sequence of C. equisetifolia by a combination of Illumina second-generation sequencing reads and Pacific Biosciences single-molecule real-time (SMRT) long reads to advance the investigation of this species. Here, we report the genome assembly, which contains approximately 300 megabases (Mb) and scaffold size of N50 is 1.06 Mb. Additionally, gene annotation, assisted by a combination of prediction and RNA-seq data, generated 29 827 annotated protein-coding genes and 1983 non-coding genes, respectively. Furthermore, we found that the total number of repetitive sequences account for one-third of the genome assembly. Here we also construct the genome-wide map of DNA modification, such as two novel forms N6 -adenine (6mA) and N4-methylcytosine (4mC) at the level of single-nucleotide resolution using single-molecule real-time (SMRT) sequencing. Interestingly, we found that 17% of 6mA modification genes and 15% of 4mC modification genes also included alternative splicing events. Finally, we investigated cellulose, hemicellulose, and lignin-related genes, which were associated with secondary growth and contained different DNA modifications. The high-quality genome sequence and annotation of C. equisetifolia in this study provide a valuable resource to strengthen our understanding of the diverse traits of trees. © 2018 The Authors The Plant Journal © 2018 John Wiley & Sons Ltd.


April 21, 2020

Conventional culture methods with commercially available media unveil the presence of novel culturable bacteria.

Recent metagenomic analysis has revealed that our gut microbiota plays an important role in not only the maintenance of our health but also various diseases such as obesity, diabetes, inflammatory bowel disease, and allergy. However, most intestinal bacteria are considered ‘unculturable’ bacteria, and their functions remain unknown. Although culture-independent genomic approaches have enabled us to gain insight into their potential roles, culture-based approaches are still required to understand their characteristic features and phenotypes. To date, various culturing methods have been attempted to obtain these ‘unculturable’ bacteria, but most such methods require advanced techniques. Here, we have tried to isolate possible unculturable bacteria from a healthy Japanese individual by using commercially available media. A 16S rRNA (ribosomal RNA) gene metagenomic analysis revealed that each culture medium showed bacterial growth depending on its selective features and a possibility of the presence of novel bacterial species. Whole genome sequencing of these candidate strains suggested the isolation of 8 novel bacterial species classified in the Actinobacteria and Firmicutes phyla. Our approach indicates that a number of intestinal bacteria hitherto considered unculturable are potentially culturable and can be cultured on commercially available media. We have obtained novel gut bacteria from a healthy Japanese individual using a combination of comprehensive genomics and conventional culturing methods. We would expect that the discovery of such novel bacteria could illuminate pivotal roles for the gut microbiota in association with human health.


April 21, 2020

Iso-Seq Allows Genome-Independent Transcriptome Profiling of Grape Berry Development.

Transcriptomics has been widely applied to study grape berry development. With few exceptions, transcriptomic studies in grape are performed using the available genome sequence, PN40024, as reference. However, differences in gene content among grape accessions, which contribute to phenotypic differences among cultivars, suggest that a single reference genome does not represent the species’ entire gene space. Though whole genome assembly and annotation can reveal the relatively unique or “private” gene space of any particular cultivar, transcriptome reconstruction is a more rapid, less costly, and less computationally intensive strategy to accomplish the same goal. In this study, we used single molecule-real time sequencing (SMRT) to sequence full-length cDNA (Iso-Seq) and reconstruct the transcriptome of Cabernet Sauvignon berries during berry ripening. In addition, short reads from ripening berries were used to error-correct low-expression isoforms and to profile isoform expression. By comparing the annotated gene space of Cabernet Sauvignon to other grape cultivars, we demonstrate that the transcriptome reference built with Iso-Seq data represents most of the expressed genes in the grape berries and includes 1,501 cultivar-specific genes. Iso-Seq produced transcriptome profiles similar to those obtained after mapping on a complete genome reference. Together, these results justify the application of Iso-Seq to identify cultivar-specific genes and build a comprehensive reference for transcriptional profiling that circumvents the necessity of a genome reference with its associated costs and computational weight.Copyright © 2019 Minio et al.


April 21, 2020

The Genome of Armadillidium vulgare (Crustacea, Isopoda) Provides Insights into Sex Chromosome Evolution in the Context of Cytoplasmic Sex Determination.

The terrestrial isopod Armadillidium vulgare is an original model to study the evolution of sex determination and symbiosis in animals. Its sex can be determined by ZW sex chromosomes, or by feminizing Wolbachia bacterial endosymbionts. Here, we report the sequence and analysis of the ZW female genome of A. vulgare. A distinguishing feature of the 1.72 gigabase assembly is the abundance of repeats (68% of the genome). We show that the Z and W sex chromosomes are essentially undifferentiated at the molecular level and the W-specific region is extremely small (at most several hundreds of kilobases). Our results suggest that recombination suppression has not spread very far from the sex-determining locus, if at all. This is consistent with A. vulgare possessing evolutionarily young sex chromosomes. We characterized multiple Wolbachia nuclear inserts in the A. vulgare genome, none of which is associated with the W-specific region. We also identified several candidate genes that may be involved in the sex determination or sexual differentiation pathways. The A. vulgare genome serves as a resource for studying the biology and evolution of crustaceans, one of the most speciose and emblematic metazoan groups. © The Author(s) 2019. Published by Oxford University Press on behalf of the Society for Molecular Biology and Evolution.


April 21, 2020

Tools and Strategies for Long-Read Sequencing and De Novo Assembly of Plant Genomes.

The commercial release of third-generation sequencing technologies (TGSTs), giving long and ultra-long sequencing reads, has stimulated the development of new tools for assembling highly contiguous genome sequences with unprecedented accuracy across complex repeat regions. We survey here a wide range of emerging sequencing platforms and analytical tools for de novo assembly, provide background information for each of their steps, and discuss the spectrum of available options. Our decision tree recommends workflows for the generation of a high-quality genome assembly when used in combination with the specific needs and resources of a project.Copyright © 2019 Elsevier Ltd. All rights reserved.


April 21, 2020

Megaphylogeny resolves global patterns of mushroom evolution.

Mushroom-forming fungi (Agaricomycetes) have the greatest morphological diversity and complexity of any group of fungi. They have radiated into most niches and fulfil diverse roles in the ecosystem, including wood decomposers, pathogens or mycorrhizal mutualists. Despite the importance of mushroom-forming fungi, large-scale patterns of their evolutionary history are poorly known, in part due to the lack of a comprehensive and dated molecular phylogeny. Here, using multigene and genome-based data, we assemble a 5,284-species phylogenetic tree and infer ages and broad patterns of speciation/extinction and morphological innovation in mushroom-forming fungi. Agaricomycetes started a rapid class-wide radiation in the Jurassic, coinciding with the spread of (sub)tropical coniferous forests and a warming climate. A possible mass extinction, several clade-specific adaptive radiations and morphological diversification of fruiting bodies followed during the Cretaceous and the Paleogene, convergently giving rise to the classic toadstool morphology, with a cap, stalk and gills (pileate-stipitate morphology). This morphology is associated with increased rates of lineage diversification, suggesting it represents a key innovation in the evolution of mushroom-forming fungi. The increase in mushroom diversity started during the Mesozoic-Cenozoic radiation event, an era of humid climate when terrestrial communities dominated by gymnosperms and reptiles were also expanding.


Talk with an expert

If you have a question, need to check the status of an order, or are interested in purchasing an instrument, we're here to help.