Menu
April 21, 2020

Long-read sequence and assembly of segmental duplications.

We have developed a computational method based on polyploid phasing of long sequence reads to resolve collapsed regions of segmental duplications within genome assemblies. Segmental Duplication Assembler (SDA; https://github.com/mvollger/SDA ) constructs graphs in which paralogous sequence variants define the nodes and long-read sequences provide attraction and repulsion edges, enabling the partition and assembly of long reads corresponding to distinct paralogs. We apply it to single-molecule, real-time sequence data from three human genomes and recover 33-79 megabase pairs (Mb) of duplications in which approximately half of the loci are diverged (<99.8%) compared to the reference genome. We show that the corresponding sequence is highly accurate (>99.9%) and that the diverged sequence corresponds to copy-number-variable paralogs that are absent from the human reference genome. Our method can be applied to other complex genomes to resolve the last gene-rich gaps, improve duplicate gene annotation, and better understand copy-number-variant genetic diversity at the base-pair level.


April 21, 2020

FLAM-seq: full-length mRNA sequencing reveals principles of poly(A) tail length control.

Although messenger RNAs are key molecules for understanding life, until now, no method has existed to determine the full-length sequence of endogenous mRNAs including their poly(A) tails. Moreover, although non-A nucleotides can be incorporated in poly(A) tails, there also exists no method to accurately sequence them. Here, we present full-length poly(A) and mRNA sequencing (FLAM-seq), a rapid and simple method for high-quality sequencing of entire mRNAs. We report a complementary DNA library preparation method coupled to single-molecule sequencing to perform FLAM-seq. Using human cell lines, brain organoids and Caenorhabditis elegans we show that FLAM-seq delivers high-quality full-length mRNA sequences for thousands of different genes per sample. We find that 3′ untranslated region length is correlated with poly(A) tail length, that alternative polyadenylation sites and alternative promoters for the same gene are linked to different tail lengths, and that tails contain a substantial number of cytosines.


April 21, 2020

Comprehensive evaluation of non-hybrid genome assembly tools for third-generation PacBio long-read sequence data.

Long reads obtained from third-generation sequencing platforms can help overcome the long-standing challenge of the de novo assembly of sequences for the genomic analysis of non-model eukaryotic organisms. Numerous long-read-aided de novo assemblies have been published recently, which exhibited superior quality of the assembled genomes in comparison with those achieved using earlier second-generation sequencing technologies. Evaluating assemblies is important in guiding the appropriate choice for specific research needs. In this study, we evaluated 10 long-read assemblers using a variety of metrics on Pacific Biosciences (PacBio) data sets from different taxonomic categories with considerable differences in genome size. The results allowed us to narrow down the list to a few assemblers that can be effectively applied to eukaryotic assembly projects. Moreover, we highlight how best to use limited genomic resources for effectively evaluating the genome assemblies of non-model organisms. © The Author 2017. Published by Oxford University Press.


April 21, 2020

Whole genome sequencing of NDM-1-producing serotype K1 ST23 hypervirulent Klebsiella pneumoniae in China.

The emergence and spread of carbapenem-resistant hypervirulent Klebsiella pneumoniae (CR-hvKP) is causing worldwide concern, whereas NDM-producing hvKP is still rare. Here we report the complete genome sequence characteristics of an NDM-1-producing ST23 type clinical hvKP in PR China.Capsular polysaccharide serotyping was performed by PCR. The complete genome sequence of isolate 3214 was obtained using both the Illumina Hiseq platform and Pacbio RS platform. Multilocus sequence type was identified by submitting the genome sequence to mlst 2.0 and the antimicrobial resistance genes and plasmid replicons were identified using ResFinder and PlasmidFinder, respectively. Transferability of the blaNDM-1-bearing plasmid was determined by conjugation experiment, S1 pulsed-field gel electrophoresis and Southern hybridization.Isolate 3214 was classified to ST23 and belonged to the K1 capsular serotype. The isolate’s total genome size was 6 171 644?bp with a G+C content of 56.39 %, consisting of a 5 448 209?bp chromosome and seven plasmids. The resistome included 18 types of antibiotic resistance genes. Fourteen resistance genes including blaNDM-1 and blaCTX-M-14 were located on plasmids and five also including blaCTX-M-14 were in the chromosome. Plasmid pNDM_3214 carrying blaNDM-1 harboured six types of resistance genes surrounded by insertion sequences and was conjugative. The worldwide pLVPK-like virulence plasmid harbouring rmpA2 and rmpA was also found in this isolate.This study provides basic information of phenotypic and genomic features of ST23 CR-hvKP isolate 3214. Our data highlights the potential risk of spread of NDM-1-producing ST23 hvKP.


April 21, 2020

Transmission of ciprofloxacin resistance in Salmonella mediated by a novel type of conjugative helper plasmids.

Ciprofloxacin resistance in Salmonella has been increasingly reported due to the emergence and dissemination of multiple Plasmid-Mediated Quinolone Resistance (PMQR) determinants, which are mainly located in non-conjugative plasmids or chromosome. In this study, we aimed to depict the molecular mechanisms underlying the rare phenomenon of horizontal transfer of ciprofloxacin resistance phenotype in Salmonella by conjugation experiments, S1-PFGE and complete plasmid sequencing. Two types of non-conjugative plasmids, namely an IncX1 type carrying a qnrS1 gene, and an IncH1 plasmid carrying the oqxAB-qnrS gene, both ciprofloxacin resistance determinants in Salmonella, were recovered from two Salmonella strains. Importantly, these non-conjugative plasmids could be fused with a novel Incl1 type conjugative helper plasmid, which could target insertion sequence (IS) elements located in the non-conjugative, ciprofloxacin-resistance-encoding plasmid through replicative transcription, eventually forming a hybrid conjugative plasmid transmissible among members of Enterobacteriaceae. Since our data showed that such conjugative helper plasmids are commonly detectable among clinical Salmonella strains, particularly S. Typhimurium, fusion events leading to generation and enhanced dissemination of conjugative ciprofloxacin resistance-encoding plasmids in Salmonella are expected to result in a sharp increase in the incidence of resistance to fluoroquinolone, the key choice for treating life-threatening Salmonella infections, thereby posing a serious public health threat.


April 21, 2020

Streptococcus periodonticum sp. nov., Isolated from Human Subgingival Dental Plaque of Periodontitis Lesion.

A novel facultative anaerobic and Gram-stain-positive coccus, designated strain ChDC F135T, was isolated from human subgingival dental plaque of periodontitis lesion and was characterized by polyphasic taxonomic analysis. The 16S rRNA gene (16S rDNA) sequence of strain ChDC F135T was closest to that of Streptococcus sinensis HKU4T (98.2%), followed by Streptococcus intermedia SK54T (97.0%), Streptococcus constellatus NCTC11325T (96.0%), and Streptococcus anginosus NCTC 10713T (95.7%). In contrast, phylogenetic analysis based on the superoxide dismutase gene (sodA) and the RNA polymerase beta-subunit gene (rpoB) showed that the nucleotide sequence similarities of strain ChDC F135T were highly similar to the corresponding genes of S. anginosus NCTC 10713T (99.2% and 97.6%, respectively), S. constellatus NCTC11325T (87.8% and 91.4%, respectively), and S. intermedia SK54T (85.8% and 91.2%, respectively) rather than those of S. sinensis HKU4T (80.5% and 82.6%). The complete genome of strain ChDC F135T consisted of 1,901,251 bp and the G+C content was 38.9 mol %. Average nucleotide identity value between strain ChDC F135T and S. sinensis HKU4T or S. anginosus NCTC 10713T were 75.7% and 95.6%, respectively. The C14:0 composition of the cellular fatty acids of strain ChDC F135T (32.8%) was different from that of S. intermedia (6-8%), S. constellatus (6-13%), and S. anginosus (13-20%). Based on the results of phylogenetic and phenotypic analysis, strain ChDC F135T (=?KCOM 2412T?=?JCM 33300T) was classified as a type strain of a novel species of the genus Streptococcus, for which we proposed the name Streptococcus periodonticum sp. nov.


April 21, 2020

Assembly of allele-aware, chromosomal-scale autopolyploid genomes based on Hi-C data.

Construction of chromosome-level assembly is a vital step in achieving the goal of a ‘Platinum’ genome, but it remains a major challenge to assemble and anchor sequences to chromosomes in autopolyploid or highly heterozygous genomes. High-throughput chromosome conformation capture (Hi-C) technology serves as a robust tool to dramatically advance chromosome scaffolding; however, existing approaches are mostly designed for diploid genomes and often with the aim of reconstructing a haploid representation, thereby having limited power to reconstruct chromosomes for autopolyploid genomes. We developed a novel algorithm (ALLHiC) that is capable of building allele-aware, chromosomal-scale assembly for autopolyploid genomes using Hi-C paired-end reads with innovative ‘prune’ and ‘optimize’ steps. Application on simulated data showed that ALLHiC can phase allelic contigs and substantially improve ordering and orientation when compared to other mainstream Hi-C assemblers. We applied ALLHiC on an autotetraploid and an autooctoploid sugar-cane genome and successfully constructed the phased chromosomal-level assemblies, revealing allelic variations present in these two genomes. The ALLHiC pipeline enables de novo chromosome-level assembly of autopolyploid genomes, separating each allele. Haplotype chromosome-level assembly of allopolyploid and heterozygous diploid genomes can be achieved using ALLHiC, overcoming obstacles in assembling complex genomes.


April 21, 2020

Streptococcus gwangjuense sp. nov., Isolated from Human Pericoronitis.

A novel facultative anaerobic, Gram-stain-negative coccus, designated strain ChDC B345T, was isolated from human pericoronitis lesion and was characterized by polyphasic taxonomic analysis. The 16S ribosomal RNA gene (16S rDNA) sequence revealed that the strain belonged to the genus Streptococcus. The 16S rDNA sequence of strain ChDC B345T was most closely related to those of  Streptococcus mitis NCTC 12261T (99.5%) and Streptococcus pseudopneumoniae ATCC BAA-960T (99.5%). Complete genome of strain ChDC B345T was 1,972,471 bp in length and the G?+?C content was 40.2 mol%. Average nucleotide identity values between strain ChDC B345T and S. pseudopneumoniae ATCC BAA-960T or S. mitis NCTC 12261T were 92.17% and 93.63%, respectively. Genome-to-genome distance values between strain ChDC B345T and S. pseudopneumoniae ATCC BAA-960T or S. mitis NCTC 12261T were 47.8% (45.2-50.4%) and 53.0% (51.0-56.4%), respectively. Based on these results, strain ChDC B345T (=?KCOM 1679T?=?JCM 33299T) should be classified as a novel species of genus Streptococcus, for which we propose the name Streptococcus gwangjuense sp. nov.


April 21, 2020

Conventional culture methods with commercially available media unveil the presence of novel culturable bacteria.

Recent metagenomic analysis has revealed that our gut microbiota plays an important role in not only the maintenance of our health but also various diseases such as obesity, diabetes, inflammatory bowel disease, and allergy. However, most intestinal bacteria are considered ‘unculturable’ bacteria, and their functions remain unknown. Although culture-independent genomic approaches have enabled us to gain insight into their potential roles, culture-based approaches are still required to understand their characteristic features and phenotypes. To date, various culturing methods have been attempted to obtain these ‘unculturable’ bacteria, but most such methods require advanced techniques. Here, we have tried to isolate possible unculturable bacteria from a healthy Japanese individual by using commercially available media. A 16S rRNA (ribosomal RNA) gene metagenomic analysis revealed that each culture medium showed bacterial growth depending on its selective features and a possibility of the presence of novel bacterial species. Whole genome sequencing of these candidate strains suggested the isolation of 8 novel bacterial species classified in the Actinobacteria and Firmicutes phyla. Our approach indicates that a number of intestinal bacteria hitherto considered unculturable are potentially culturable and can be cultured on commercially available media. We have obtained novel gut bacteria from a healthy Japanese individual using a combination of comprehensive genomics and conventional culturing methods. We would expect that the discovery of such novel bacteria could illuminate pivotal roles for the gut microbiota in association with human health.


April 21, 2020

Polysaccharide utilization loci of North Sea Flavobacteriia as basis for using SusC/D-protein expression for predicting major phytoplankton glycans.

Marine algae convert a substantial fraction of fixed carbon dioxide into various polysaccharides. Flavobacteriia that are specialized on algal polysaccharide degradation feature genomic clusters termed polysaccharide utilization loci (PULs). As knowledge on extant PUL diversity is sparse, we sequenced the genomes of 53 North Sea Flavobacteriia and obtained 400 PULs. Bioinformatic PUL annotations suggest usage of a large array of polysaccharides, including laminarin, a-glucans, and alginate as well as mannose-, fucose-, and xylose-rich substrates. Many of the PULs exhibit new genetic architectures and suggest substrates rarely described for marine environments. The isolates’ PUL repertoires often differed considerably within genera, corroborating ecological niche-associated glycan partitioning. Polysaccharide uptake in Flavobacteriia is mediated by SusCD-like transporter complexes. Respective protein trees revealed clustering according to polysaccharide specificities predicted by PUL annotations. Using the trees, we analyzed expression of SusC/D homologs in multiyear phytoplankton bloom-associated metaproteomes and found indications for profound changes in microbial utilization of laminarin, a-glucans, ß-mannan, and sulfated xylan. We hence suggest the suitability of SusC/D-like transporter protein expression within heterotrophic bacteria as a proxy for the temporal utilization of discrete polysaccharides.


April 21, 2020

FadR1, a pathway-specific activator of fidaxomicin biosynthesis in Actinoplanes deccanensis Yp-1.

Fidaxomicin, an 18-membered macrolide antibiotic, is highly active against Clostridium difficile, the most common cause of diarrhea in hospitalized patients. Though the biosynthetic mechanism of fidaxomicin has been well studied, little is known about its regulatory mechanism. Here, we reported that FadR1, a LAL family transcriptional regulator in the fidaxomicin cluster of Actinoplanes deccanensis Yp-1, acts as an activator for fidaxomicin biosynthesis. The disruption of fadR1 abolished the ability to synthesize fidaxomicin, and production could be restored by reintegrating a single copy of fadR1. Overexpression of fadR1 resulted in an approximately 400 % improvement in fidaxomicin production. Electrophoretic mobility shift assays indicated that fidaxomicin biosynthesis is under the control of FadR1 through its binding to the promoter regions of fadM, fadA1-fadP2, fadS2-fadC, and fadE-fadF, respectively. And the conserved binding sites of FadR1 within the four promoter regions were determined by footprinting experiment. All results indicated that fadR1 encodes a pathway-specific positive regulator of fidaxomicin biosynthesis and upregulates the transcription levels of most of genes by binding to the four above intergenic regions. In summary, we not only clearly elucidate the regulatory mechanism of FadR1 but also provide strategies for the construction of industrial high-yield strain of fidaxomicin.


April 21, 2020

Genetic variation in the conjugative plasmidome of a hospital effluent multidrug resistant Escherichia coli strain.

Bacteria harboring conjugative plasmids have the potential for spreading antibiotic resistance through horizontal gene transfer. It is described that the selection and dissemination of antibiotic resistance is enhanced by stressors, like metals or antibiotics, which can occur as environmental contaminants. This study aimed at unveiling the composition of the conjugative plasmidome of a hospital effluent multidrug resistant Escherichia coli strain (H1FC54) under different mating conditions. To meet this objective, plasmid pulsed field gel electrophoresis, optical mapping analyses and DNA sequencing were used in combination with phenotype analysis. Strain H1FC54 was observed to harbor five plasmids, three of which were conjugative and two of these, pH1FC54_330 and pH1FC54_140, contained metal and antibiotic resistance genes. Transconjugants obtained in the absence or presence of tellurite (0.5?µM or 5?µM), arsenite (0.5?µM, 5?µM or 15?µM) or ceftazidime (10?mg/L) and selected in the presence of sodium azide (100?mg/L) and tetracycline (16?mg/L) presented distinct phenotypes, associated with the acquisition of different plasmid combinations, including two co-integrate plasmids, of 310 kbp and 517 kbp. The variable composition of the conjugative plasmidome, the formation of co-integrates during conjugation, as well as the transfer of non-transferable plasmids via co-integration, and the possible association between antibiotic, arsenite and tellurite tolerance was demonstrated. These evidences bring interesting insights into the comprehension of the molecular and physiological mechanisms that underlie antibiotic resistance propagation in the environment. Copyright © 2019 Elsevier Ltd. All rights reserved.


April 21, 2020

Single-Molecule Sequencing: Towards Clinical Applications.

In the past several years, single-molecule sequencing platforms, such as those by Pacific Biosciences and Oxford Nanopore Technologies, have become available to researchers and are currently being tested for clinical applications. They offer exceptionally long reads that permit direct sequencing through regions of the genome inaccessible or difficult to analyze by short-read platforms. This includes disease-causing long repetitive elements, extreme GC content regions, and complex gene loci. Similarly, these platforms enable structural variation characterization at previously unparalleled resolution and direct detection of epigenetic marks in native DNA. Here, we review how these technologies are opening up new clinical avenues that are being applied to pathogenic microorganisms and viruses, constitutional disorders, pharmacogenomics, cancer, and more.Copyright © 2018 Elsevier Ltd. All rights reserved.


April 21, 2020

The Genome of Armadillidium vulgare (Crustacea, Isopoda) Provides Insights into Sex Chromosome Evolution in the Context of Cytoplasmic Sex Determination.

The terrestrial isopod Armadillidium vulgare is an original model to study the evolution of sex determination and symbiosis in animals. Its sex can be determined by ZW sex chromosomes, or by feminizing Wolbachia bacterial endosymbionts. Here, we report the sequence and analysis of the ZW female genome of A. vulgare. A distinguishing feature of the 1.72 gigabase assembly is the abundance of repeats (68% of the genome). We show that the Z and W sex chromosomes are essentially undifferentiated at the molecular level and the W-specific region is extremely small (at most several hundreds of kilobases). Our results suggest that recombination suppression has not spread very far from the sex-determining locus, if at all. This is consistent with A. vulgare possessing evolutionarily young sex chromosomes. We characterized multiple Wolbachia nuclear inserts in the A. vulgare genome, none of which is associated with the W-specific region. We also identified several candidate genes that may be involved in the sex determination or sexual differentiation pathways. The A. vulgare genome serves as a resource for studying the biology and evolution of crustaceans, one of the most speciose and emblematic metazoan groups. © The Author(s) 2019. Published by Oxford University Press on behalf of the Society for Molecular Biology and Evolution.


April 21, 2020

Double PIK3CA mutations in cis increase oncogenicity and sensitivity to PI3Ka inhibitors.

Activating mutations in PIK3CA are frequent in human breast cancer, and phosphoinositide 3-kinase alpha (PI3Ka) inhibitors have been approved for therapy. To characterize determinants of sensitivity to these agents, we analyzed PIK3CA-mutant cancer genomes and observed the presence of multiple PIK3CA mutations in 12 to 15% of breast cancers and other tumor types, most of which (95%) are double mutations. Double PIK3CA mutations are in cis on the same allele and result in increased PI3K activity, enhanced downstream signaling, increased cell proliferation, and tumor growth. The biochemical mechanisms of dual mutations include increased disruption of p110a binding to the inhibitory subunit p85a, which relieves its catalytic inhibition, and increased p110a membrane lipid binding. Double PIK3CA mutations predict increased sensitivity to PI3Ka inhibitors compared with single-hotspot mutations.Copyright © 2019 The Authors, some rights reserved; exclusive licensee American Association for the Advancement of Science. No claim to original U.S. Government Works.


Talk with an expert

If you have a question, need to check the status of an order, or are interested in purchasing an instrument, we're here to help.