Menu
September 22, 2019

Resolving the complexity of human skin metagenomes using single-molecule sequencing.

Deep metagenomic shotgun sequencing has emerged as a powerful tool to interrogate composition and function of complex microbial communities. Computational approaches to assemble genome fragments have been demonstrated to be an effective tool for de novo reconstruction of genomes from these communities. However, the resultant “genomes” are typically fragmented and incomplete due to the limited ability of short-read sequence data to assemble complex or low-coverage regions. Here, we use single-molecule, real-time (SMRT) sequencing to reconstruct a high-quality, closed genome of a previously uncharacterized Corynebacterium simulans and its companion bacteriophage from a skin metagenomic sample. Considerable improvement in assembly quality occurs in hybrid approaches incorporating short-read data, with even relatively small amounts of long-read data being sufficient to improve metagenome reconstruction. Using short-read data to evaluate strain variation of this C. simulans in its skin community at single-nucleotide resolution, we observed a dominant C. simulans strain with moderate allelic heterozygosity throughout the population. We demonstrate the utility of SMRT sequencing and hybrid approaches in metagenome quantitation, reconstruction, and annotation.The species comprising a microbial community are often difficult to deconvolute due to technical limitations inherent to most short-read sequencing technologies. Here, we leverage new advances in sequencing technology, single-molecule sequencing, to significantly improve reconstruction of a complex human skin microbial community. With this long-read technology, we were able to reconstruct and annotate a closed, high-quality genome of a previously uncharacterized skin species. We demonstrate that hybrid approaches with short-read technology are sufficiently powerful to reconstruct even single-nucleotide polymorphism level variation of species in this a community. Copyright © 2016 Tsai et al.


September 22, 2019

Comparative genome and methylome analysis reveals restriction/modification system diversity in the gut commensal Bifidobacterium breve.

Bifidobacterium breve represents one of the most abundant bifidobacterial species in the gastro-intestinal tract of breast-fed infants, where their presence is believed to exert beneficial effects. In the present study whole genome sequencing, employing the PacBio Single Molecule, Real-Time (SMRT) sequencing platform, combined with comparative genome analysis allowed the most extensive genetic investigation of this taxon. Our findings demonstrate that genes encoding Restriction/Modification (R/M) systems constitute a substantial part of the B. breve variable gene content (or variome). Using the methylome data generated by SMRT sequencing, combined with targeted Illumina bisulfite sequencing (BS-seq) and comparative genome analysis, we were able to detect methylation recognition motifs and assign these to identified B. breve R/M systems, where in several cases such assignments were confirmed by restriction analysis. Furthermore, we show that R/M systems typically impose a very significant barrier to genetic accessibility of B. breve strains, and that cloning of a methyltransferase-encoding gene may overcome such a barrier, thus allowing future functional investigations of members of this species.


September 22, 2019

Culture-facilitated comparative genomics of the facultative symbiont Hamiltonella defensa.

Many insects host facultative, bacterial symbionts that confer conditional fitness benefits to their hosts. Hamiltonella defensa is a common facultative symbiont of aphids that provides protection against parasitoid wasps. Protection levels vary among strains of H. defensa that are also differentially infected by bacteriophages named APSEs. However, little is known about trait variation among strains because only one isolate has been fully sequenced. Generating complete genomes for facultative symbionts is hindered by relatively large genome sizes but low abundances in hosts like aphids that are very small. Here, we took advantage of methods for culturing H. defensa outside of aphids to generate complete genomes and transcriptome data for four strains of H. defensa from the pea aphid Acyrthosiphon pisum. Chosen strains also spanned the breadth of the H. defensa phylogeny and differed in strength of protection conferred against parasitoids. Results indicated that strains shared most genes with roles in nutrient acquisition, metabolism, and essential housekeeping functions. In contrast, the inventory of mobile genetic elements varied substantially, which generated strain specific differences in gene content and genome architecture. In some cases, specific traits correlated with differences in protection against parasitoids, but in others high variation between strains obscured identification of traits with likely roles in defense. Transcriptome data generated continuous distributions to genome assemblies with some genes that were highly expressed and others that were not. Single molecule real-time sequencing further identified differences in DNA methylation patterns and restriction modification systems that provide defense against phage infection.


September 22, 2019

2′-O-methylation in mRNA disrupts tRNA decoding during translation elongation.

Chemical modifications of mRNA may regulate many aspects of mRNA processing and protein synthesis. Recently, 2′-O-methylation of nucleotides was identified as a frequent modification in translated regions of human mRNA, showing enrichment in codons for certain amino acids. Here, using single-molecule, bulk kinetics and structural methods, we show that 2′-O-methylation within coding regions of mRNA disrupts key steps in codon reading during cognate tRNA selection. Our results suggest that 2′-O-methylation sterically perturbs interactions of ribosomal-monitoring bases (G530, A1492 and A1493) with cognate codon-anticodon helices, thereby inhibiting downstream GTP hydrolysis by elongation factor Tu (EF-Tu) and A-site tRNA accommodation, leading to excessive rejection of cognate aminoacylated tRNAs in initial selection and proofreading. Our current and prior findings highlight how chemical modifications of mRNA tune the dynamics of protein synthesis at different steps of translation elongation.


September 22, 2019

Comparative genomics of Spiraeoideae-infecting Erwinia amylovora strains provides novel insight to genetic diversity and identifies the genetic basis of a low-virulence strain.

Erwinia amylovora is the causal agent of fire blight, one of the most devastating diseases of apple and pear. Erwinia amylovora is thought to have originated in North America and has now spread to at least 50 countries worldwide. An understanding of the diversity of the pathogen population and the transmission to different geographical regions is important for the future mitigation of this disease. In this research, we performed an expanded comparative genomic study of the Spiraeoideae-infecting (SI) E. amylovora population in North America and Europe. We discovered that, although still highly homogeneous, the genetic diversity of 30 E. amylovora genomes examined was about 30 times higher than previously determined. These isolates belong to four distinct clades, three of which display geographical clustering and one of which contains strains from various geographical locations (‘Widely Prevalent’ clade). Furthermore, we revealed that strains from the Widely Prevalent clade displayed a higher level of recombination with strains from a clade strictly from the eastern USA, which suggests that the Widely Prevalent clade probably originated from the eastern USA before it spread to other locations. Finally, we detected variations in virulence in the SI E. amylovora strains on immature pear, and identified the genetic basis of one of the low-virulence strains as being caused by a single nucleotide polymorphism in hfq, a gene encoding an important virulence regulator. Our results provide insights into the population structure, distribution and evolution of SI E. amylovora in North America and Europe.© 2017 BSPP AND JOHN WILEY & SONS LTD.


September 22, 2019

MultiMotifMaker: a multi-thread tool for identifying DNA methylation motifs from Pacbio reads.

The methylation of DNA is important mechanism to control biological processes. Recently, the Pacbio SMRT technology provides a new way to identify base methylation in the genome. MotifMaker is a tool developed by Pacbio for discovering DNA methylation motifs from methylated DNA sequences. However, MotifMaker is single-threaded and computational expensive for identifying methylation motifs from large genomes. Here, we present an efficient motif finding algorithm (MultiMotifMaker) by implementing multi threads of the MotifMaker. The MultiMotifMaker, speeds up the motif search about 8-9 times on a 32 core computer comparing to MotifMaker. MultiMotifMaker makes it possible to identify methylation motifs from Pacbio reads for large genomes.


September 22, 2019

Comparative genomics of Salmonella enterica serovar Montevideo reveals lineage-specific gene differences that may influence ecological niche association.

Salmonella enterica serovar Montevideo has been linked to recent foodborne illness outbreaks resulting from contamination of products such as fruits, vegetables, seeds and spices. Studies have shown that Montevideo also is frequently associated with healthy cattle and can be isolated from ground beef, yet human salmonellosis outbreaks of Montevideo associated with ground beef contamination are rare. This disparity fuelled our interest in characterizing the genomic differences between Montevideo strains isolated from healthy cattle and beef products, and those isolated from human patients and outbreak sources. To that end, we sequenced 13 Montevideo strains to completion, producing high-quality genome assemblies of isolates from human patients (n=8) or from healthy cattle at slaughter (n=5). Comparative analysis of sequence data from this study and publicly available sequences (n=72) shows that Montevideo falls into four previously established clades, differentially occupied by cattle and human strains. The results of these analyses reveal differences in metabolic islands, environmental adhesion determinants and virulence factors within each clade, and suggest explanations for the infrequent association between bovine isolates and human illnesses.


September 22, 2019

Thermosipho spp. immune system differences affect variation in genome size and geographical distributions.

Thermosipho species inhabit thermal environments such as marine hydrothermal vents, petroleum reservoirs, and terrestrial hot springs. A 16S rRNA phylogeny of available Thermosipho spp. sequences suggested habitat specialists adapted to living in hydrothermal vents only, and habitat generalists inhabiting oil reservoirs, hydrothermal vents, and hotsprings. Comparative genomics of 15 Thermosipho genomes separated them into three distinct species with different habitat distributions: The widely distributed T. africanus and the more specialized, T. melanesiensis and T. affectus. Moreover, the species can be differentiated on the basis of genome size (GS), genome content, and immune system composition. For instance, the T. africanus genomes are largest and contained the most carbohydrate metabolism genes, which could explain why these isolates were obtained from ecologically more divergent habitats. Nonetheless, all the Thermosipho genomes, like other Thermotogae genomes, show evidence of genome streamlining. GS differences between the species could further be correlated to differences in defense capacities against foreign DNA, which influence recombination via HGT. The smallest genomes are found in T. affectus that contain both CRISPR-cas Type I and III systems, but no RM system genes. We suggest that this has caused these genomes to be almost devoid of mobile elements, contrasting the two other species genomes that contain a higher abundance of mobile elements combined with different immune system configurations. Taken together, the comparative genomic analyses of Thermosipho spp. revealed genetic variation allowing habitat differentiation within the genus as well as differentiation with respect to invading mobile DNA.


September 22, 2019

N6-methyladenine DNA modification in Xanthomonas oryzae pv. oryzicola genome.

DNA N6-methyladenine (6mA) modifications expand the information capacity of DNA and have long been known to exist in bacterial genomes. Xanthomonas oryzae pv. Oryzicola (Xoc) is the causative agent of bacterial leaf streak, an emerging and destructive disease in rice worldwide. However, the genome-wide distribution patterns and potential functions of 6mA in Xoc are largely unknown. In this study, we analyzed the levels and global distribution patterns of 6mA modification in genomic DNA of seven Xoc strains (BLS256, BLS279, CFBP2286, CFBP7331, CFBP7341, L8 and RS105). The 6mA modification was found to be widely distributed across the seven Xoc genomes, accounting for percent of 3.80, 3.10, 3.70, 4.20, 3.40, 2.10, and 3.10 of the total adenines in BLS256, BLS279, CFBP2286, CFBP7331, CFBP7341, L8, and RS105, respectively. Notably, more than 82% of 6mA sites were located within gene bodies in all seven strains. Two specific motifs for 6?mA modification, ARGT and AVCG, were prevalent in all seven strains. Comparison of putative DNA methylation motifs from the seven strains reveals that Xoc have a specific DNA methylation system. Furthermore, the 6?mA modification of rpfC dramatically decreased during Xoc infection indicates the important role for Xoc adaption to environment.


September 22, 2019

Identification of DNA base modifications by means of Pacific Biosciences RS Sequencing technology.

Whole phage genomes can be sequenced readily using one or a combination of next generation sequencing (NGS) technologies. One of the most recently developed NGS platforms, the so-called Single-Molecule Real-Time (SMRT) sequencing approach provided by the PacBio RS platform, is particularly useful in providing complete (i.e., un-gapped) genome sequences, but differs from other technologies in that the platform also allows for downstream analysis to identify nucleotides that have been modified by DNA methylation. Here, we describe the methodological approach for the detection of genomic methylation motifs by means of SMRT sequencing.


September 22, 2019

DNA Methylation by Restriction Modification Systems Affects the Global Transcriptome Profile in Borrelia burgdorferi.

Prokaryote restriction modification (RM) systems serve to protect bacteria from potentially detrimental foreign DNA. Recent evidence suggests that DNA methylation by the methyltransferase (MTase) components of RM systems can also have effects on transcriptome profiles. The type strain of the causative agent of Lyme disease, Borrelia burgdorferi B31, possesses two RM systems with N6-methyladenosine (m6A) MTase activity, which are encoded by the bbe02 gene located on linear plasmid lp25 and bbq67 on lp56. The specific recognition and/or methylation sequences had not been identified for either of these B. burgdorferi MTases, and it was not previously known whether these RM systems influence transcript levels. In the current study, single-molecule real-time sequencing was utilized to map genome-wide m6A sites and to identify consensus modified motifs in wild-type B. burgdorferi as well as MTase mutants lacking either the bbe02 gene alone or both bbe02 and bbq67 genes. Four novel conserved m6A motifs were identified and were fully attributable to the presence of specific MTases. Whole-genome transcriptome changes were observed in conjunction with the loss of MTase enzymes, indicating that DNA methylation by the RM systems has effects on gene expression. Genes with altered transcription in MTase mutants include those involved in vertebrate host colonization (e.g., rpoS regulon) and acquisition by/transmission from the tick vector (e.g., rrp1 and pdeB). The results of this study provide a comprehensive view of the DNA methylation pattern in B. burgdorferi, and the accompanying gene expression profiles add to the emerging body of research on RM systems and gene regulation in bacteria.IMPORTANCE Lyme disease is the most prevalent vector-borne disease in North America and is classified by the Centers for Disease Control and Prevention (CDC) as an emerging infectious disease with an expanding geographical area of occurrence. Previous studies have shown that the causative bacterium, Borrelia burgdorferi, methylates its genome using restriction modification systems that enable the distinction from foreign DNA. Although much research has focused on the regulation of gene expression in B. burgdorferi, the effect of DNA methylation on gene regulation has not been evaluated. The current study characterizes the patterns of DNA methylation by restriction modification systems in B. burgdorferi and evaluates the resulting effects on gene regulation in this important pathogen. Copyright © 2018 American Society for Microbiology.


September 22, 2019

Achieving Accurate Sequence and Annotation Data for Caulobacter vibrioides CB13.

Annotated sequence data are instrumental in nearly all realms of biology. However, the advent of next-generation sequencing has rapidly facilitated an imbalance between accurate sequence data and accurate annotation data. To increase the annotation accuracy of the Caulobacter vibrioides CB13b1a (CB13) genome, we compared the PGAP and RAST annotations of the CB13 genome. A total of 64 unique genes were identified in the PGAP annotation that were either completely or partially absent in the RAST annotation, and a total of 16 genes were identified in the RAST annotation that were not included in the PGAP annotation. Moreover, PGAP identified 73 frameshifted genes and 22 genes with an internal stop. In contrast, RAST annotated the larger segment of these frameshifted genes without indicating a change in reading frame may have occurred. The RAST annotation did not include any genes with internal stop codons, since it chose start codons that were after the internal stop. To confirm the discrepancies between the two annotations and verify the accuracy of the CB13 genome sequence data, we re-sequenced and re-annotated the entire genome and obtained an identical sequence, except in a small number of homopolymer regions. A genome sequence comparison between the two versions allowed us to determine the correct number of bases in each homopolymer region, which eliminated frameshifts for 31 genes annotated as frameshifted genes and removed 24 pseudogenes from the PGAP annotation. Both annotation systems correctly identified genes that were missed by the other system. In addition, PGAP identified conserved gene fragments that represented the beginning of genes, but it employed no corrective method to adjust the reading frame of frameshifted genes or the start sites of genes harboring an internal stop codon. In doing so, the PGAP annotation identified a large number of pseudogenes, which may reflect evolutionary history but likely do not produce gene products. These results demonstrate that re-sequencing and annotation comparisons can be used to increase the accuracy of genomic data and the corresponding gene annotation.


September 21, 2019

DNA-guided delivery of single molecules into zero-mode waveguides.

Zero-mode waveguides (ZMWs) are powerful analytical tools corresponding to optical nanostructures fabricated in a thin metallic film capable of confining an excitation volume to the range of attoliters. This small volume of confinement allows single-molecule fluorescence experiments to be performed at physiologically relevant concentrations of fluorescently labeled biomolecules. Exactly one molecule to be studied must be attached at the floor of the ZMW for signal detection and analysis; however, the massive parallelism of these nanoarrays suffers from a Poissonian-limited distribution of these biomolecules. To date, there is no method available that provides full single-molecule occupancy of massively arrayed ZMWs. Here we report the performance of a DNA-guided method that uses steric exclusion properties of large DNA molecules to bias the Poissonian-limited delivery of single molecules. Non-Poissonian statistics were obtained with DNA molecules that contain a free-biotinylated extremity for efficient binding to the floor of the ZMW, which resulted in a decrease of accessibility for a second molecule. Both random-coiled and condensed DNA conformations drove non-Poissonian single-molecule delivery into ZMW arrays. The results suggest that an optimal balance between the rigidity and flexibility of the macromolecule is critical for favorable accessibility and single occupancy. The optimized method provides a means for full exploitation of these massively parallelized analytical tools.


September 21, 2019

The advantages of SMRT sequencing.

Of the current next-generation sequencing technologies, SMRT sequencing is sometimes overlooked. However, attributes such as long reads, modified base detection and high accuracy make SMRT a useful technology and an ideal approach to the complete sequencing of small genomes.


July 19, 2019

Characterization of DNA methyltransferase specificities using single-molecule, real-time DNA sequencing.

DNA methylation is the most common form of DNA modification in prokaryotic and eukaryotic genomes. We have applied the method of single-molecule, real-time (SMRT) DNA sequencing that is capable of direct detection of modified bases at single-nucleotide resolution to characterize the specificity of several bacterial DNA methyltransferases (MTases). In addition to previously described SMRT sequencing of N6-methyladenine and 5-methylcytosine, we show that N4-methylcytosine also has a specific kinetic signature and is therefore identifiable using this approach. We demonstrate for all three prokaryotic methylation types that SMRT sequencing confirms the identity and position of the methylated base in cases where the MTase specificity was previously established by other methods. We then applied the method to determine the sequence context and methylated base identity for three MTases with unknown specificities. In addition, we also find evidence of unanticipated MTase promiscuity with some enzymes apparently also modifying sequences that are related, but not identical, to the cognate site.


Talk with an expert

If you have a question, need to check the status of an order, or are interested in purchasing an instrument, we're here to help.