Menu
July 7, 2019

What caused the outbreak of ESBL-producing Klebsiella pneumoniae in a neonatal intensive care unit, Germany 2009 to 2012? Reconstructing transmission with epidemiological analysis and whole-genome sequencing.

We aimed to retrospectively reconstruct the timing of transmission events and pathways in order to understand why extensive preventive measures and investigations were not sufficient to prevent new cases.We extracted available information from patient charts to describe cases and to compare them to the normal population of the ward. We conducted a cohort study to identify risk factors for pathogen acquisition. We sequenced the available isolates to determine the phylogenetic relatedness of Klebsiella pneumoniae isolates on the basis of their genome sequences.The investigation comprises 37 cases and the 10 cases with ESBL (extended-spectrum beta-lactamase)-producing K. pneumoniae bloodstream infection. Descriptive epidemiology indicated that a continuous transmission from person to person was most likely. Results from the cohort study showed that ‘frequent manipulation’ (a proxy for increased exposure to medical procedures) was significantly associated with being a case (RR 1.44, 95% CI 1.02 to 2.19). Genome sequences revealed that all 48 bacterial isolates available for sequencing from 31 cases were closely related (maximum genetic distance, 12 single nucleotide polymorphisms). Based on our calculation of evolutionary rate and sequence diversity, we estimate that the outbreak strain was endemic since 2008. Epidemiological and phylogenetic analyses consistently indicated that there were additional, undiscovered cases prior to the onset of microbiological screening and that the spread of the pathogen remained undetected over several years, driven predominantly by person-to-person transmission. Whole-genome sequencing provided valuable information on the onset, course and size of the outbreak, and on possible ways of transmission. Published by the BMJ Publishing Group Limited. For permission to use (where not already granted under a licence) please go to http://group.bmj.com/group/rights-licensing/permissions.


July 7, 2019

Complete genome sequence of ER2796, a DNA methyltransferase-deficient strain of Escherichia coli K-12.

We report the complete sequence of ER2796, a laboratory strain of Escherichia coli K-12 that is completely defective in DNA methylation. Because of its lack of any native methylation, it is extremely useful as a host into which heterologous DNA methyltransferase genes can be cloned and the recognition sequences of their products deduced by Pacific Biosciences Single-Molecule Real Time (SMRT) sequencing. The genome was itself sequenced from a long-insert library using the SMRT platform, resulting in a single closed contig devoid of methylated bases. Comparison with K-12 MG1655, the first E. coli K-12 strain to be sequenced, shows an essentially co-linear relationship with no major rearrangements despite many generations of laboratory manipulation. The comparison revealed a total of 41 insertions and deletions, and 228 single base pair substitutions. In addition, the long-read approach facilitated the surprising discovery of four gene conversion events, three involving rRNA operons and one between two cryptic prophages. Such events thus contribute both to genomic homogenization and to bacteriophage diversification. As one of relatively few laboratory strains of E. coli to be sequenced, the genome also reveals the sequence changes underlying a number of classical mutant alleles including those affecting the various native DNA methylation systems.


July 7, 2019

Genome expansion via lineage splitting and genome reduction in the cicada endosymbiont Hodgkinia.

Comparative genomics from mitochondria, plastids, and mutualistic endosymbiotic bacteria has shown that the stable establishment of a bacterium in a host cell results in genome reduction. Although many highly reduced genomes from endosymbiotic bacteria are stable in gene content and genome structure, organelle genomes are sometimes characterized by dramatic structural diversity. Previous results from Candidatus Hodgkinia cicadicola, an endosymbiont of cicadas, revealed that some lineages of this bacterium had split into two new cytologically distinct yet genetically interdependent species. It was hypothesized that the long life cycle of cicadas in part enabled this unusual lineage-splitting event. Here we test this hypothesis by investigating the structure of the Ca. Hodgkinia genome in one of the longest-lived cicadas, Magicicada tredecim. We show that the Ca. Hodgkinia genome from M. tredecim has fragmented into multiple new chromosomes or genomes, with at least some remaining partitioned into discrete cells. We also show that this lineage-splitting process has resulted in a complex of Ca. Hodgkinia genomes that are 1.1-Mb pairs in length when considered together, an almost 10-fold increase in size from the hypothetical single-genome ancestor. These results parallel some examples of genome fragmentation and expansion in organelles, although the mechanisms that give rise to these extreme genome instabilities are likely different.


July 7, 2019

It’s more than stamp collecting: how genome sequencing can unify biological research.

The availability of reference genome sequences, especially the human reference, has revolutionized the study of biology. However, while the genomes of some species have been fully sequenced, a wide range of biological problems still cannot be effectively studied for lack of genome sequence information. Here, I identify neglected areas of biology and describe how both targeted species sequencing and more broad taxonomic surveys of the tree of life can address important biological questions. I enumerate the significant benefits that would accrue from sequencing a broader range of taxa, as well as discuss the technical advances in sequencing and assembly methods that would allow for wide-ranging application of whole-genome analysis. Finally, I suggest that in addition to ‘big science’ survey initiatives to sequence the tree of life, a modified infrastructure-funding paradigm would better support reference genome sequence generation for research communities most in need. Copyright © 2015 Elsevier Ltd. All rights reserved.


July 7, 2019

GAML: genome assembly by maximum likelihood.

Resolution of repeats and scaffolding of shorter contigs are critical parts of genome assembly. Modern assemblers usually perform such steps by heuristics, often tailored to a particular technology for producing paired or long reads.We propose a new framework that allows systematic combination of diverse sequencing datasets into a single assembly. We achieve this by searching for an assembly with the maximum likelihood in a probabilistic model capturing error rate, insert lengths, and other characteristics of the sequencing technology used to produce each dataset. We have implemented a prototype genome assembler GAML that can use any combination of insert sizes with Illumina or 454 reads, as well as PacBio reads. Our experiments show that we can assemble short genomes with N50 sizes and error rates comparable to ALLPATHS-LG or Cerulean. While ALLPATHS-LG and Cerulean require each a specific combination of datasets, GAML works on any combination.We have introduced a new probabilistic approach to genome assembly and demonstrated that this approach can lead to superior results when used to combine diverse set of datasets from different sequencing technologies. Data and software is available at http://compbio.fmph.uniba.sk/gaml.


July 7, 2019

Complete genome sequence of endophytic nitrogen-fixing Klebsiella variicola strain DX120E.

Klebsiella variicola strain DX120E (=CGMCC 1.14935) is an endophytic nitrogen-fixing bacterium isolated from sugarcane crops grown in Guangxi, China and promotes sugarcane growth. Here we summarize the features of the strain DX120E and describe its complete genome sequence. The genome contains one circular chromosome and two plasmids, and contains 5,718,434 nucleotides with 57.1% GC content, 5,172 protein-coding genes, 25 rRNA genes, 87 tRNA genes, 7 ncRNA genes, 25 pseudo genes, and 2 CRISPR repeats.


July 7, 2019

Complete genome sequence of Actinobacillus equuli subspecies equuli ATCC 19392(T).

Actinobacillus equuli subsp. equuli is a member of the family Pasteurellaceae that is a common resident of the oral cavity and alimentary tract of healthy horses. At the same time, it can also cause a fatal septicemia in foals, commonly known as sleepy foal disease or joint ill disease. In addition, A. equuli subsp. equuli has recently been reported to act as a primary pathogen in breeding sows and piglets. To better understand how A. equuli subsp. equuli can cause disease, the genome of the type strain of A. equuli subsp. equuli, ATCC 19392(T), was sequenced using the PacBio RS II sequencing system. Its genome is comprised of 2,431,533 bp and is predicted to encode 2,264 proteins and 82 RNAs.


July 7, 2019

Draft genome sequence of Raoultella terrigena R1Gly, a diazotrophic endophyte.

Raoultella terrigena R1Gly is a diazotrophic endophyte isolated from surface-sterilized roots of Nicotiana tabacum. The whole-genome sequence was obtained to investigate the endophytic characteristics of this organism at the genetic level, as well as to compare this strain with its close relatives. To our knowledge, this is the first genome obtained from the Raoultella terrigena species and only the third genome from the Raoultella genus, after Raoultella ornitholytic and Raoultella planticola. This genome will provide a foundation for further comparative genomic, metagenomic, and functional studies of this genus. Copyright © 2015 Schicklberger et al.


July 7, 2019

Comparative analyses of clinical and environmental populations of Cryptococcus neoformans in Botswana.

Cryptococcus neoformans var. grubii (Cng) is the most common cause of fungal meningitis, and its prevalence is highest in sub-Saharan Africa. Patients become infected by inhaling airborne spores or desiccated yeast cells from the environment, where the fungus thrives in avian droppings, trees and soil. To investigate the prevalence and population structure of Cng in southern Africa, we analysed isolates from 77 environmental samples and 64 patients. We detected significant genetic diversity among isolates and strong evidence of geographic structure at the local level. High proportions of isolates with the rare MATa allele were observed in both clinical and environmental isolates; however, the mating-type alleles were unevenly distributed among different subpopulations. Nearly equal proportions of the MATa and MATa mating types were observed among all clinical isolates and in one environmental subpopulation from the eastern part of Botswana. As previously reported, there was evidence of both clonality and recombination in different geographic areas. These results provide a foundation for subsequent genomewide association studies to identify genes and genotypes linked to pathogenicity in humans. © 2015 The Authors. Molecular Ecology published by John Wiley & Sons Ltd.


July 7, 2019

Sequence type 1 group B Streptococcus, an emerging cause of invasive disease in adults, evolves by small genetic changes.

The molecular mechanisms underlying pathogen emergence in humans is a critical but poorly understood area of microbiologic investigation. Serotype V group B Streptococcus (GBS) was first isolated from humans in 1975, and rates of invasive serotype V GBS disease significantly increased starting in the early 1990s. We found that 210 of 229 serotype V GBS strains (92%) isolated from the bloodstream of nonpregnant adults in the United States and Canada between 1992 and 2013 were multilocus sequence type (ST) 1. Elucidation of the complete genome of a 1992 ST-1 strain revealed that this strain had the highest homology with a GBS strain causing cow mastitis and that the 1992 ST-1 strain differed from serotype V strains isolated in the late 1970s by acquisition of cell surface proteins and antimicrobial resistance determinants. Whole-genome comparison of 202 invasive ST-1 strains detected significant recombination in only eight strains. The remaining 194 strains differed by an average of 97 SNPs. Phylogenetic analysis revealed a temporally dependent mode of genetic diversification consistent with the emergence in the 1990s of ST-1 GBS as major agents of human disease. Thirty-one loci were identified as being under positive selective pressure, and mutations at loci encoding polysaccharide capsule production proteins, regulators of pilus expression, and two-component gene regulatory systems were shown to affect the bacterial phenotype. These data reveal that phenotypic diversity among ST-1 GBS is mainly driven by small genetic changes rather than extensive recombination, thereby extending knowledge into how pathogens adapt to humans.


July 7, 2019

Genome sequence of the alkaline-tolerant Cellulomonas sp. strain FA1.

We present the genome of the cellulose-degrading Cellulomonas sp. strain FA1 isolated from an actively serpentinizing highly alkaline spring. Knowledge of this genome will enable studies into the molecular basis of plant material degradation in alkaline environments and inform the development of lignocellulose bioprocessing procedures for biofuel production. Copyright © 2015 Cohen et al.


July 7, 2019

Genome sequence of Porticoccus hydrocarbonoclasticus strain MCTG13d, an obligate polycyclic aromatic hydrocarbon-degrading bacterium associated with marine eukaryotic phytoplankton.

Porticoccus hydrocarbonoclasticus strain MCTG13d is a recently discovered bacterium that is associated with marine eukaryotic phytoplankton and that almost exclusively utilizes polycyclic aromatic hydrocarbons (PAHs) as the sole source of carbon and energy. Here, we present the genome sequence of this strain, which is 2,474,654 bp with 2,385 genes and has an average G+C content of 53.1%. Copyright © 2015 Gutierrez et al.


July 7, 2019

The Streptomyces leeuwenhoekii genome: de novo sequencing and assembly in single contigs of the chromosome, circular plasmid pSLE1 and linear plasmid pSLE2.

Next Generation DNA Sequencing (NGS) and genome mining of actinomycetes and other microorganisms is currently one of the most promising strategies for the discovery of novel bioactive natural products, potentially revealing novel chemistry and enzymology involved in their biosynthesis. This approach also allows rapid insights into the biosynthetic potential of microorganisms isolated from unexploited habitats and ecosystems, which in many cases may prove difficult to culture and manipulate in the laboratory. Streptomyces leeuwenhoekii (formerly Streptomyces sp. strain C34) was isolated from the hyper-arid high-altitude Atacama Desert in Chile and shown to produce novel polyketide antibiotics.Here we present the de novo sequencing of the S. leeuwenhoekii linear chromosome (8 Mb) and two extrachromosomal replicons, the circular pSLE1 (86 kb) and the linear pSLE2 (132 kb), all in single contigs, obtained by combining Pacific Biosciences SMRT (PacBio) and Illumina MiSeq technologies. We identified the biosynthetic gene clusters for chaxamycin, chaxalactin, hygromycin A and desferrioxamine E, metabolites all previously shown to be produced by this strain (J Nat Prod, 2011, 74:1965) and an additional 31 putative gene clusters for specialised metabolites. As well as gene clusters for polyketides and non-ribosomal peptides, we also identified three gene clusters encoding novel lasso-peptides.The S. leeuwenhoekii genome contains 35 gene clusters apparently encoding the biosynthesis of specialised metabolites, most of them completely novel and uncharacterised. This project has served to evaluate the current state of NGS for efficient and effective genome mining of high GC actinomycetes. The PacBio technology now permits the assembly of actinomycete replicons into single contigs with >99 % accuracy. The assembled Illumina sequence permitted not only the correction of omissions found in GC homopolymers in the PacBio assembly (exacerbated by the high GC content of actinomycete DNA) but it also allowed us to obtain the sequences of the termini of the chromosome and of a linear plasmid that were not assembled by PacBio. We propose an experimental pipeline that uses the Illumina assembled contigs, in addition to just the reads, to complement the current limitations of the PacBio sequencing technology and assembly software.


Talk with an expert

If you have a question, need to check the status of an order, or are interested in purchasing an instrument, we're here to help.