Menu
September 22, 2019

Comparison of phasing strategies for whole human genomes.

Humans are a diploid species that inherit one set of chromosomes paternally and one homologous set of chromosomes maternally. Unfortunately, most human sequencing initiatives ignore this fact in that they do not directly delineate the nucleotide content of the maternal and paternal copies of the 23 chromosomes individuals possess (i.e., they do not ‘phase’ the genome) often because of the costs and complexities of doing so. We compared 11 different widely-used approaches to phasing human genomes using the publicly available ‘Genome-In-A-Bottle’ (GIAB) phased version of the NA12878 genome as a gold standard. The phasing strategies we compared included laboratory-based assays that prepare DNA in unique ways to facilitate phasing as well as purely computational approaches that seek to reconstruct phase information from general sequencing reads and constructs or population-level haplotype frequency information obtained through a reference panel of haplotypes. To assess the performance of the 11 approaches, we used metrics that included, among others, switch error rates, haplotype block lengths, the proportion of fully phase-resolved genes, phasing accuracy and yield between pairs of SNVs. Our comparisons suggest that a hybrid or combined approach that leverages: 1. population-based phasing using the SHAPEIT software suite, 2. either genome-wide sequencing read data or parental genotypes, and 3. a large reference panel of variant and haplotype frequencies, provides a fast and efficient way to produce highly accurate phase-resolved individual human genomes. We found that for population-based approaches, phasing performance is enhanced with the addition of genome-wide read data; e.g., whole genome shotgun and/or RNA sequencing reads. Further, we found that the inclusion of parental genotype data within a population-based phasing strategy can provide as much as a ten-fold reduction in phasing errors. We also considered a majority voting scheme for the construction of a consensus haplotype combining multiple predictions for enhanced performance and site coverage. Finally, we also identified DNA sequence signatures associated with the genomic regions harboring phasing switch errors, which included regions of low polymorphism or SNV density.


September 22, 2019

Massive lateral transfer of genes encoding plant cell wall-degrading enzymes to the mycoparasitic fungus Trichoderma from its plant-associated hosts.

Unlike most other fungi, molds of the genus Trichoderma (Hypocreales, Ascomycota) are aggressive parasites of other fungi and efficient decomposers of plant biomass. Although nutritional shifts are common among hypocrealean fungi, there are no examples of such broad substrate versatility as that observed in Trichoderma. A phylogenomic analysis of 23 hypocrealean fungi (including nine Trichoderma spp. and the related Escovopsis weberi) revealed that the genus Trichoderma has evolved from an ancestor with limited cellulolytic capability that fed on either fungi or arthropods. The evolutionary analysis of Trichoderma genes encoding plant cell wall-degrading carbohydrate-active enzymes and auxiliary proteins (pcwdCAZome, 122 gene families) based on a gene tree / species tree reconciliation demonstrated that the formation of the genus was accompanied by an unprecedented extent of lateral gene transfer (LGT). Nearly one-half of the genes in Trichoderma pcwdCAZome (41%) were obtained via LGT from plant-associated filamentous fungi belonging to different classes of Ascomycota, while no LGT was observed from other potential donors. In addition to the ability to feed on unrelated fungi (such as Basidiomycota), we also showed that Trichoderma is capable of endoparasitism on a broad range of Ascomycota, including extant LGT donors. This phenomenon was not observed in E. weberi and rarely in other mycoparasitic hypocrealean fungi. Thus, our study suggests that LGT is linked to the ability of Trichoderma to parasitize taxonomically related fungi (up to adelphoparasitism in strict sense). This may have allowed primarily mycotrophic Trichoderma fungi to evolve into decomposers of plant biomass.


September 22, 2019

Transposable element genomic fissuring in Pyrenophora teres is associated with genome expansion and dynamics of host-pathogen genetic interactions.

Pyrenophora teres, P. teres f. teres (PTT) and P. teres f. maculata (PTM) cause significant diseases in barley, but little is known about the large-scale genomic differences that may distinguish the two forms. Comprehensive genome assemblies were constructed from long DNA reads, optical and genetic maps. As repeat masking in fungal genomes influences the final gene annotations, an accurate and reproducible pipeline was developed to ensure comparability between isolates. The genomes of the two forms are highly collinear, each composed of 12 chromosomes. Genome evolution in P. teres is characterized by genome fissuring through the insertion and expansion of transposable elements (TEs), a process that isolates blocks of genic sequence. The phenomenon is particularly pronounced in PTT, which has a larger, more repetitive genome than PTM and more recent transposon activity measured by the frequency and size of genome fissures. PTT has a longer cultivated host association and, notably, a greater range of host-pathogen genetic interactions compared to other Pyrenophora spp., a property which associates better with genome size than pathogen lifestyle. The two forms possess similar complements of TE families with Tc1/Mariner and LINE-like Tad-1 elements more abundant in PTT. Tad-1 was only detectable as vestigial fragments in PTM and, within the forms, differences in genome sizes and the presence and absence of several TE families indicated recent lineage invasions. Gene differences between P. teres forms are mainly associated with gene-sparse regions near or within TE-rich regions, with many genes possessing characteristics of fungal effectors. Instances of gene interruption by transposons resulting in pseudogenization were detected in PTT. In addition, both forms have a large complement of secondary metabolite gene clusters indicating significant capacity to produce an array of different molecules. This study provides genomic resources for functional genetics to help dissect factors underlying the host-pathogen interactions.


September 22, 2019

Genome evolution across 1,011 Saccharomyces cerevisiae isolates.

Large-scale population genomic surveys are essential to explore the phenotypic diversity of natural populations. Here we report the whole-genome sequencing and phenotyping of 1,011 Saccharomyces cerevisiae isolates, which together provide an accurate evolutionary picture of the genomic variants that shape the species-wide phenotypic landscape of this yeast. Genomic analyses support a single ‘out-of-China’ origin for this species, followed by several independent domestication events. Although domesticated isolates exhibit high variation in ploidy, aneuploidy and genome content, genome evolution in wild isolates is mainly driven by the accumulation of single nucleotide polymorphisms. A common feature is the extensive loss of heterozygosity, which represents an essential source of inter-individual variation in this mainly asexual species. Most of the single nucleotide polymorphisms, including experimentally identified functional polymorphisms, are present at very low frequencies. The largest numbers of variants identified by genome-wide association are copy-number changes, which have a greater phenotypic effect than do single nucleotide polymorphisms. This resource will guide future population genomics and genotype-phenotype studies in this classic model system.


September 22, 2019

SvABA: genome-wide detection of structural variants and indels by local assembly.

Structural variants (SVs), including small insertion and deletion variants (indels), are challenging to detect through standard alignment-based variant calling methods. Sequence assembly offers a powerful approach to identifying SVs, but is difficult to apply at scale genome-wide for SV detection due to its computational complexity and the difficulty of extracting SVs from assembly contigs. We describe SvABA, an efficient and accurate method for detecting SVs from short-read sequencing data using genome-wide local assembly with low memory and computing requirements. We evaluated SvABA’s performance on the NA12878 human genome and in simulated and real cancer genomes. SvABA demonstrates superior sensitivity and specificity across a large spectrum of SVs and substantially improves detection performance for variants in the 20-300 bp range, compared with existing methods. SvABA also identifies complex somatic rearrangements with chains of short (<1000 bp) templated-sequence insertions copied from distant genomic regions. We applied SvABA to 344 cancer genomes from 11 cancer types and found that short templated-sequence insertions occur in ~4% of all somatic rearrangements. Finally, we demonstrate that SvABA can identify sites of viral integration and cancer driver alterations containing medium-sized (50-300 bp) SVs.© 2018 Wala et al.; Published by Cold Spring Harbor Laboratory Press.


September 22, 2019

Evolution of sequence type 4821 clonal complex meningococcal strains in China from prequinolone to quinolone era, 1972-2013.

The expansion of hypervirulent sequence type 4821 clonal complex (CC4821) lineage Neisseria meningitidis bacteria has led to a shift in meningococcal disease epidemiology in China, from serogroup A (MenA) to MenC. Knowledge of the evolution and genetic origin of the emergent MenC strains is limited. In this study, we subjected 76 CC4821 isolates collected across China during 1972-1977 and 2005-2013 to phylogenetic analysis, traditional genotyping, or both. We show that successive recombination events within genes encoding surface antigens and acquisition of quinolone resistance mutations possibly played a role in the emergence of CC4821 as an epidemic clone in China. MenC and MenB CC4821 strains have spread across China and have been detected in several countries in different continents. Capsular switches involving serogroups B and C occurred among epidemic strains, raising concerns regarding possible increases in MenB disease, given that vaccines in use in China do not protect against MenB.


September 22, 2019

Complete genome sequences of seven Vibrio anguillarum strains as derived from PacBio sequencing.

We report here the complete genome sequences of seven Vibrio anguillarum strains isolated from multiple geographic locations, thus increasing the total number of genomes of finished quality to 11. The genomes were de novo assembled from long-sequence PacBio reads. Including draft genomes, a total of 44?V. anguillarum genomes are currently available in the genome databases. They represent an important resource in the study of, for example, genetic variations and for identifying virulence determinants. In this article, we present the genomes and basic genome comparisons of the 11 complete genomes, including a BRIG analysis, and pan genome calculation. We also describe some structural features of superintegrons on chromosome 2?s, and associated insertion sequence (IS) elements, including 18 new ISs (ISVa3?-?ISVa20), both of importance in the complement of V. anguillarum genomes.


September 22, 2019

Developing collaborative works for faster progress on fungal respiratory infections in cystic fibrosis.

Cystic fibrosis (CF) is the major genetic inherited disease in Caucasian populations. The respiratory tract of CF patients displays a sticky viscous mucus, which allows for the entrapment of airborne bacteria and fungal spores and provides a suitable environment for growth of microorganisms, including numerous yeast and filamentous fungal species. As a consequence, respiratory infections are the major cause of morbidity and mortality in this clinical context. Although bacteria remain the most common agents of these infections, fungal respiratory infections have emerged as an important cause of disease. Therefore, the International Society for Human and Animal Mycology (ISHAM) has launched a working group on Fungal respiratory infections in Cystic Fibrosis (Fri-CF) in October 2006, which was subsequently approved by the European Confederation of Medical Mycology (ECMM). Meetings of this working group, comprising both clinicians and mycologists involved in the follow-up of CF patients, as well as basic scientists interested in the fungal species involved, provided the opportunity to initiate collaborative works aimed to improve our knowledge on these infections to assist clinicians in patient management. The current review highlights the outcomes of some of these collaborative works in clinical surveillance, pathogenesis and treatment, giving special emphasis to standardization of culture procedures, improvement of species identification methods including the development of nonculture-based diagnostic methods, microbiome studies and identification of new biological markers, and the description of genotyping studies aiming to differentiate transient carriage and chronic colonization of the airways. The review also reports on the breakthrough in sequencing the genomes of the main Scedosporium species as basis for a better understanding of the pathogenic mechanisms of these fungi, and discusses treatment options of infections caused by multidrug resistant microorganisms, such as Scedosporium and Lomentospora species and members of the Rasamsonia argillacea species complex.


September 22, 2019

Haemophilus influenzae genome evolution during persistence in the human airways in chronic obstructive pulmonary disease.

Nontypeable Haemophilus influenzae (NTHi) exclusively colonize and infect humans and are critical to the pathogenesis of chronic obstructive pulmonary disease (COPD). In vitro and animal models do not accurately capture the complex environments encountered by NTHi during human infection. We conducted whole-genome sequencing of 269 longitudinally collected cleared and persistent NTHi from a 15-y prospective study of adults with COPD. Genome sequences were used to elucidate the phylogeny of NTHi isolates, identify genomic changes that occur with persistence in the human airways, and evaluate the effect of selective pressure on 12 candidate vaccine antigens. Strains persisted in individuals with COPD for as long as 1,422 d. Slipped-strand mispairing, mediated by changes in simple sequence repeats in multiple genes during persistence, regulates expression of critical virulence functions, including adherence, nutrient uptake, and modification of surface molecules, and is a major mechanism for survival in the hostile environment of the human airways. A subset of strains underwent a large 400-kb inversion during persistence. NTHi does not undergo significant gene gain or loss during persistence, in contrast to other persistent respiratory tract pathogens. Amino acid sequence changes occurred in 8 of 12 candidate vaccine antigens during persistence, an observation with important implications for vaccine development. These results indicate that NTHi alters its genome during persistence by regulation of critical virulence functions primarily by slipped-strand mispairing, advancing our understanding of how a bacterial pathogen that plays a critical role in COPD adapts to survival in the human respiratory tract.


September 22, 2019

Heterologous expression guides identification of the biosynthetic gene cluster of chuangxinmycin, an indole alkaloid antibiotic.

The indole alkaloid antibiotic chuangxinmycin, from Actinobacteria Actinoplanes tsinanensis, containing a unique thiopyrano[4,3,2- cd]indole scaffold, is a potent and selective inhibitor of bacterial tryptophanyl-tRNA synthetase. The chuangxinmycin biosynthetic gene cluster was identified by in silico analysis of the genome sequence, then verified by heterologous expression. Systemic gene inactivation and intermediate identification determined the minimum set of genes for unique thiopyrano[4,3,2- cd]indole formation and the concerted action of a radical S-adenosylmethionine protein plus an unknown protein for addition of the 3-methyl group. These findings set a solid foundation for comprehensively investigating the biosynthesis, optimizing yield, and generating new analogues of chuangxinmycin.


September 22, 2019

Distinct evolutionary patterns of Neisseria meningitidis serogroup B disease outbreaks at two universities in the USA.

Neisseria meningitidis serogroup B (MnB) was responsible for two independent meningococcal disease outbreaks at universities in the USA during 2013. The first at University A in New Jersey included nine confirmed cases reported between March 2013 and March 2014. The second outbreak occurred at University B in California, with four confirmed cases during November 2013. The public health response to these outbreaks included the approval and deployment of a serogroup B meningococcal vaccine that was not yet licensed in the USA. This study investigated the use of whole-genome sequencing(WGS) to examine the genetic profile of the disease-causing outbreak isolates at each university. Comparative WGS revealed differences in evolutionary patterns between the two disease outbreaks. The University A outbreak isolates were very closely related, with differences primarily attributed to single nucleotide polymorphisms/insertion-deletion (SNP/indel) events. In contrast, the University B outbreak isolates segregated into two phylogenetic clades, differing in large part due to recombination events covering extensive regions (>30?kb) of the genome including virulence factors. This high-resolution comparison of two meningococcal disease outbreaks further demonstrates the genetic complexity of meningococcal bacteria as related to evolution and disease virulence.


September 22, 2019

Somatic hypermutation of T cell receptor a chain contributes to selection in nurse shark thymus.

Since the discovery of the T cell receptor (TcR), immunologists have assigned somatic hypermutation (SHM) as a mechanism employed solely by B cells to diversify their antigen receptors. Remarkably, we found SHM acting in the thymus on a chain locus of shark TcR. SHM in developing shark T cells likely is catalyzed by activation-induced cytidine deaminase (AID) and results in both point and tandem mutations that accumulate non-conservative amino acid replacements within complementarity-determining regions (CDRs). Mutation frequency at TcRa was as high as that seen at B cell receptor loci (BcR) in sharks and mammals, and the mechanism of SHM shares unique characteristics first detected at shark BcR loci. Additionally, fluorescence in situ hybridization showed the strongest AID expression in thymic corticomedullary junction and medulla. We suggest that TcRa utilizes SHM to broaden diversification of the primary aß T cell repertoire in sharks, the first reported use in vertebrates.© 2018, Ott et al.


September 22, 2019

Multi-omics approach identifies novel pathogen-derived prognostic biomarkers in patients with Pseudomonas aeruginosa bloodstream infection

Pseudomonas aeruginosa is a human pathogen that causes health-care associated blood stream infections (BSI). Although P. aeruginosa BSI are associated with high mortality rates, the clinical relevance of pathogen-derived prognostic biomarker to identify patients at risk for unfavorable outcome remains largely unexplored. We found novel pathogen-derived prognostic biomarker candidates by applying a multi-omics approach on a multicenter sepsis patient cohort. Multi-level Cox regression was used to investigate the relation between patient characteristics and pathogen features (2298 accessory genes, 1078 core protein levels, 107 parsimony-informative variations in reported virulence factors) with 30-day mortality. Our analysis revealed that presence of the helP gene encoding a putative DEAD-box helicase was independently associated with a fatal outcome (hazard ratio 2.01, p = 0.05). helP is located within a region related to the pathogenicity island PAPI-1 in close proximity to a pil gene cluster, which has been associated with horizontal gene transfer. Besides helP, elevated protein levels of the bacterial flagellum protein FliL (hazard ratio 3.44, p < 0.001) and of a bacterioferritin-like protein (hazard ratio 1.74, p = 0.003) increased the risk of death, while high protein levels of a putative aminotransferase were associated with an improved outcome (hazard ratio 0.12, p < 0.001). The prognostic potential of biomarker candidates and clinical factors was confirmed with different machine learning approaches using training and hold-out datasets. The helP genotype appeared the most attractive biomarker for clinical risk stratification due to its relevant predictive power and ease of detection.


September 22, 2019

Comparative genomics of bdelloid rotifers: Insights from desiccating and nondesiccating species.

Bdelloid rotifers are a class of microscopic invertebrates that have existed for millions of years apparently without sex or meiosis. They inhabit a variety of temporary and permanent freshwater habitats globally, and many species are remarkably tolerant of desiccation. Bdelloids offer an opportunity to better understand the evolution of sex and recombination, but previous work has emphasised desiccation as the cause of several unusual genomic features in this group. Here, we present high-quality whole-genome sequences of 3 bdelloid species: Rotaria macrura and R. magnacalcarata, which are both desiccation intolerant, and Adineta ricciae, which is desiccation tolerant. In combination with the published assembly of A. vaga, which is also desiccation tolerant, we apply a comparative genomics approach to evaluate the potential effects of desiccation tolerance and asexuality on genome evolution in bdelloids. We find that ancestral tetraploidy is conserved among all 4 bdelloid species, but homologous divergence in obligately aquatic Rotaria genomes is unexpectedly low. This finding is contrary to current models regarding the role of desiccation in shaping bdelloid genomes. In addition, we find that homologous regions in A. ricciae are largely collinear and do not form palindromic repeats as observed in the published A. vaga assembly. Consequently, several features interpreted as genomic evidence for long-term ameiotic evolution are not general to all bdelloid species, even within the same genus. Finally, we substantiate previous findings of high levels of horizontally transferred nonmetazoan genes in both desiccating and nondesiccating bdelloid species and show that this unusual feature is not shared by other animal phyla, even those with desiccation-tolerant representatives. These comparisons call into question the proposed role of desiccation in mediating horizontal genetic transfer.


September 22, 2019

Genome-wide comparison reveals a probiotic strain Lactococcus lactis WFLU12 isolated from the gastrointestinal tract of olive flounder (Paralichthys Olivaceus) harboring genes supporting probiotic action.

Our previous study has shown that dietary supplementation with Lactococcus lactis WFLU12 can enhance the growth of olive flounder and its resistance against streptococcal infection. The objective of the present study was to use comparative genomics tools to investigate genomic characteristics of strain WFLU12 and the presence of genes supporting its probiotic action using sequenced genomes of L. lactis strains. Dispensable and singleton genes of strain WFLU12 were found to be more enriched in genes associated with metabolism (e.g., energy production and conversion, and carbohydrate transport and metabolism) than pooled dispensable and singleton genes in other L. lactis strains, reflecting WFLU12 strain-specific ecosystem origin and its ability to metabolize different energy sources. Strain WFLU12 produced antimicrobial compounds that could inhibit several bacterial fish pathogens. It possessed the nisin gene cluster (nisZBTCIPRKFEG) and genes encoding lysozyme and colicin V. However, only three other strains (CV56, IO-1, and SO) harbor a complete nisin gene cluster. We also found that L. lactis WFLU12 possessed many other important functional genes involved in stress responses to the gastrointestinal tract environment, dietary energy extraction, and metabolism to support the probiotic action of this strain found in our previous study. This strongly indicates that not all L. lactis strains can be used as probiotics. This study highlights comparative genomics approaches as very useful and powerful tools to select probiotic candidates and predict their probiotic effects.


Talk with an expert

If you have a question, need to check the status of an order, or are interested in purchasing an instrument, we're here to help.