Menu
September 22, 2019

Genome-wide identification of simple sequence repeats and development of polymorphic SSR markers for genetic studies in tea plant (Camellia sinensis)

The tea plant (Camellia sinensis (L.) O. Kuntze) is one of the most popular non-alcoholic beverage crops worldwide. The availability of complete genome sequences for the Camellia sinensis var. ‘Shuchazao’ has provided the opportunity to identify all types of simple sequence repeat (SSR) markers by genome-wide scan. In this study, a total of 667,980 SSRs were identified in the ~?3.08 Gb genome, with an overall density of 216.88 SSRs/Mb. Dinucleotide repeats were predominant among microsatellites (72.25%), followed by trinucleotide repeats (15.35%), while the remaining SSRs accounted for less than 13%. The motif AG/CT (49.96%) and AT/TA (40.14%) were the most and the second most abundant among all identified SSR motifs, respectively; meanwhile, AAT/ATT (41.29%) and AAAT/ATTT (67.47%) were the most common among trinucleotides and tetranucleotides, respectively. A total of 300 primer pairs were designed to screen six tea cultivars for polymorphisms of SSR markers using the five selected repeat types of microsatellite sequences. The resulting 96 SSR markers that yielded polymorphic and unambiguous bands were further deployed on 47 tea cultivars for genetic diversity assessment, demonstrating high polymorphism of these SSR markers. Remarkably, the dendrogram revealed that the phylogenetic relationships among these tea cultivars are highly consistent with their genetic backgrounds or places of origin. The identified genome-wide SSRs and newly developed SSR markers will provide a powerful means for genetic researches in tea plant, including genetic diversity and evolutionary origin analysis, fingerprinting, QTL mapping, and marker-assisted selection for breeding.


September 22, 2019

Evaluation of WGS based approaches for investigating a food-borne outbreak caused by Salmonella enterica serovar Derby in Germany.

In Germany salmonellosis still represents the 2nd most common bacterial foodborne disease. The majority of infections are caused by Salmonella (S.) Typhimurium and S. Enteritidis followed by a variety of other broad host-range serovars. Salmonella Derby is one of the five top-ranked serovars isolated from humans and it represents one of the most prevalent serovars in pigs, thus bearing the potential risk for transmission to humans upon consumption of pig meat and products thereof. From November 2013 to January 2014 S. Derby caused a large outbreak that affected 145 primarily elderly people. Epidemiological investigations identified raw pork sausage as the probable source of infection, which was confirmed by microbiological evidence. During the outbreak isolates from patients, food specimen and asymptomatic carriers were investigated by conventional typing methods. However, the quantity and quality of available microbiological and epidemiological data made this outbreak highly suitable for retrospective investigation by Whole Genome Sequencing (WGS) and subsequent evaluation of different bioinformatics approaches for cluster definition. Overall the WGS-based methods confirmed the results of the conventional typing but were of significant higher discriminatory power. That was particularly beneficial for strains with incomplete epidemiological data. For our data set both, single nucleotide polymorphism (SNP)- and core genome multilocus sequence typing (cgMLST)-based methods proved to be appropriate tools for cluster definition. Copyright © 2017 Elsevier Ltd. All rights reserved.


September 22, 2019

The antibody loci of the domestic goat (Capra hircus).

The domestic goat (Capra hircus) is an important ruminant species both as a source of antibody-based reagents for research and biomedical applications and as an economically important animal for agriculture, particularly for developing nations that maintain most of the global goat population. Characterization of the loci encoding the goat immune repertoire would be highly beneficial for both vaccine and immune reagent development. However, in goat and other species whose reference genomes were generated using short-read sequencing technologies, the immune loci are poorly assembled as a result of their repetitive nature. Our recent construction of a long-read goat genome assembly (ARS1) has facilitated characterization of all three antibody loci with high confidence and comparative analysis to cattle. We observed broad similarity of goat and cattle antibody-encoding loci but with notable differences that likely influence formation of the functional antibody repertoire. The goat heavy-chain locus is restricted to only four functional and nearly identical IGHV genes, in contrast to the ten observed in cattle. Repertoire analysis indicates that light-chain usage is more balanced in goats, with greater representation of kappa light chains (~ 20-30%) compared to that in cattle (~ 5%). The present study represents the first characterization of the goat antibody loci and will help inform future investigations of their antibody responses to disease and vaccination.


September 22, 2019

Discovery of gorilla MHC-C expressing C1 ligand for KIR.

In comparison to humans and chimpanzees, gorillas show low diversity at MHC class I genes (Gogo), as reflected by an overall reduced level of allelic variation as well as the absence of a functionally important sequence motif that interacts with killer cell immunoglobulin-like receptors (KIR). Here, we use recently generated large-scale genomic sequence data for a reassessment of allelic diversity at Gogo-C, the gorilla orthologue of HLA-C. Through the combination of long-range amplifications and long-read sequencing technology, we obtained, among the 35 gorillas reanalyzed, three novel full-length genomic sequences including a coding region sequence that has not been previously described. The newly identified Gogo-C*03:01 allele has a divergent recombinant structure that sets it apart from other Gogo-C alleles. Domain-by-domain phylogenetic analysis shows that Gogo-C*03:01 has segments in common with Gogo-B*07, the additional B-like gene that is present on some gorilla MHC haplotypes. Identified in ~ 50% of the gorillas analyzed, the Gogo-C*03:01 allele exclusively encodes the C1 epitope among Gogo-C allotypes, indicating its important function in controlling natural killer cell (NK cell) responses via KIR. We further explored the hypothesis whether gorillas experienced a selective sweep which may have resulted in a general reduction of the gorilla MHC class I repertoire. Our results provide little support for a selective sweep but rather suggest that the overall low Gogo class I diversity can be best explained by drastic demographic changes gorillas experienced in the ancient and recent past.


September 22, 2019

Insights into platypus population structure and history from whole-genome sequencing.

The platypus is an egg-laying mammal which, alongside the echidna, occupies a unique place in the mammalian phylogenetic tree. Despite widespread interest in its unusual biology, little is known about its population structure or recent evolutionary history. To provide new insights into the dispersal and demographic history of this iconic species, we sequenced the genomes of 57 platypuses from across the whole species range in eastern mainland Australia and Tasmania. Using a highly improved reference genome, we called over 6.7?M SNPs, providing an informative genetic data set for population analyses. Our results show very strong population structure in the platypus, with our sampling locations corresponding to discrete groupings between which there is no evidence for recent gene flow. Genome-wide data allowed us to establish that 28 of the 57 sampled individuals had at least a third-degree relative among other samples from the same river, often taken at different times. Taking advantage of a sampled family quartet, we estimated the de novo mutation rate in the platypus at 7.0?×?10-9/bp/generation (95% CI 4.1?×?10-9-1.2?×?10-8/bp/generation). We estimated effective population sizes of ancestral populations and haplotype sharing between current groupings, and found evidence for bottlenecks and long-term population decline in multiple regions, and early divergence between populations in different regions. This study demonstrates the power of whole-genome sequencing for studying natural populations of an evolutionarily important species.


September 22, 2019

Spread of plasmid-encoded NDM-1 and GES-5 carbapenemases among extensively drug-resistant and pandrug-resistant clinical Enterobacteriaceae in Durban, South Africa.

Whole-genome sequence analyses revealed the presence of blaNDM-1 (n = 31), blaGES-5 (n = 8), blaOXA-232 (n = 1), or blaNDM-5 (n = 1) in extensively drug-resistant and pandrug-resistant Enterobacteriaceae organisms isolated from in-patients in 10 private hospitals (2012 to 2013) in Durban, South Africa. Two novel NDM-1-encoding plasmids from Klebsiella pneumoniae were circularized by PacBio sequencing. In p19-10_01 [IncFIB(K); 223.434 bp], blaNDM-1 was part of a Tn1548-like structure (16.276 bp) delineated by IS26 The multireplicon plasmid p18-43_01 [IncR_1/IncFIB(pB171)/IncFII(Yp); 212.326 bp] shared an 80-kb region with p19-10_01, not including the blaNDM-1-containing region. The two plasmids were used as references for tracing NDM-1-encoding plasmids in the other genome assemblies. The p19-10_01 sequence was detected in K. pneumoniae (n = 7) only, whereas p18-43_01 was tracked to K. pneumoniae (n = 4), Klebsiella michiganensis (n = 1), Serratia marcescens (n = 11), Enterobacter spp. (n = 7), and Citrobacter freundii (n = 1), revealing horizontal spread of this blaNDM-1-bearing plasmid structure. Global phylogeny showed clustering of the K. pneumoniae (18/20) isolates together with closely related carbapenemase-negative ST101 isolates from other geographical origins. The South African isolates were divided into three phylogenetic subbranches, where each group had distinct resistance and replicon profiles, carrying either p19-10_01, p18-10_01, or pCHE-A1 (8,201 bp). The latter plasmid carried blaGES-5 and aacA4 within an integron mobilization unit. Our findings imply independent plasmid acquisition followed by local dissemination. Additionally, we detected blaOXA-232 carried by pPKPN4 in K. pneumoniae (ST14) and blaNDM-5 contained by a pNDM-MGR194-like genetic structure in Escherichia coli (ST167), adding even more complexity to the multilayer molecular mechanisms behind nosocomial spread of carbapenem-resistant Enterobacteriaceae in Durban, South Africa. Copyright © 2018 American Society for Microbiology.


September 22, 2019

CagY-dependent regulation of type IV secretion in Helicobacter pylori is associated with alterations in integrin binding.

Strains of Helicobacter pylori that cause ulcer or gastric cancer typically express a type IV secretion system (T4SS) encoded by the cag pathogenicity island (cagPAI). CagY is an ortholog of VirB10 that, unlike other VirB10 orthologs, has a large middle repeat region (MRR) with extensive repetitive sequence motifs, which undergo CD4+ T cell-dependent recombination during infection of mice. Recombination in the CagY MRR reduces T4SS function, diminishes the host inflammatory response, and enables the bacteria to colonize at a higher density. Since CagY is known to bind human a5ß1 integrin, we tested the hypothesis that recombination in the CagY MRR regulates T4SS function by modulating binding to a5ß1 integrin. Using a cell-free microfluidic assay, we found that H. pylori binding to a5ß1 integrin under shear flow is dependent on the CagY MRR, but independent of the presence of the T4SS pili, which are only formed when H. pylori is in contact with host cells. Similarly, expression of CagY in the absence of other T4SS genes was necessary and sufficient for whole bacterial cell binding to a5ß1 integrin. Bacteria with variant cagY alleles that reduced T4SS function showed comparable reduction in binding to a5ß1 integrin, although CagY was still expressed on the bacterial surface. We speculate that cagY-dependent modulation of H. pylori T4SS function is mediated by alterations in binding to a5ß1 integrin, which in turn regulates the host inflammatory response so as to maximize persistent infection.IMPORTANCE Infection with H. pylori can cause peptic ulcers and is the most important risk factor for gastric cancer, the third most common cause of cancer death worldwide. The major H. pylori virulence factor that determines whether infection causes disease or asymptomatic colonization is the type IV secretion system (T4SS), a sort of molecular syringe that injects bacterial products into gastric epithelial cells and alters host cell physiology. We previously showed that recombination in CagY, an essential T4SS component, modulates the function of the T4SS. Here we found that these recombination events produce parallel changes in specific binding to a5ß1 integrin, a host cell receptor that is essential for T4SS-dependent translocation of bacterial effectors. We propose that CagY-dependent binding to a5ß1 integrin acts like a molecular rheostat that alters T4SS function and modulates the host immune response to promote persistent infection. Copyright © 2018 Skoog et al.


September 22, 2019

Characterization of phenotypic variation and genome aberrations observed among Phytophthora ramorum isolates from diverse hosts.

Accumulating evidence suggests that genome plasticity allows filamentous plant pathogens to adapt to changing environments. Recently, the generalist plant pathogen Phytophthora ramorum has been documented to undergo irreversible phenotypic alterations accompanied by chromosomal aberrations when infecting trunks of mature oak trees (genus Quercus). In contrast, genomes and phenotypes of the pathogen derived from the foliage of California bay (Umbellularia californica) are usually stable. We define this phenomenon as host-induced phenotypic diversification (HIPD). P. ramorum also causes a severe foliar blight in some ornamental plants such as Rhododendron spp. and Viburnum spp., and isolates from these hosts occasionally show phenotypes resembling those from oak trunks that carry chromosomal aberrations. The aim of this study was to investigate variations in phenotypes and genomes of P. ramorum isolates from non-oak hosts and substrates to determine whether HIPD changes may be equivalent to those among isolates from oaks.We analyzed genomes of diverse non-oak isolates including those taken from foliage of Rhododendron and other ornamental plants, as well as from natural host species, soil, and water. Isolates recovered from artificially inoculated oak logs were also examined. We identified diverse chromosomal aberrations including copy neutral loss of heterozygosity (cnLOH) and aneuploidy in isolates from non-oak hosts. Most identified aberrations in non-oak hosts were also common among oak isolates; however, trisomy, a frequent type of chromosomal aberration in oak isolates was not observed in isolates from Rhododendron.This work cross-examined phenotypic variation and chromosomal aberrations in P. ramorum isolates from oak and non-oak hosts and substrates. The results suggest that HIPD comparable to that occurring in oak hosts occurs in non-oak environments such as in Rhododendron leaves. Rhododendron leaves are more easily available than mature oak stems and thus can potentially serve as a model host for the investigation of HIPD, the newly described plant-pathogen interaction.


September 22, 2019

NextSV: a meta-caller for structural variants from low-coverage long-read sequencing data.

Structural variants (SVs) in human genomes are implicated in a variety of human diseases. Long-read sequencing delivers much longer read lengths than short-read sequencing and may greatly improve SV detection. However, due to the relatively high cost of long-read sequencing, it is unclear what coverage is needed and how to optimally use the aligners and SV callers.In this study, we developed NextSV, a meta-caller to perform SV calling from low coverage long-read sequencing data. NextSV integrates three aligners and three SV callers and generates two integrated call sets (sensitive/stringent) for different analysis purposes. We evaluated SV calling performance of NextSV under different PacBio coverages on two personal genomes, NA12878 and HX1. Our results showed that, compared with running any single SV caller, NextSV stringent call set had higher precision and balanced accuracy (F1 score) while NextSV sensitive call set had a higher recall. At 10X coverage, the recall of NextSV sensitive call set was 93.5 to 94.1% for deletions and 87.9 to 93.2% for insertions, indicating that ~10X coverage might be an optimal coverage to use in practice, considering the balance between the sequencing costs and the recall rates. We further evaluated the Mendelian errors on an Ashkenazi Jewish trio dataset.Our results provide useful guidelines for SV detection from low coverage whole-genome PacBio data and we expect that NextSV will facilitate the analysis of SVs on long-read sequencing data.


September 22, 2019

Multiple convergent supergene evolution events in mating-type chromosomes.

Convergent adaptation provides unique insights into the predictability of evolution and ultimately into processes of biological diversification. Supergenes (beneficial gene linkage) are striking examples of adaptation, but little is known about their prevalence or evolution. A recent study on anther-smut fungi documented supergene formation by rearrangements linking two key mating-type loci, controlling pre- and post-mating compatibility. Here further high-quality genome assemblies reveal four additional independent cases of chromosomal rearrangements leading to regions of suppressed recombination linking these mating-type loci in closely related species. Such convergent transitions in genomic architecture of mating-type determination indicate strong selection favoring linkage of mating-type loci into cosegregating supergenes. We find independent evolutionary strata (stepwise recombination suppression) in several species, with extensive rearrangements, gene losses, and transposable element accumulation. We thus show remarkable convergence in mating-type chromosome evolution, recurrent supergene formation, and repeated evolution of similar phenotypes through different genomic changes.


September 22, 2019

Mutant phenotypes for thousands of bacterial genes of unknown function.

One-third of all protein-coding genes from bacterial genomes cannot be annotated with a function. Here, to investigate the functions of these genes, we present genome-wide mutant fitness data from 32 diverse bacteria across dozens of growth conditions. We identified mutant phenotypes for 11,779 protein-coding genes that had not been annotated with a specific function. Many genes could be associated with a specific condition because the gene affected fitness only in that condition, or with another gene in the same bacterium because they had similar mutant phenotypes. Of the poorly annotated genes, 2,316 had associations that have high confidence because they are conserved in other bacteria. By combining these conserved associations with comparative genomics, we identified putative DNA repair proteins; in addition, we propose specific functions for poorly annotated enzymes and transporters and for uncharacterized protein families. Our study demonstrates the scalability of microbial genetics and its utility for improving gene annotations.


September 22, 2019

Genome-wide analysis of Mycoplasma bovirhinis GS01 reveals potential virulence factors and phylogenetic relationships.

Mycoplasma bovirhinis is a significant etiology in bovine pneumonia and mastitis, but our knowledge about the genetic and pathogenic mechanisms of M. bovirhinis is very limited. In this study, we sequenced the complete genome of M. bovirhinis strain GS01 isolated from the nasal swab of pneumonic calves in Gansu, China, and we found that its genome forms a 847,985 bp single circular chromosome with a GC content of 27.57% and with 707 protein-coding genes. The putative virulence determinants of M. bovirhinis were then analyzed. Results showed that three genomic islands and 16 putative virulence genes, including one adhesion gene enolase, seven surface lipoproteins, proteins involved in glycerol metabolism, and cation transporters, might be potential virulence factors. Glycerol and pyruvate metabolic pathways were defective. Comparative analysis revealed remarkable genome variations between GS01 and a recently reported HAZ141_2 strain, and extremely low homology with others mycoplasma species. Phylogenetic analysis demonstrated that M. bovirhinis was most genetically close to M. canis, distant from other bovine Mycoplasma species. Genomic dissection may provide useful information on the pathogenic mechanisms and genetics of M. bovirhinis. Copyright © 2018 Chen et al.


September 22, 2019

Whole genome analysis reveals the diversity and evolutionary relationships between necrotic enteritis-causing strains of Clostridium perfringens.

Clostridium perfringens causes a range of diseases in animals and humans including necrotic enteritis in chickens and food poisoning and gas gangrene in humans. Necrotic enteritis is of concern in commercial chicken production due to the cost of the implementation of infection control measures and to productivity losses. This study has focused on the genomic analysis of a range of chicken-derived C. perfringens isolates, from around the world and from different years. The genomes were sequenced and compared with 20 genomes available from public databases, which were from a diverse collection of isolates from chickens, other animals, and humans. We used a distance based phylogeny that was constructed based on gene content rather than sequence identity. Similarity between strains was defined as the number of genes that they have in common divided by their total number of genes. In this type of phylogenetic analysis, evolutionary distance can be interpreted in terms of evolutionary events such as acquisition and loss of genes, whereas the underlying properties (the gene content) can be interpreted in terms of function. We also compared these methods to the sequence-based phylogeny of the core genome.Distinct pathogenic clades of necrotic enteritis-causing C. perfringens were identified. They were characterised by variable regions encoded on the chromosome, with predicted roles in capsule production, adhesion, inhibition of related strains, phage integration, and metabolism. Some strains have almost identical genomes, even though they were isolated from different geographic regions at various times, while other highly distant genomes appear to result in similar outcomes with regard to virulence and pathogenesis.The high level of diversity in chicken isolates suggests there is no reliable factor that defines a chicken strain of C. perfringens, however, disease-causing strains can be defined by the presence of netB-encoding plasmids. This study reveals that horizontal gene transfer appears to play a significant role in genetic variation of the C. perfringens chromosome as well as the plasmid content within strains.


September 22, 2019

A whole genome assembly of the horn fly, Haematobia irritans, and prediction of genes with roles in metabolism and sex determination.

Haematobia irritans, commonly known as the horn fly, is a globally distributed blood-feeding pest of cattle that is responsible for significant economic losses to cattle producers. Chemical insecticides are the primary means for controlling this pest but problems with insecticide resistance have become common in the horn fly. To provide a foundation for identification of genomic loci for insecticide resistance and for discovery of new control technology, we report the sequencing, assembly, and annotation of the horn fly genome. The assembled genome is 1.14 Gb, comprising 76,616 scaffolds with N50 scaffold length of 23 Kb. Using RNA-Seq data, we have predicted 34,413 gene models of which 19,185 have been assigned functional annotations. Comparative genomics analysis with the Dipteran flies Musca domestica L., Drosophila melanogaster, and Lucilia cuprina, show that the horn fly is most closely related to M. domestica, sharing 8,748 orthologous clusters followed by D. melanogaster and L. cuprina, sharing 7,582 and 7,490 orthologous clusters respectively. We also identified a gene locus for the sodium channel protein in which mutations have been previously reported that confers target site resistance to the most common class of pesticides used in fly control. Additionally, we identified 276 genomic loci encoding members of metabolic enzyme gene families such as cytochrome P450s, esterases and glutathione S-transferases, and several genes orthologous to sex determination pathway genes in other Dipteran species. Copyright © 2018 Konganti et al.


September 22, 2019

Whole-genome analysis of three yeast strains used for production of sherry-like wines revealed genetic traits specific to Flor yeasts.

Flor yeast strains represent a specialized group of Saccharomyces cerevisiae yeasts used for biological wine aging. We have sequenced the genomes of three flor strains originated from different geographic regions and used for production of sherry-like wines in Russia. According to the obtained phylogeny of 118 yeast strains, flor strains form very tight cluster adjacent to the main wine clade. SNP analysis versus available genomes of wine and flor strains revealed 2,270 genetic variants in 1,337 loci specific to flor strains. Gene ontology analysis in combination with gene content evaluation revealed a complex landscape of possibly adaptive genetic changes in flor yeast, related to genes associated with cell morphology, mitotic cell cycle, ion homeostasis, DNA repair, carbohydrate metabolism, lipid metabolism, and cell wall biogenesis. Pangenomic analysis discovered the presence of several well-known “non-reference” loci of potential industrial importance. Events of gene loss included deletions of asparaginase genes, maltose utilization locus, and FRE-FIT locus involved in iron transport. The latter in combination with a flor-yeast-specific mutation in the Aft1 transcription factor gene is likely to be responsible for the discovered phenotype of increased iron sensitivity and improved iron uptake of analyzed strains. Expansion of the coding region of the FLO11 flocullin gene and alteration of the balance between members of the FLO gene family are likely to positively affect the well-known propensity of flor strains for velum formation. Our study provides new insights in the nature of genetic variation in flor yeast strains and demonstrates that different adaptive properties of flor yeast strains could have evolved through different mechanisms of genetic variation.


Talk with an expert

If you have a question, need to check the status of an order, or are interested in purchasing an instrument, we're here to help.