Menu
April 21, 2020

Population Genomic Analysis and De Novo Assembly Reveal the Origin of Weedy Rice as an Evolutionary Game.

Crop weediness, especially that of weedy rice (Oryza sativa f. spontanea), remains mysterious. Weedy rice possesses robust ecological adaptability; however, how this strain originated and gradually formed proprietary genetic features remains unclear. Here, we demonstrate that weedy rice at Asian high latitudes (WRAH) is phylogenetically well defined and possesses unselected genomic characteristics in many divergence regions between weedy and cultivated rice. We also identified novel quantitative trait loci underlying weedy-specific traits, and revealed that a genome block on the end of chromosome 1 is associated with rice weediness. To identify the genomic modifications underlying weedy rice evolution, we generated the first de novo assembly of a high-quality weedy rice genome (WR04-6), and conducted a comparative genomics study between WR04-6 with other rice reference genomes. Multiple lines of evidence, including the results of demographic scenario comparisons, suggest that differentiation between weedy rice and cultivated rice was initiated by genetic improvement of cultivated rice and that the essence of weediness arose through semi-domestication. A plant height model further implied that the origin of WRAH can be modeled as an evolutionary game and indicated that strategy-based selection driven by fitness shaped its genomic diversity.Copyright © 2019 The Authors. Published by Elsevier Inc. All rights reserved.


April 21, 2020

De novo assembly of white poplar genome and genetic diversity of white poplar population in Irtysh River basin in China.

The white poplar (Populus alba) is widely distributed in Central Asia and Europe. There are natural populations of white poplar in Irtysh River basin in China. It also can be cultivated and grown well in northern China. In this study, we sequenced the genome of P. alba by single-molecule real-time technology. De novo assembly of P. alba had a genome size of 415.99 Mb with a contig N50 of 1.18 Mb. A total of 32,963 protein-coding genes were identified. 45.16% of the genome was annotated as repetitive elements. Genome evolution analysis revealed that divergence between P. alba and Populus trichocarpa (black cottonwood) occurred ~5.0 Mya (3.0, 7.1). Fourfold synonymous third-codon transversion (4DTV) and synonymous substitution rate (ks) distributions supported the occurrence of the salicoid WGD event (~ 65 Mya). Twelve natural populations of P. alba in the Irtysh River basin in China were sequenced to explore the genetic diversity. Average pooled heterozygosity value of P. alba populations was 0.170±0.014, which was lower than that in Italy (0.271±0.051) and Hungary (0.264±0.054). Tajima’s D values showed a negative distribution, which might signify an excess of low frequency polymorphisms and a bottleneck with later expansion of P. alba populations examined.


April 21, 2020

Potential KPC-2 carbapenemase reservoir of environmental Aeromonas hydrophila and Aeromonas caviae isolates from the effluent of an urban wastewater treatment plant in Japan.

Aeromonas hydrophila and Aeromonas caviae adapt to saline water environments and are the most predominant Aeromonas species isolated from estuaries. Here, we isolated antimicrobial-resistant (AMR) Aeromonas strains (A. hydrophila GSH8-2 and A. caviae GSH8M-1) carrying the carabapenemase blaKPC-2 gene from a wastewater treatment plant (WWTP) effluent in Tokyo Bay (Japan) and determined their complete genome sequences. GSH8-2 and GSH8M-1 were classified as newly assigned sequence types ST558 and ST13, suggesting no supportive evidence of clonal dissemination. The strains appear to have acquired blaKPC-2 -positive IncP-6-relative plasmids (pGSH8-2 and pGSH8M-1-2) that share a common backbone with plasmids in Aeromonas sp. ASNIH3 isolated from hospital wastewater in the United States, A. hydrophila WCHAH045096 isolated from sewage in China, other clinical isolates (Klebsiella, Enterobacter and Escherichia coli), and wastewater isolates (Citrobacter, Pseudomonas and other Aeromonas spp.). In addition to blaKPC-2 , pGSH8M-1-2 carries an IS26-mediated composite transposon including a macrolide resistance gene, mph(A). Although Aeromonas species are opportunistic pathogens, they could serve as potential environmental reservoir bacteria for carbapenemase and AMR genes. AMR monitoring from WWTP effluents will contribute to the detection of ongoing AMR dissemination in the environment and might provide an early warning of potential dissemination in clinical settings and communities. © 2019 The Authors. Environmental Microbiology Reports published by Society for Applied Microbiology and John Wiley & Sons Ltd.


April 21, 2020

A New Species of the ?-Proteobacterium Francisella, F. adeliensis Sp. Nov., Endocytobiont in an Antarctic Marine Ciliate and Potential Evolutionary Forerunner of Pathogenic Species.

The study of the draft genome of an Antarctic marine ciliate, Euplotes petzi, revealed foreign sequences of bacterial origin belonging to the ?-proteobacterium Francisella that includes pathogenic and environmental species. TEM and FISH analyses confirmed the presence of a Francisella endocytobiont in E. petzi. This endocytobiont was isolated and found to be a new species, named F. adeliensis sp. nov.. F. adeliensis grows well at wide ranges of temperature, salinity, and carbon dioxide concentrations implying that it may colonize new organisms living in deeply diversified habitats. The F. adeliensis genome includes the igl and pdp gene sets (pdpC and pdpE excepted) of the Francisella pathogenicity island needed for intracellular growth. Consistently with an F. adeliensis ancient symbiotic lifestyle, it also contains a single insertion-sequence element. Instead, it lacks genes for the biosynthesis of essential amino acids such as cysteine, lysine, methionine, and tyrosine. In a genome-based phylogenetic tree, F. adeliensis forms a new early branching clade, basal to the evolution of pathogenic species. The correlations of this clade with the other clades raise doubts about a genuine free-living nature of the environmental Francisella species isolated from natural and man-made environments, and suggest to look at F. adeliensis as a pioneer in the Francisella colonization of eukaryotic organisms.


April 21, 2020

Genomic analysis provides insights into the transmission and pathogenicity of Talaromyces marneffei.

Talaromyces marneffei (T. marneffei) is a medically important opportunistic dimorphic fungus that infects both humans and bamboo rats. However, the mechanisms of transmission and pathogenicity of T. marneffei are poorly understood. In our study, we combined Illumina and PacBio sequencing technologies to sequence and assemble a complete genome of T. marneffei. To elucidate the transmission route and source, we sequenced three additional T. marneffei isolates using Illumina sequencing technology. Variations among isolates were used to develop a multilocus sequence typing (MLST) system comprising five housekeeping genes that can be used to discriminate between isolates derived from different sources. Our analysis revealed that human and bamboo rat share identical genotypes in these five loci. Thus, we hypothesized that T. marneffei is transmitted to humans through inhalation of spores in the surrounding environment into the lungs and that the bamboo rat can serve as an important natural reservoir for pathogens. Furthermore, we also identified temperature-dependent polyketide synthases, non-ribosomal peptide synthetases and secreted proteins as putative pathogenicity-related factors. In addition, we identified antifungal drug targets that can be investigated in future studies to elucidate the mechanisms underlying drug resistance. In summary, our study presents the basic features of the T. marneffei genome and provides insights into the transmission and pathogenicity of T. marneffei, which warrant fundamental experimental research.Copyright © 2019 Elsevier Inc. All rights reserved.


April 21, 2020

The Genome of Cucurbita argyrosperma (Silver-Seed Gourd) Reveals Faster Rates of Protein-Coding Gene and Long Noncoding RNA Turnover and Neofunctionalization within Cucurbita.

Whole-genome duplications are an important source of evolutionary novelties that change the mode and tempo at which genetic elements evolve within a genome. The Cucurbita genus experienced a whole-genome duplication around 30 million years ago, although the evolutionary dynamics of the coding and noncoding genes in this genus have not yet been scrutinized. Here, we analyzed the genomes of four Cucurbita species, including a newly assembled genome of Cucurbita argyrosperma, and compared the gene contents of these species with those of five other members of the Cucurbitaceae family to assess the evolutionary dynamics of protein-coding and long intergenic noncoding RNA (lincRNA) genes after the genome duplication. We report that Cucurbita genomes have a higher protein-coding gene birth-death rate compared with the genomes of the other members of the Cucurbitaceae family. C. argyrosperma gene families associated with pollination and transmembrane transport had significantly faster evolutionary rates. lincRNA families showed high levels of gene turnover throughout the phylogeny, and 67.7% of the lincRNA families in Cucurbita showed evidence of birth from the neofunctionalization of previously existing protein-coding genes. Collectively, our results suggest that the whole-genome duplication in Cucurbita resulted in faster rates of gene family evolution through the neofunctionalization of duplicated genes. Copyright © 2019 The Author. Published by Elsevier Inc. All rights reserved.


April 21, 2020

Complete genome sequence of Pseudomonas frederiksbergensis ERDD5:01 revealed genetic bases for survivability at high altitude ecosystem and bioprospection potential.

Pseudomonas frederiksbergensis ERDD5:01 is a psychrotrophic bacteria isolated from the glacial stream flowing from East Rathong glacier in Sikkim Himalaya. The strain showed survivability at high altitude stress conditions like freezing, frequent freeze-thaw cycles, and UV-C radiations. The complete genome of 5,746,824?bp circular chromosome and a plasmid of 371,027?bp was sequenced to understand the genetic basis of its survival strategy. Multiple copies of cold-associated genes encoding cold active chaperons, general stress response, osmotic stress, oxidative stress, membrane/cell wall alteration, carbon storage/starvation and, DNA repair mechanisms supported its survivability at extreme cold and radiations corroborating with the bacterial physiological findings. The molecular cold adaptation analysis in comparison with the genome of 15 mesophilic Pseudomonas species revealed functional insight into the strategies of cold adaptation. The genomic data also revealed the presence of industrially important enzymes.Copyright © 2018 Elsevier Inc. All rights reserved.


April 21, 2020

Endogenous pararetrovirus sequences are widely present in Citrinae genomes.

Endogenous pararetroviruses (EPRVs) are characterized in several plant genomes and their biological effects have been reported. In this study, hundreds of EPRV segments were identified in six Citrinae genomes. A total of 1034 EPRV segments were identified in the genomes of sweet orange, 2036 in pummelo, 598 in clementine mandarin, 752 in Ichang papeda, 2060 in citron and 245 in atalantia. Genomic analysis indicated that EPRV segments tend to cluster as hot spots in the genomes, particularly on chromosome 2 and 5. Large numbers of simple repeats and transposable elements were identified in the 2-kb flanking regions of the EPRV segments. Comparative genomic analysis and PCR experiments showed that there are highly conserved EPRV segments and species-specific EPRV segments between the Citrinae genomes. Phylogenetic analysis suggested that the integration events of EPRVs could initiate in a common progenitor of Citrinae species and repeatedly occur during the Citrinae divergence.Copyright © 2018 Elsevier B.V. All rights reserved.


April 21, 2020

Phylogenetic barriers to horizontal transfer of antimicrobial peptide resistance genes in the human gut microbiota.

The human gut microbiota has adapted to the presence of antimicrobial peptides (AMPs), which are ancient components of immune defence. Despite its medical importance, it has remained unclear whether AMP resistance genes in the gut microbiome are available for genetic exchange between bacterial species. Here, we show that AMP resistance and antibiotic resistance genes differ in their mobilization patterns and functional compatibilities with new bacterial hosts. First, whereas AMP resistance genes are widespread in the gut microbiome, their rate of horizontal transfer is lower than that of antibiotic resistance genes. Second, gut microbiota culturing and functional metagenomics have revealed that AMP resistance genes originating from phylogenetically distant bacteria have only a limited potential to confer resistance in Escherichia coli, an intrinsically susceptible species. Taken together, functional compatibility with the new bacterial host emerges as a key factor limiting the genetic exchange of AMP resistance genes. Finally, our results suggest that AMPs induce highly specific changes in the composition of the human microbiota, with implications for disease risks.


April 21, 2020

Diversity of phytobeneficial traits revealed by whole-genome analysis of worldwide-isolated phenazine-producing Pseudomonas spp.

Plant-beneficial Pseudomonas spp. competitively colonize the rhizosphere and display plant-growth promotion and/or disease-suppression activities. Some strains within the P. fluorescens species complex produce phenazine derivatives, such as phenazine-1-carboxylic acid. These antimicrobial compounds are broadly inhibitory to numerous soil-dwelling plant pathogens and play a role in the ecological competence of phenazine-producing Pseudomonas spp. We assembled a collection encompassing 63 strains representative of the worldwide diversity of plant-beneficial phenazine-producing Pseudomonas spp. In this study, we report the sequencing of 58 complete genomes using PacBio RS II sequencing technology. Distributed among four subgroups within the P. fluorescens species complex, the diversity of our collection is reflected by the large pangenome which accounts for 25 413 protein-coding genes. We identified genes and clusters encoding for numerous phytobeneficial traits, including antibiotics, siderophores and cyclic lipopeptides biosynthesis, some of which were previously unknown in these microorganisms. Finally, we gained insight into the evolutionary history of the phenazine biosynthetic operon. Given its diverse genomic context, it is likely that this operon was relocated several times during Pseudomonas evolution. Our findings acknowledge the tremendous diversity of plant-beneficial phenazine-producing Pseudomonas spp., paving the way for comparative analyses to identify new genetic determinants involved in biocontrol, plant-growth promotion and rhizosphere competence. © 2018 Society for Applied Microbiology and John Wiley & Sons Ltd.


April 21, 2020

Genetic map-guided genome assembly reveals a virulence-governing minichromosome in the lentil anthracnose pathogen Colletotrichum lentis.

Colletotrichum lentis causes anthracnose, which is a serious disease on lentil and can account for up to 70% crop loss. Two pathogenic races, 0 and 1, have been described in the C. lentis population from lentil. To unravel the genetic control of virulence, an isolate of the virulent race 0 was sequenced at 1481-fold genomic coverage. The 56.10-Mb genome assembly consists of 50 scaffolds with N50 scaffold length of 4.89 Mb. A total of 11 436 protein-coding gene models was predicted in the genome with 237 coding candidate effectors, 43 secondary metabolite biosynthetic enzymes and 229 carbohydrate-active enzymes (CAZymes), suggesting a contraction of the virulence gene repertoire in C. lentis. Scaffolds were assigned to 10 core and two minichromosomes using a population (race 0 × race 1, n = 94 progeny isolates) sequencing-based, high-density (14 312 single nucleotide polymorphisms) genetic map. Composite interval mapping revealed a single quantitative trait locus (QTL), qClVIR-11, located on minichromosome 11, explaining 85% of the variability in virulence of the C. lentis population. The QTL covers a physical distance of 0.84 Mb with 98 genes, including seven candidate effector and two secondary metabolite genes. Taken together, the study provides genetic and physical evidence for the existence of a minichromosome controlling the C. lentis virulence on lentil. © 2018 The Authors. New Phytologist © 2018 New Phytologist Trust.


April 21, 2020

Bioinformatic analysis of the complete genome sequence of Pectobacterium carotovorum subsp. brasiliense BZA12 and candidate effector screening

AbstractPectobacterium carotovorum subsp. brasiliense (Pcb) is a gram-negative, plant pathogenic bacterium of the soft rot Enterobacteriaceae (SRE) family. We present the complete genome sequence of Pcb strain BZA12, which reveals that Pcb strain BZA12 carries a single 4,924,809 bp chromosome with 51.97% GC content and comprises 4508 predicted protein-coding genes.Geneannotationofthese genes utilizedGO, KEGG,and COG databases.Incomparison withthree closely related soft-rot pathogens, strain BZA12 has 3797 gene families, among which 3107 gene families are identified as orthologous with those of both P. carotovorum subsp. carotovorum PCC21 and P. carotovorum subsp. odoriferum BCS7, as well as 36 putative Unique Gene Families. We selected five putative effectors from the BZA12 genome and transiently expressed them in Nicotiana benthamiana. Candidate effector A12GL002483 was localized in the cell nucleus and induced cell death. This study provides a foundation for a better understanding of the genomic structure and function of Pcb, particularly in the discovery of potential pathogenic factors and for the development of more effective strategies against this pathogen.


April 21, 2020

A 12-kb structural variation in progressive myoclonic epilepsy was newly identified by long-read whole-genome sequencing.

We report a family with progressive myoclonic epilepsy who underwent whole-exome sequencing but was negative for pathogenic variants. Similar clinical courses of a devastating neurodegenerative phenotype of two affected siblings were highly suggestive of a genetic etiology, which indicates that the survey of genetic variation by whole-exome sequencing was not comprehensive. To investigate the presence of a variant that remained unrecognized by standard genetic testing, PacBio long-read sequencing was performed. Structural variant (SV) detection using low-coverage (6×) whole-genome sequencing called 17,165 SVs (7,216 deletions and 9,949 insertions). Our SV selection narrowed down potential candidates to only five SVs (two deletions and three insertions) on the genes tagged with autosomal recessive phenotypes. Among them, a 12.4-kb deletion involving the CLN6 gene was the top candidate because its homozygous abnormalities cause neuronal ceroid lipofuscinosis. This deletion included the initiation codon and was found in a GC-rich region containing multiple repetitive elements. These results indicate the presence of a causal variant in a difficult-to-sequence region and suggest that such variants that remain enigmatic after the application of current whole-exome sequencing technology could be uncovered by unbiased application of long-read whole-genome sequencing.


April 21, 2020

Characterization of the genome of a Nocardia strain isolated from soils in the Qinghai-Tibetan Plateau that specifically degrades crude oil and of this biodegradation.

A strain of Nocardia isolated from crude oil-contaminated soils in the Qinghai-Tibetan Plateau degrades nearly all components of crude oil. This strain was identified as Nocardia soli Y48, and its growth conditions were determined. Complete genome sequencing showed that N. soli Y48 has a 7.3?Mb genome and many genes responsible for hydrocarbon degradation, biosurfactant synthesis, emulsification and other hydrocarbon degradation-related metabolisms. Analysis of the clusters of orthologous groups (COGs) and genomic islands (GIs) revealed that Y48 has undergone significant gene transfer events to adapt to changing environmental conditions (crude oil contamination). The structural features of the genome might provide a competitive edge for the survival of N. soli Y48 in oil-polluted environments and reflect the adaptation of coexisting bacteria to distinct nutritional niches.Copyright © 2018. Published by Elsevier Inc.


April 21, 2020

Fast and accurate genomic analyses using genome graphs.

The human reference genome serves as the foundation for genomics by providing a scaffold for alignment of sequencing reads, but currently only reflects a single consensus haplotype, thus impairing analysis accuracy. Here we present a graph reference genome implementation that enables read alignment across 2,800 diploid genomes encompassing 12.6 million SNPs and 4.0 million insertions and deletions (indels). The pipeline processes one whole-genome sequencing sample in 6.5?h using a system with 36?CPU cores. We show that using a graph genome reference improves read mapping sensitivity and produces a 0.5% increase in variant calling recall, with unaffected specificity. Structural variations incorporated into a graph genome can be genotyped accurately under a unified framework. Finally, we show that iterative augmentation of graph genomes yields incremental gains in variant calling accuracy. Our implementation is an important advance toward fulfilling the promise of graph genomes to radically enhance the scalability and accuracy of genomic analyses.


Talk with an expert

If you have a question, need to check the status of an order, or are interested in purchasing an instrument, we're here to help.