Menu
April 21, 2020

Comparative Genomics of Marine Sponge-Derived Streptomyces spp. Isolates SM17 and SM18 With Their Closest Terrestrial Relatives Provides Novel Insights Into Environmental Niche Adaptations and Secondary Metabolite Biosynthesis Potential.

The emergence of antibiotic resistant microorganisms has led to an increased need for the discovery and development of novel antimicrobial compounds. Frequent rediscovery of the same natural products (NPs) continues to decrease the likelihood of the discovery of new compounds from soil bacteria. Thus, efforts have shifted toward investigating microorganisms and their secondary metabolite biosynthesis potential, from diverse niche environments, such as those isolated from marine sponges. Here we investigated at the genomic level two Streptomyces spp. strains, namely SM17 and SM18, isolated from the marine sponge Haliclona simulans, with previously reported antimicrobial activity against clinically relevant pathogens; using single molecule real-time (SMRT) sequencing. We performed a series of comparative genomic analyses on SM17 and SM18 with their closest terrestrial relatives, namely S. albus J1074 and S. pratensis ATCC 33331 respectively; in an effort to provide further insights into potential environmental niche adaptations (ENAs) of marine sponge-associated Streptomyces, and on how these adaptations might be linked to their secondary metabolite biosynthesis potential. Prediction of secondary metabolite biosynthetic gene clusters (smBGCs) indicated that, even though the marine isolates are closely related to their terrestrial counterparts at a genomic level; they potentially produce different compounds. SM17 and SM18 displayed a better ability to grow in high salinity medium when compared to their terrestrial counterparts, and further analysis of their genomes indicated that they possess a pool of 29 potential ENA genes that are absent in S. albus J1074 and S. pratensis ATCC 33331. This ENA gene pool included functional categories of genes that are likely to be related to niche adaptations and which could be grouped based on potential biological functions such as osmotic stress, defense; transcriptional regulation; symbiotic interactions; antimicrobial compound production and resistance; ABC transporters; together with horizontal gene transfer and defense-related features.


April 21, 2020

Bioinformatic discovery of a toxin family in Chryseobacterium piperi with sequence similarity to botulinum neurotoxins.

Clostridial neurotoxins (CNTs), which include botulinum neurotoxins (BoNTs) and tetanus neurotoxin (TeNT), are the most potent toxins known to science and are the causative agents of botulism and tetanus, respectively. The evolutionary origins of CNTs and their relationships to other proteins remains an intriguing question. Here we present a large-scale bioinformatic screen for putative toxin genes in all currently available genomes. We detect a total of 311 protein sequences displaying at least partial homology to BoNTs, including 161 predicted toxin sequences that have never been characterized. We focus on a novel toxin family from Chryseobacterium piperi with homology to BoNTs. We resequenced the genome of C. piperi to confirm and further analyze the genomic context of these toxins, and also examined their potential toxicity by expression of the protease domain of one C. piperi toxin in human cells. Our analysis suggests that these C. piperi sequences encode a novel family of metalloprotease toxins that are distantly related to BoNTs with similar domain architecture. These toxins target a yet unknown class of substrates, potentially reflecting divergence in substrate specificity between the metalloprotease domains of these toxins and the related metalloprotease domain of clostridial neurotoxins.


April 21, 2020

A high-quality apple genome assembly reveals the association of a retrotransposon and red fruit colour.

A complete and accurate genome sequence provides a fundamental tool for functional genomics and DNA-informed breeding. Here, we assemble a high-quality genome (contig N50 of 6.99?Mb) of the apple anther-derived homozygous line HFTH1, including 22 telomere sequences, using a combination of PacBio single-molecule real-time (SMRT) sequencing, chromosome conformation capture (Hi-C) sequencing, and optical mapping. In comparison to the Golden Delicious reference genome, we identify 18,047 deletions, 12,101 insertions and 14 large inversions. We reveal that these extensive genomic variations are largely attributable to activity of transposable elements. Interestingly, we find that a long terminal repeat (LTR) retrotransposon insertion upstream of MdMYB1, a core transcriptional activator of anthocyanin biosynthesis, is associated with red-skinned phenotype. This finding provides insights into the molecular mechanisms underlying red fruit coloration, and highlights the utility of this high-quality genome assembly in deciphering agriculturally important trait in apple.


April 21, 2020

A study of the extraordinarily strong and tough silk produced by bagworms.

Global ecological damage has heightened the demand for silk as ‘a structural material made from sustainable resources’. Scientists have earnestly searched for stronger and tougher silks. Bagworm silk might be a promising candidate considering its superior capacity to dangle a heavy weight, summed up by the weights of the larva and its house. However, detailed mechanical and structural studies on bagworm silks have been lacking. Herein, we show the superior potential of the silk produced by Japan’s largest bagworm, Eumeta variegata. This bagworm silk is extraordinarily strong and tough, and its tensile deformation behaviour is quite elastic. The outstanding mechanical property is the result of a highly ordered hierarchical structure, which remains unchanged until fracture. Our findings demonstrate how the hierarchical structure of silk proteins plays an important role in the mechanical property of silk fibres.


April 21, 2020

Complete Assembly of the Genome of an Acidovorax citrulli Strain Reveals a Naturally Occurring Plasmid in This Species.

Acidovorax citrulli is the causal agent of bacterial fruit blotch (BFB), a serious threat to cucurbit crop production worldwide. Based on genetic and phenotypic properties, A. citrulli strains are divided into two major groups: group I strains have been generally isolated from melon and other non-watermelon cucurbits, while group II strains are closely associated with watermelon. In a previous study, we reported the genome of the group I model strain, M6. At that time, the M6 genome was sequenced by MiSeq Illumina technology, with reads assembled into 139 contigs. Here, we report the assembly of the M6 genome following sequencing with PacBio technology. This approach not only allowed full assembly of the M6 genome, but it also revealed the occurrence of a ~53 kb plasmid. The M6 plasmid, named pACM6, was further confirmed by plasmid extraction, Southern-blot analysis of restricted fragments and obtention of M6-derivative cured strains. pACM6 occurs at low copy numbers (average of ~4.1 ± 1.3 chromosome equivalents) in A. citrulli M6 and contains 63 open reading frames (ORFs), most of which (55.6%) encoding hypothetical proteins. The plasmid contains several genes encoding type IV secretion components, and typical plasmid-borne genes involved in plasmid maintenance, replication and transfer. The plasmid also carries an operon encoding homologs of a Fic-VbhA toxin-antitoxin (TA) module. Transcriptome data from A. citrulli M6 revealed that, under the tested conditions, the genes encoding the components of this TA system are among the highest expressed genes in pACM6. Whether this TA module plays a role in pACM6 maintenance is still to be determined. Leaf infiltration and seed transmission assays revealed that, under tested conditions, the loss of pACM6 did not affect the virulence of A. citrulli M6. We also show that pACM6 or similar plasmids are present in several group I strains, but absent in all tested group II strains of A. citrulli.


April 21, 2020

Mobilome of Brevibacterium aurantiacum Sheds Light on Its Genetic Diversity and Its Adaptation to Smear-Ripened Cheeses.

Brevibacterium aurantiacum is an actinobacterium that confers key organoleptic properties to washed-rind cheeses during the ripening process. Although this industrially relevant species has been gaining an increasing attention in the past years, its genome plasticity is still understudied due to the unavailability of complete genomic sequences. To add insights on the mobilome of this group, we sequenced the complete genomes of five dairy Brevibacterium strains and one non-dairy strain using PacBio RSII. We performed phylogenetic and pan-genome analyses, including comparisons with other publicly available Brevibacterium genomic sequences. Our phylogenetic analysis revealed that these five dairy strains, previously identified as Brevibacterium linens, belong instead to the B. aurantiacum species. A high number of transposases and integrases were observed in the Brevibacterium spp. strains. In addition, we identified 14 and 12 new insertion sequences (IS) in B. aurantiacum and B. linens genomes, respectively. Several stretches of homologous DNA sequences were also found between B. aurantiacum and other cheese rind actinobacteria, suggesting horizontal gene transfer (HGT). A HGT region from an iRon Uptake/Siderophore Transport Island (RUSTI) and an iron uptake composite transposon were found in five B. aurantiacum genomes. These findings suggest that low iron availability in milk is a driving force in the adaptation of this bacterial species to this niche. Moreover, the exchange of iron uptake systems suggests cooperative evolution between cheese rind actinobacteria. We also demonstrated that the integrative and conjugative element BreLI (Brevibacterium Lanthipeptide Island) can excise from B. aurantiacum SMQ-1417 chromosome. Our comparative genomic analysis suggests that mobile genetic elements played an important role into the adaptation of B. aurantiacum to cheese ecosystems.


April 21, 2020

A reference-grade wild soybean genome.

Efficient crop improvement depends on the application of accurate genetic information contained in diverse germplasm resources. Here we report a reference-grade genome of wild soybean accession W05, with a final assembled genome size of 1013.2?Mb and a contig N50 of 3.3?Mb. The analytical power of the W05 genome is demonstrated by several examples. First, we identify an inversion at the locus determining seed coat color during domestication. Second, a translocation event between chromosomes 11 and 13 of some genotypes is shown to interfere with the assignment of QTLs. Third, we find a region containing copy number variations of the Kunitz trypsin inhibitor (KTI) genes. Such findings illustrate the power of this assembly in the analysis of large structural variations in soybean germplasm collections. The wild soybean genome assembly has wide applications in comparative genomic and evolutionary studies, as well as in crop breeding and improvement programs.


April 21, 2020

Genomic analyses of two Alteromonas stellipolaris strains reveal traits with potential biotechnological applications.

The Alteromonas stellipolaris strains PQQ-42 and PQQ-44, previously isolated from a fish hatchery, have been selected on the basis of their strong quorum quenching (QQ) activity, as well as their ability to reduce Vibrio-induced mortality on the coral Oculina patagonica. In this study, the genome sequences of both strains were determined and analyzed in order to identify the mechanism responsible for QQ activity. Both PQQ-42 and PQQ-44 were found to degrade a wide range of N-acylhomoserine lactone (AHL) QS signals, possibly due to the presence of an aac gene which encodes an AHL amidohydrolase. In addition, the different colony morphologies exhibited by the strains could be related to the differences observed in genes encoding cell wall biosynthesis and exopolysaccharide (EPS) production. The PQQ-42 strain produces more EPS (0.36?g?l-1) than the PQQ-44 strain (0.15?g?l-1), whose chemical compositions also differ. Remarkably, PQQ-44 EPS contains large amounts of fucose, a sugar used in high-value biotechnological applications. Furthermore, the genome of strain PQQ-42 contained a large non-ribosomal peptide synthase (NRPS) cluster with a previously unknown genetic structure. The synthesis of enzymes and other bioactive compounds were also identified, indicating that PQQ-42 and PQQ-44 could have biotechnological applications.


April 21, 2020

In-Depth Genomic and Phenotypic Characterization of the Antarctic Psychrotolerant Strain Pseudomonas sp. MPC6 Reveals Unique Metabolic Features, Plasticity, and Biotechnological Potential.

We obtained the complete genome sequence of the psychrotolerant extremophile Pseudomonas sp. MPC6, a natural Polyhydroxyalkanoates (PHAs) producing bacterium able to rapidly grow at low temperatures. Genomic and phenotypic analyses allowed us to situate this isolate inside the Pseudomonas fluorescens phylogroup of pseudomonads as well as to reveal its metabolic versatility and plasticity. The isolate possesses the gene machinery for metabolizing a variety of toxic aromatic compounds such as toluene, phenol, chloroaromatics, and TNT. In addition, it can use both C6- and C5-carbon sugars like xylose and arabinose as carbon substrates, an uncommon feature for bacteria of this genus. Furthermore, Pseudomonas sp. MPC6 exhibits a high-copy number of genes encoding for enzymes involved in oxidative and cold-stress response that allows it to cope with high concentrations of heavy metals (As, Cd, Cu) and low temperatures, a finding that was further validated experimentally. We then assessed the growth performance of MPC6 on glycerol using a temperature range from 0 to 45°C, the latter temperature corresponding to the limit at which this Antarctic isolate was no longer able to propagate. On the other hand, the MPC6 genome comprised considerably less virulence and drug resistance factors as compared to pathogenic Pseudomonas strains, thus supporting its safety. Unexpectedly, we found five PHA synthases within the genome of MPC6, one of which clustered separately from the other four. This PHA synthase shared only 40% sequence identity at the amino acid level against the only PHA polymerase described for Pseudomonas (63-1 strain) able to produce copolymers of short- and medium-chain length PHAs. Batch cultures for PHA synthesis in Pseudomonas sp. MPC6 using sugars, decanoate, ethylene glycol, and organic acids as carbon substrates result in biopolymers with different monomer compositions. This indicates that the PHA synthases play a critical role in defining not only the final chemical structure of the biosynthesized PHA, but also the employed biosynthetic pathways. Based on the results obtained, we conclude that Pseudomonas sp. MPC6 can be exploited as a bioremediator and biopolymer factory, as well as a model strain to unveil molecular mechanisms behind adaptation to cold and extreme environments.


April 21, 2020

A multi-task convolutional deep neural network for variant calling in single molecule sequencing.

The accurate identification of DNA sequence variants is an important, but challenging task in genomics. It is particularly difficult for single molecule sequencing, which has a per-nucleotide error rate of ~5-15%. Meeting this demand, we developed Clairvoyante, a multi-task five-layer convolutional neural network model for predicting variant type (SNP or indel), zygosity, alternative allele and indel length from aligned reads. For the well-characterized NA12878 human sample, Clairvoyante achieves 99.67, 95.78, 90.53% F1-score on 1KP common variants, and 98.65, 92.57, 87.26% F1-score for whole-genome analysis, using Illumina, PacBio, and Oxford Nanopore data, respectively. Training on a second human sample shows Clairvoyante is sample agnostic and finds variants in less than 2?h on a standard server. Furthermore, we present 3,135 variants that are missed using Illumina but supported independently by both PacBio and Oxford Nanopore reads. Clairvoyante is available open-source ( https://github.com/aquaskyline/Clairvoyante ), with modules to train, utilize and visualize the model.


April 21, 2020

Characterization of the Castanopsis carlesii Deadwood Mycobiome by Pacbio Sequencing of the Full-Length Fungal Nuclear Ribosomal Internal Transcribed Spacer (ITS)

Short-read Next Generation Sequencing (NGS) platforms can easily and quickly generate thousands to hundreds of thousands of sequences per sample. However, the limited length of these sequences can cause problems during fungal taxonomic identification. Here we validate the use of Pacbio sequencing, a long-read NGS method, for characterizing the fungal community (mycobiome) of Castanopsis carlesii deadwood. We report the successful use of Pacbio sequencing to generate long-read sequences of the full-length (500 – 780 bp) fungal ITS regions of the Castanopsis carlesii mycobiome. Our results show that the studied deadwood mycobiome is taxonomically and functionally diverse, with an average of 85 fungal OTUs representing five functional groups (animal endosymbionts, endophytes, mycoparasites, plant pathogens, and saprotrophs). Based on relative abundance data, Basidiomycota were the most frequently detected phyla (50% of total sequences), followed by unidentified phyla and Ascomycota. However, based on presence/absence data, the most OTU-rich phyla were Ascomycota (58% of total OTUs, 72 OTUs) followed by Basidiomycota and unidentified phyla. The majority of fungal OTUs were identified as saprotrophs (70% of successfully function-assigned OTUs) followed by plant pathogens. Finally, we used phylogenetic analysis based on the full-length ITS sequences to confirm the species identification of 14/36 OTUs with high bootstrap support (99 – 100%). Based on the numbers of sequence reads obtained per sample, which ranged from 3,047 to 13,463, we conclude that Pacbio sequencing can be a powerful tool for characterizing moderate- and possibly high-complexity fungal communities.


April 21, 2020

Study of the whole genome, methylome and transcriptome of Cordyceps militaris.

The complete genome of Cordyceps militaris was sequenced using single-molecule real-time (SMRT) sequencing technology at a coverage over 300×. The genome size was 32.57?Mb, and 14 contigs ranging from 0.35 to 4.58?Mb with an N50 of 2.86?Mb were assembled, including 4 contigs with telomeric sequences on both ends and an additional 8 contigs with telomeric sequences on either the 5′ or 3′ end. A methylome database of the genome was constructed using SMRT and m4C and m6A methylated nucleotides, and many unknown modification types were identified. The major m6A methylation motif is GA and GGAG, and the major m4C methylation motif is GC or CG/GC. In the C. militaris genome DNA, there were four types of methylated nucleotides that we confirmed using high-resolution LCMS-IT-TOF. Using PacBio Iso-Seq, a total of 31,133 complete cDNA sequences were obtained in the fruiting body. The conserved domains of the nontranscribed regions of the genome include TATA boxes, which are the initial regions of genome replication. There were 406 structural variants between the HN and CM01 strains, and there were 1,114 structural variants between the HN and ATCC strains.


April 21, 2020

Genome Features and Secondary Metabolites Biosynthetic Potential of the Class Ktedonobacteria.

The prevalence of antibiotic resistance and the decrease in novel antibiotic discovery in recent years necessitates the identification of potentially novel microbial resources to produce natural products. Ktedonobacteria, a class of deeply branched bacterial lineage in the ancient phylum Chloroflexi, are ubiquitous in terrestrial environments and characterized by their large genome size and complex life cycle. These characteristics indicate Ktedonobacteria as a potential active producer of bioactive compounds. In this study, we observed the existence of a putative “megaplasmid,” multiple copies of ribosomal RNA operons, and high ratio of hypothetical proteins with unknown functions in the class Ktedonobacteria. Furthermore, a total of 104 antiSMASH-predicted putative biosynthetic gene clusters (BGCs) for secondary metabolites with high novelty and diversity were identified in nine Ktedonobacteria genomes. Our investigation of domain composition and organization of the non-ribosomal peptide synthetase and polyketide synthase BGCs further supports the concept that class Ktedonobacteria may produce compounds structurally different from known natural products. Furthermore, screening of bioactive compounds from representative Ktedonobacteria strains resulted in the identification of broad antimicrobial activities against both Gram-positive and Gram-negative tested bacterial strains. Based on these findings, we propose the ancient, ubiquitous, and spore-forming Ktedonobacteria as a versatile and promising microbial resource for natural product discovery.


April 21, 2020

Biodegradation of naphthalene, BTEX, and aliphatic hydrocarbons by Paraburkholderia aromaticivorans BN5 isolated from petroleum-contaminated soil.

To isolate bacteria responsible for the biodegradation of naphthalene, BTEX (benzene, toluene, ethylbenzene, and o-, m-, and p-xylene), and aliphatic hydrocarbons in petroleum-contaminated soil, three enrichment cultures were established using soil extract as the medium supplemented with naphthalene, BTEX, or n-hexadecane. Community analyses showed that Paraburkholderia species were predominant in naphthalene and BTEX, but relatively minor in n-hexadecane. Paraburkholderia aromaticivorans BN5 was able to degrade naphthalene and all BTEX compounds, but not n-hexadecane. The genome of strain BN5 harbors genes encoding 29 monooxygenases including two alkane 1-monooxygenases and 54 dioxygenases, indicating that strain BN5 has versatile metabolic capabilities, for diverse organic compounds: the ability of strain BN5 to degrade short chain aliphatic hydrocarbons was verified experimentally. The biodegradation pathways of naphthalene and BTEX compounds were bioinformatically predicted and verified experimentally through the analysis of their metabolic intermediates. Some genomic features including the encoding of the biodegradation genes on a plasmid and the low sequence homologies of biodegradation-related genes suggest that biodegradation potentials of strain BN5 may have been acquired via horizontal gene transfers and/or gene duplication, resulting in enhanced ecological fitness by enabling strain BN5 to degrade all compounds including naphthalene, BTEX, and short aliphatic hydrocarbons in contaminated soil.


April 21, 2020

Sequence properties of certain GC rich avian genes, their origins and absence from genome assemblies: case studies.

More and more eukaryotic genomes are sequenced and assembled, most of them presented as a complete model in which missing chromosomal regions are filled by Ns and where a few chromosomes may be lacking. Avian genomes often contain sequences with high GC content, which has been hypothesized to be at the origin of many missing sequences in these genomes. We investigated features of these missing sequences to discover why some may not have been integrated into genomic libraries and/or sequenced.The sequences of five red jungle fowl cDNA models with high GC content were used as queries to search publicly available datasets of Illumina and Pacbio sequencing reads. These were used to reconstruct the leptin, TNFa, MRPL52, PCP2 and PET100 genes, all of which are absent from the red jungle fowl genome model. These gene sequences displayed elevated GC contents, had intron sizes that were sometimes larger than non-avian orthologues, and had non-coding regions that contained numerous tandem and inverted repeat sequences with motifs able to assemble into stable G-quadruplexes and intrastrand dyadic structures. Our results suggest that Illumina technology was unable to sequence the non-coding regions of these genes. On the other hand, PacBio technology was able to sequence these regions, but with dramatically lower efficiency than would typically be expected.High GC content was not the principal reason why numerous GC-rich regions of avian genomes are missing from genome assembly models. Instead, it is the presence of tandem repeats containing motifs capable of assembling into very stable secondary structures that is likely responsible.


Talk with an expert

If you have a question, need to check the status of an order, or are interested in purchasing an instrument, we're here to help.