Sugar pine (Pinus lambertiana Douglas) is within the subgenus Strobus with an estimated genome size of 31 Gbp. Transcriptomic resources are of particular interest in conifers due to the challenges presented in their megagenomes for gene identification. In this study, we present the first comprehensive survey of the P. lambertiana transcriptome through deep sequencing of a variety of tissue types to generate more than 2.5 billion short reads. Third generation, long reads generated through PacBio Iso-Seq has been included for the first time in conifers to combat the challenges associated with de novo transcriptome assembly. A technology comparison is provided…
Anaerobic bacterial biosynthesis of toluene from phenylacetate was reported more than two decades ago, but the biochemistry underlying this novel metabolism has never been elucidated. Here we report results of in vitro characterization studies of a novel phenylacetate decarboxylase from an anaerobic, sewage-derived enrichment culture that quantitatively produces toluene from phenylacetate; complementary metagenomic and metaproteomic analyses are also presented. Among the noteworthy findings is that this enzyme is not the well-characterized clostridial p-hydroxyphenylacetate decarboxylase (CsdBC). However, the toluene synthase under study appears to be able to catalyze both phenylacetate and p-hydroxyphenylacetate decarboxylation. Observations suggesting that phenylacetate and p-hydroxyphenylacetate decarboxylation in…
Single-molecule, real-time sequencing developed by Pacific BioSciences offers longer read lengths than the second-generation sequencing (SGS) technologies, making it well-suited for unsolved problems in genome, transcriptome, and epigenetics research. The highly-contiguous de novo assemblies using PacBio sequencing can close gaps in current reference assemblies and characterize structural variation (SV) in personal genomes. With longer reads, we can sequence through extended repetitive regions and detect mutations, many of which are associated with diseases. Moreover, PacBio transcriptome sequencing is advantageous for the identification of gene isoforms and facilitates reliable discoveries of novel genes and novel isoforms of annotated genes, due to its…
Microbial toluene biosynthesis was reported in anoxic lake sediments more than three decades ago, but the enzyme catalyzing this biochemically challenging reaction has never been identified. Here we report the toluene-producing enzyme PhdB, a glycyl radical enzyme of bacterial origin that catalyzes phenylacetate decarboxylation, and its cognate activating enzyme PhdA, a radical S-adenosylmethionine enzyme, discovered in two distinct anoxic microbial communities that produce toluene. The unconventional process of enzyme discovery from a complex microbial community (>300,000 genes), rather than from a microbial isolate, involved metagenomics- and metaproteomics-enabled biochemistry, as well as in vitro confirmation of activity with recombinant enzymes. This…
Strains of Helicobacter pylori that cause ulcer or gastric cancer typically express a type IV secretion system (T4SS) encoded by the cag pathogenicity island (cagPAI). CagY is an ortholog of VirB10 that, unlike other VirB10 orthologs, has a large middle repeat region (MRR) with extensive repetitive sequence motifs, which undergo CD4+ T cell-dependent recombination during infection of mice. Recombination in the CagY MRR reduces T4SS function, diminishes the host inflammatory response, and enables the bacteria to colonize at a higher density. Since CagY is known to bind human a5ß1 integrin, we tested the hypothesis that recombination in the CagY MRR…