Annotation Archives - Page 18 of 34

April 21, 2020

De novo genome assembly of the stress tolerant forest species Casuarina equisetifolia provides insight into secondary growth.

Casuarina equisetifolia (C. equisetifolia), a conifer-like angiosperm with resistance to typhoon and stress tolerance, is mainly cultivated in the coastal areas of Australasia. C. equisetifolia, making it a valuable model to study secondary growth associated genes and stress-tolerance traits. However, the genome sequence is unavailable and therefore wood-associated growth rate and stress resistance at the molecular level is largely unexplored. We therefore constructed a high-quality draft genome sequence of C. equisetifolia by a combination of Illumina second-generation sequencing reads and Pacific Biosciences single-molecule real-time (SMRT) long reads to advance the investigation of this species. Here, we report the genome assembly, which contains approximately 300 megabases (Mb) and scaffold size of N50 is 1.06 Mb. Additionally, gene annotation, assisted by a combination of prediction and RNA-seq data, generated 29 827 annotated protein-coding genes and 1983 non-coding genes, respectively. Furthermore, we found that the total number of repetitive sequences account for one-third of the genome assembly. Here we also construct the genome-wide map of DNA modification, such as two novel forms N6 -adenine (6mA) and N4-methylcytosine (4mC) at the level of single-nucleotide resolution using single-molecule real-time (SMRT) sequencing. Interestingly, we found that 17% of 6mA modification genes and 15% of 4mC modification genes also included alternative splicing events. Finally, we investigated cellulose, hemicellulose, and lignin-related genes, which were associated with secondary growth and contained different DNA modifications. The high-quality genome sequence and annotation of C. equisetifolia in this study provide a valuable resource to strengthen our understanding of the diverse traits of trees. © 2018 The Authors The Plant Journal © 2018 John Wiley & Sons Ltd.

April 21, 2020

Conventional culture methods with commercially available media unveil the presence of novel culturable bacteria.

Recent metagenomic analysis has revealed that our gut microbiota plays an important role in not only the maintenance of our health but also various diseases such as obesity, diabetes, inflammatory bowel disease, and allergy. However, most intestinal bacteria are considered ‘unculturable’ bacteria, and their functions remain unknown. Although culture-independent genomic approaches have enabled us to gain insight into their potential roles, culture-based approaches are still required to understand their characteristic features and phenotypes. To date, various culturing methods have been attempted to obtain these ‘unculturable’ bacteria, but most such methods require advanced techniques. Here, we have tried to isolate possible unculturable bacteria from a healthy Japanese individual by using commercially available media. A 16S rRNA (ribosomal RNA) gene metagenomic analysis revealed that each culture medium showed bacterial growth depending on its selective features and a possibility of the presence of novel bacterial species. Whole genome sequencing of these candidate strains suggested the isolation of 8 novel bacterial species classified in the Actinobacteria and Firmicutes phyla. Our approach indicates that a number of intestinal bacteria hitherto considered unculturable are potentially culturable and can be cultured on commercially available media. We have obtained novel gut bacteria from a healthy Japanese individual using a combination of comprehensive genomics and conventional culturing methods. We would expect that the discovery of such novel bacteria could illuminate pivotal roles for the gut microbiota in association with human health.

April 21, 2020

Polysaccharide utilization loci of North Sea Flavobacteriia as basis for using SusC/D-protein expression for predicting major phytoplankton glycans.

Marine algae convert a substantial fraction of fixed carbon dioxide into various polysaccharides. Flavobacteriia that are specialized on algal polysaccharide degradation feature genomic clusters termed polysaccharide utilization loci (PULs). As knowledge on extant PUL diversity is sparse, we sequenced the genomes of 53 North Sea Flavobacteriia and obtained 400 PULs. Bioinformatic PUL annotations suggest usage of a large array of polysaccharides, including laminarin, a-glucans, and alginate as well as mannose-, fucose-, and xylose-rich substrates. Many of the PULs exhibit new genetic architectures and suggest substrates rarely described for marine environments. The isolates’ PUL repertoires often differed considerably within genera, corroborating ecological niche-associated glycan partitioning. Polysaccharide uptake in Flavobacteriia is mediated by SusCD-like transporter complexes. Respective protein trees revealed clustering according to polysaccharide specificities predicted by PUL annotations. Using the trees, we analyzed expression of SusC/D homologs in multiyear phytoplankton bloom-associated metaproteomes and found indications for profound changes in microbial utilization of laminarin, a-glucans, ß-mannan, and sulfated xylan. We hence suggest the suitability of SusC/D-like transporter protein expression within heterotrophic bacteria as a proxy for the temporal utilization of discrete polysaccharides.

April 21, 2020

Iso-Seq Allows Genome-Independent Transcriptome Profiling of Grape Berry Development.

Transcriptomics has been widely applied to study grape berry development. With few exceptions, transcriptomic studies in grape are performed using the available genome sequence, PN40024, as reference. However, differences in gene content among grape accessions, which contribute to phenotypic differences among cultivars, suggest that a single reference genome does not represent the species’ entire gene space. Though whole genome assembly and annotation can reveal the relatively unique or “private” gene space of any particular cultivar, transcriptome reconstruction is a more rapid, less costly, and less computationally intensive strategy to accomplish the same goal. In this study, we used single molecule-real time sequencing (SMRT) to sequence full-length cDNA (Iso-Seq) and reconstruct the transcriptome of Cabernet Sauvignon berries during berry ripening. In addition, short reads from ripening berries were used to error-correct low-expression isoforms and to profile isoform expression. By comparing the annotated gene space of Cabernet Sauvignon to other grape cultivars, we demonstrate that the transcriptome reference built with Iso-Seq data represents most of the expressed genes in the grape berries and includes 1,501 cultivar-specific genes. Iso-Seq produced transcriptome profiles similar to those obtained after mapping on a complete genome reference. Together, these results justify the application of Iso-Seq to identify cultivar-specific genes and build a comprehensive reference for transcriptional profiling that circumvents the necessity of a genome reference with its associated costs and computational weight.Copyright © 2019 Minio et al.

April 21, 2020

Genetic variation in the conjugative plasmidome of a hospital effluent multidrug resistant Escherichia coli strain.

Bacteria harboring conjugative plasmids have the potential for spreading antibiotic resistance through horizontal gene transfer. It is described that the selection and dissemination of antibiotic resistance is enhanced by stressors, like metals or antibiotics, which can occur as environmental contaminants. This study aimed at unveiling the composition of the conjugative plasmidome of a hospital effluent multidrug resistant Escherichia coli strain (H1FC54) under different mating conditions. To meet this objective, plasmid pulsed field gel electrophoresis, optical mapping analyses and DNA sequencing were used in combination with phenotype analysis. Strain H1FC54 was observed to harbor five plasmids, three of which were conjugative and two of these, pH1FC54_330 and pH1FC54_140, contained metal and antibiotic resistance genes. Transconjugants obtained in the absence or presence of tellurite (0.5?µM or 5?µM), arsenite (0.5?µM, 5?µM or 15?µM) or ceftazidime (10?mg/L) and selected in the presence of sodium azide (100?mg/L) and tetracycline (16?mg/L) presented distinct phenotypes, associated with the acquisition of different plasmid combinations, including two co-integrate plasmids, of 310 kbp and 517 kbp. The variable composition of the conjugative plasmidome, the formation of co-integrates during conjugation, as well as the transfer of non-transferable plasmids via co-integration, and the possible association between antibiotic, arsenite and tellurite tolerance was demonstrated. These evidences bring interesting insights into the comprehension of the molecular and physiological mechanisms that underlie antibiotic resistance propagation in the environment. Copyright © 2019 Elsevier Ltd. All rights reserved.

April 21, 2020

The Genome of Armadillidium vulgare (Crustacea, Isopoda) Provides Insights into Sex Chromosome Evolution in the Context of Cytoplasmic Sex Determination.

The terrestrial isopod Armadillidium vulgare is an original model to study the evolution of sex determination and symbiosis in animals. Its sex can be determined by ZW sex chromosomes, or by feminizing Wolbachia bacterial endosymbionts. Here, we report the sequence and analysis of the ZW female genome of A. vulgare. A distinguishing feature of the 1.72 gigabase assembly is the abundance of repeats (68% of the genome). We show that the Z and W sex chromosomes are essentially undifferentiated at the molecular level and the W-specific region is extremely small (at most several hundreds of kilobases). Our results suggest that recombination suppression has not spread very far from the sex-determining locus, if at all. This is consistent with A. vulgare possessing evolutionarily young sex chromosomes. We characterized multiple Wolbachia nuclear inserts in the A. vulgare genome, none of which is associated with the W-specific region. We also identified several candidate genes that may be involved in the sex determination or sexual differentiation pathways. The A. vulgare genome serves as a resource for studying the biology and evolution of crustaceans, one of the most speciose and emblematic metazoan groups. © The Author(s) 2019. Published by Oxford University Press on behalf of the Society for Molecular Biology and Evolution.

April 21, 2020

Complete genome sequence of marine Bacillus sp. Y-01, isolated from the plastics contamination in the Yellow Sea

Plastics contamination in the environment has been an increasing ecological problem. Here we present the complete genome sequence of Bacillus sp. Y-01, isolated from plastic contamination samples in the Yellow Sea, which can utilize the polypropylene as the sole carbon and energy source. The strain has one circular chromosome of 5,130,901?bp in 8 contigs with a 38.24% GC content, consisting of 4996 protein-coding genes, 118 tRNA genes, as well as 40 rRNA operons as 5S-16S-23S rRNA. The complete genome sequence of Bacillus sp. Y-01 will provide useful genetic information to further detect the molecular mechanisms behind marine microplastics degradation.

April 21, 2020

Genome of Crucihimalaya himalaica, a close relative of Arabidopsis, shows ecological adaptation to high altitude.

Crucihimalaya himalaica, a close relative of Arabidopsis and Capsella, grows on the Qinghai-Tibet Plateau (QTP) about 4,000 m above sea level and represents an attractive model system for studying speciation and ecological adaptation in extreme environments. We assembled a draft genome sequence of 234.72 Mb encoding 27,019 genes and investigated its origin and adaptive evolutionary mechanisms. Phylogenomic analyses based on 4,586 single-copy genes revealed that C. himalaica is most closely related to Capsella (estimated divergence 8.8 to 12.2 Mya), whereas both species form a sister clade to Arabidopsis thaliana and Arabidopsis lyrata, from which they diverged between 12.7 and 17.2 Mya. LTR retrotransposons in C. himalaica proliferated shortly after the dramatic uplift and climatic change of the Himalayas from the Late Pliocene to Pleistocene. Compared with closely related species, C. himalaica showed significant contraction and pseudogenization in gene families associated with disease resistance and also significant expansion in gene families associated with ubiquitin-mediated proteolysis and DNA repair. We identified hundreds of genes involved in DNA repair, ubiquitin-mediated proteolysis, and reproductive processes with signs of positive selection. Gene families showing dramatic changes in size and genes showing signs of positive selection are likely candidates for C. himalaica’s adaptation to intense radiation, low temperature, and pathogen-depauperate environments in the QTP. Loss of function at the S-locus, the reason for the transition to self-fertilization of C. himalaica, might have enabled its QTP occupation. Overall, the genome sequence of C. himalaica provides insights into the mechanisms of plant adaptation to extreme environments.Copyright © 2019 the Author(s). Published by PNAS.

April 21, 2020

Complete Genome Sequence of the Wolbachia wAlbB Endosymbiont of Aedes albopictus.

Wolbachia, an alpha-proteobacterium closely related to Rickettsia, is a maternally transmitted, intracellular symbiont of arthropods and nematodes. Aedes albopictus mosquitoes are naturally infected with Wolbachia strains wAlbA and wAlbB. Cell line Aa23 established from Ae. albopictus embryos retains only wAlbB and is a key model to study host-endosymbiont interactions. We have assembled the complete circular genome of wAlbB from the Aa23 cell line using long-read PacBio sequencing at 500× median coverage. The assembled circular chromosome is 1.48 megabases in size, an increase of more than 300 kb over the published draft wAlbB genome. The annotation of the genome identified 1,205 protein coding genes, 34 tRNA, 3 rRNA, 1 tmRNA, and 3 other ncRNA loci. The long reads enabled sequencing over complex repeat regions which are difficult to resolve with short-read sequencing. Thirteen percent of the genome comprised insertion sequence elements distributed throughout the genome, some of which cause pseudogenization. Prophage WO genes encoding some essential components of phage particle assembly are missing, while the remainder are found in five prophage regions/WO-like islands or scattered around the genome. Orthology analysis identified a core proteome of 535 orthogroups across all completed Wolbachia genomes. The majority of proteins could be annotated using Pfam and eggNOG analyses, including ankyrins and components of the Type IV secretion system. KEGG analysis revealed the absence of five genes in wAlbB which are present in other Wolbachia. The availability of a complete circular chromosome from wAlbB will enable further biochemical, molecular, and genetic analyses on this strain and related Wolbachia. © The Author(s) 2019. Published by Oxford University Press on behalf of the Society for Molecular Biology and Evolution.

April 21, 2020

Tools and Strategies for Long-Read Sequencing and De Novo Assembly of Plant Genomes.

The commercial release of third-generation sequencing technologies (TGSTs), giving long and ultra-long sequencing reads, has stimulated the development of new tools for assembling highly contiguous genome sequences with unprecedented accuracy across complex repeat regions. We survey here a wide range of emerging sequencing platforms and analytical tools for de novo assembly, provide background information for each of their steps, and discuss the spectrum of available options. Our decision tree recommends workflows for the generation of a high-quality genome assembly when used in combination with the specific needs and resources of a project.Copyright © 2019 Elsevier Ltd. All rights reserved.

April 21, 2020

Analysis of differential gene expression and alternative splicing is significantly influenced by choice of reference genome.

RNA-seq analysis has enabled the evaluation of transcriptional changes in many species including nonmodel organisms. However, in most species only a single reference genome is available and RNA-seq reads from highly divergent varieties are typically aligned to this reference. Here, we quantify the impacts of the choice of mapping genome in rice where three high-quality reference genomes are available. We aligned RNA-seq data from a popular productive rice variety to three different reference genomes and found that the identification of differentially expressed genes differed depending on which reference genome was used for mapping. Furthermore, the ability to detect differentially used transcript isoforms was profoundly affected by the choice of reference genome: Only 30% of the differentially used splicing features were detected when reads were mapped to the more commonly used, but more distantly related reference genome. This demonstrated that gene expression and splicing analysis varies considerably depending on the mapping reference genome, and that analysis of individuals that are distantly related to an available reference genome may be improved by acquisition of new genomic reference material. We observed that these differences in transcriptome analysis are, in part, due to the presence of single nucleotide polymorphisms between the sequenced individual and each respective reference genome, as well as annotation differences between the reference genomes that exist even between syntenic orthologs. We conclude that even between two closely related genomes of similar quality, using the reference genome that is most closely related to the species being sampled significantly improves transcriptome analysis. © 2019 Slabaugh et al.; Published by Cold Spring Harbor Laboratory Press for the RNA Society.

April 21, 2020

Characterizing the major structural variant alleles of the human genome.

In order to provide a comprehensive resource for human structural variants (SVs), we generated long-read sequence data and analyzed SVs for fifteen human genomes. We sequence resolved 99,604 insertions, deletions, and inversions including 2,238 (1.6 Mbp) that are shared among all discovery genomes with an additional 13,053 (6.9 Mbp) present in the majority, indicating minor alleles or errors in the reference. Genotyping in 440 additional genomes confirms the most common SVs in unique euchromatin are now sequence resolved. We report a ninefold SV bias toward the last 5 Mbp of human chromosomes with nearly 55% of all VNTRs (variable number of tandem repeats) mapping to this portion of the genome. We identify SVs affecting coding and noncoding regulatory loci improving annotation and interpretation of functional variation. These data provide the framework to construct a canonical human reference and a resource for developing advanced representations capable of capturing allelic diversity. Copyright © 2018 Elsevier Inc. All rights reserved.

April 21, 2020

Inter-chromosomal coupling between vision and pigmentation genes during genomic divergence.

Recombination between loci underlying mate choice and ecological traits is a major evolutionary force acting against speciation with gene flow. The evolution of linkage disequilibrium between such loci is therefore a fundamental step in the origin of species. Here, we show that this process can take place in the absence of physical linkage in hamlets-a group of closely related reef fishes from the wider Caribbean that differ essentially in colour pattern and are reproductively isolated through strong visually-based assortative mating. Using full-genome analysis, we identify four narrow genomic intervals that are consistently differentiated among sympatric species in a backdrop of extremely low genomic divergence. These four intervals include genes involved in pigmentation (sox10), axial patterning (hoxc13a), photoreceptor development (casz1) and visual sensitivity (SWS and LWS opsins) that develop islands of long-distance and inter-chromosomal linkage disequilibrium as species diverge. The relatively simple genomic architecture of species differences facilitates the evolution of linkage disequilibrium in the presence of gene flow.

April 21, 2020

Complete genome sequences of a H2O2-resistant psychrophilic bacterium Colwellia sp. Arc7-D isolated from Arctic Ocean sediment

Colwellia sp. Arc7-D, a psychrophilic H2O2-resisitant bacterium, was isolated from Arctic Ocean sediment. Here we describe the complete genome of Colwellia sp. Arc7-D. The genome has one circular chromosome of 4,305,442?bp (37.67?mol%?G?+?C content), consisting of 3526 coding genes, 77 tRNA genes, as well as five rRNA operons as 16S–23S-5S rRNA and one rRNA operon as 16S-23S-5S-5S. According to KEGG analysis, strain Arc7-D encodes 23 genes related with antioxidant activity including superoxide dismutase, glutathione peroxidase, glutathione reductase and catalase. However, many additional genes affiliated with anti-oxidative stress were also identified, such as aconitase, thioredoxin and ascorbic acid.

April 21, 2020

Genomic and Functional Characterization of the Endophytic Bacillus subtilis 7PJ-16 Strain, a Potential Biocontrol Agent of Mulberry Fruit Sclerotiniose.

Bacillus sp. 7PJ-16, an endophytic bacterium isolated from a healthy mulberry stem and previously identified as Bacillus tequilensis 7PJ-16, exhibits strong antifungal activity and has the capacity to promote plant growth. This strain was studied for its effectiveness as a biocontrol agent to reduce mulberry fruit sclerotiniose in the field and as a growth-promoting agent for mulberry in the greenhouse. In field studies, the cell suspension and supernatant of strain 7PJ-16 exhibited biocontrol efficacy and the lowest disease incidence was reduced down to only 0.80%. In greenhouse experiments, the cell suspension (1.0?×?106 and 1.0?×?105 CFU/mL) and the cell-free supernatant (100-fold and 1000-fold dilution) stimulated mulberry seed germination and promoted mulberry seedling growth. In addition, to accurately identify the 7PJ-16 strain and further explore the mechanisms of its antifungal and growth-promoting properties, the complete genome of this strain was sequenced and annotated. The 7PJ-16 genome is comprised of two circular plasmids and a 4,209,045-bp circular chromosome, containing 4492 protein-coding genes and 116 RNA genes. This strain was ultimately designed as Bacillus subtilis based on core genome sequence analyses using a phylogenomic approach. In this genome, we identified a series of gene clusters that function in the synthesis of non-ribosomal peptides (surfactin, fengycin, bacillibactin, and bacilysin) as well as the ribosome-dependent synthesis of tasA and bacteriocins (subtilin, subtilosin A), which are responsible for the biosynthesis of numerous antimicrobial metabolites. Additionally, several genes with function that promote plant growth, such as indole-3-acetic acid biosynthesis, the production of volatile substances, and siderophores synthesis, were also identified. The information described in this study has established a good foundation for understanding the beneficial interactions between endophytes and host plants, and facilitates the further application of B. subtilis 7PJ-16 as an agricultural biofertilizer and biocontrol agent.

Auto Tag: Annotation

De novo genome assembly of the stress tolerant forest species Casuarina equisetifolia provides insight into secondary growth.

Conventional culture methods with commercially available media unveil the presence of novel culturable bacteria.

Polysaccharide utilization loci of North Sea Flavobacteriia as basis for using SusC/D-protein expression for predicting major phytoplankton glycans.

Iso-Seq Allows Genome-Independent Transcriptome Profiling of Grape Berry Development.

Genetic variation in the conjugative plasmidome of a hospital effluent multidrug resistant Escherichia coli strain.

The Genome of Armadillidium vulgare (Crustacea, Isopoda) Provides Insights into Sex Chromosome Evolution in the Context of Cytoplasmic Sex Determination.

Complete genome sequence of marine Bacillus sp. Y-01, isolated from the plastics contamination in the Yellow Sea

Genome of Crucihimalaya himalaica, a close relative of Arabidopsis, shows ecological adaptation to high altitude.

Complete Genome Sequence of the Wolbachia wAlbB Endosymbiont of Aedes albopictus.

Tools and Strategies for Long-Read Sequencing and De Novo Assembly of Plant Genomes.

Analysis of differential gene expression and alternative splicing is significantly influenced by choice of reference genome.

Characterizing the major structural variant alleles of the human genome.

Inter-chromosomal coupling between vision and pigmentation genes during genomic divergence.

Complete genome sequences of a H2O2-resistant psychrophilic bacterium Colwellia sp. Arc7-D isolated from Arctic Ocean sediment

Genomic and Functional Characterization of the Endophytic Bacillus subtilis 7PJ-16 Strain, a Potential Biocontrol Agent of Mulberry Fruit Sclerotiniose.

Subscribe for blog updates:

Filter by topic

Talk with an expert

Antimicrobial resistance research

Subscribe for blog updates:

Filter by topic

Talk with an expert