Menu
April 21, 2020

The red bayberry genome and genetic basis of sex determination.

Morella rubra, red bayberry, is an economically important fruit tree in south China. Here, we assembled the first high-quality genome for both a female and a male individual of red bayberry. The genome size was 313-Mb, and 90% sequences were assembled into eight pseudo chromosome molecules, with 32 493 predicted genes. By whole-genome comparison between the female and male and association analysis with sequences of bulked and individual DNA samples from female and male, a 59-Kb region determining female was identified and located on distal end of pseudochromosome 8, which contains abundant transposable element and seven putative genes, four of them are related to sex floral development. This 59-Kb female-specific region was likely to be derived from duplication and rearrangement of paralogous genes and retained non-recombinant in the female-specific region. Sex-specific molecular markers developed from candidate genes co-segregated with sex in a genetically diverse female and male germplasm. We propose sex determination follow the ZW model of female heterogamety. The genome sequence of red bayberry provides a valuable resource for plant sex chromosome evolution and also provides important insights for molecular biology, genetics and modern breeding in Myricaceae family. © 2018 The Authors. Plant Biotechnology Journal published by Society for Experimental Biology and The Association of Applied Biologists and John Wiley & Sons Ltd.


April 21, 2020

Bioinformatic analysis of the complete genome sequence of Pectobacterium carotovorum subsp. brasiliense BZA12 and candidate effector screening

AbstractPectobacterium carotovorum subsp. brasiliense (Pcb) is a gram-negative, plant pathogenic bacterium of the soft rot Enterobacteriaceae (SRE) family. We present the complete genome sequence of Pcb strain BZA12, which reveals that Pcb strain BZA12 carries a single 4,924,809 bp chromosome with 51.97% GC content and comprises 4508 predicted protein-coding genes.Geneannotationofthese genes utilizedGO, KEGG,and COG databases.Incomparison withthree closely related soft-rot pathogens, strain BZA12 has 3797 gene families, among which 3107 gene families are identified as orthologous with those of both P. carotovorum subsp. carotovorum PCC21 and P. carotovorum subsp. odoriferum BCS7, as well as 36 putative Unique Gene Families. We selected five putative effectors from the BZA12 genome and transiently expressed them in Nicotiana benthamiana. Candidate effector A12GL002483 was localized in the cell nucleus and induced cell death. This study provides a foundation for a better understanding of the genomic structure and function of Pcb, particularly in the discovery of potential pathogenic factors and for the development of more effective strategies against this pathogen.


April 21, 2020

Investigating the bacterial microbiota of traditional fermented dairy products using propidium monoazide with single-molecule real-time sequencing.

Traditional fermented dairy foods have been the major components of the Mongolian diet for millennia. In this study, we used propidium monoazide (PMA; binds to DNA of nonviable cells so that only viable cells are enumerated) and single-molecule real-time sequencing (SMRT) technology to investigate the total and viable bacterial compositions of 19 traditional fermented dairy foods, including koumiss from Inner Mongolia (KIM), koumiss from Mongolia (KM), and fermented cow milk from Mongolia (CM); sample groups treated with PMA were designated PKIM, PKM, and PCM. Full-length 16S rRNA sequencing identified 195 bacterial species in 121 genera and 13 phyla in PMA-treated and untreated samples. The PMA-treated and untreated samples differed significantly in their bacterial community composition and a-diversity values. The predominant species in KM, KIM, and CM were Lactobacillus helveticus, Streptococcus parauberis, and Lactobacillus delbrueckii, whereas the predominant species in PKM, PKIM, and PCM were Enterobacter xiangfangensis, Lactobacillus helveticus, and E. xiangfangensis, respectively. Weighted and unweighted principal coordinate analyses showed a clear clustering pattern with good separation and only minor overlapping. In addition, a pure culture method was performed to obtain lactic acid bacteria resources in dairy samples according to the results of SMRT sequencing. A total of 102 LAB strains were identified and Lb. helveticus (68.63%) was the most abundant, in agreement with SMRT sequencing results. Our results revealed that the bacterial communities of traditional dairy foods are complex and vary by type of fermented dairy product. The PMA treatment induced significant changes in bacterial community structure.Copyright © 2019 American Dairy Science Association. Published by Elsevier Inc. All rights reserved.


April 21, 2020

Nodule bacteria from the cultured legume Phaseolus dumosus (belonging to the Phaseolus vulgaris cross-inoculation group) with common tropici phenotypic characteristics and symbiovar but distinctive phylogenomic position and chromid.

Phaseolus dumosus is an endemic species from mountain tops in Mexico that was found in traditional agriculture areas in Veracruz, Mexico. P. dumosus plants were identified by ITS sequences and their nodules were collected from agricultural fields or from trap plant experiments in the laboratory. Bacteria from P. dumosus nodules were identified as belonging to the phaseoli-etli-leguminosarum (PEL) or to the tropici group by 16S rRNA gene sequences. We obtained complete closed genomes from two P. dumosus isolates CCGE531 and CCGE532 that were phylogenetically placed within the tropici group but with a distinctive phylogenomic position and low average nucleotide identity (ANI). CCGE531 and CCGE532 had common phenotypic characteristics with tropici type B rhizobial symbionts. Genome synteny analysis and ANI showed that P. dumosus isolates had different chromids and our analysis suggests that chromids have independently evolved in different lineages of the Rhizobium genus. Finally, we considered that P. dumosus and Phaseolus vulgaris plants belong to the same cross-inoculation group since they have conserved symbiotic affinites for rhizobia.Copyright © 2018 Elsevier GmbH. All rights reserved.


April 21, 2020

Characterization of the genome of a Nocardia strain isolated from soils in the Qinghai-Tibetan Plateau that specifically degrades crude oil and of this biodegradation.

A strain of Nocardia isolated from crude oil-contaminated soils in the Qinghai-Tibetan Plateau degrades nearly all components of crude oil. This strain was identified as Nocardia soli Y48, and its growth conditions were determined. Complete genome sequencing showed that N. soli Y48 has a 7.3?Mb genome and many genes responsible for hydrocarbon degradation, biosurfactant synthesis, emulsification and other hydrocarbon degradation-related metabolisms. Analysis of the clusters of orthologous groups (COGs) and genomic islands (GIs) revealed that Y48 has undergone significant gene transfer events to adapt to changing environmental conditions (crude oil contamination). The structural features of the genome might provide a competitive edge for the survival of N. soli Y48 in oil-polluted environments and reflect the adaptation of coexisting bacteria to distinct nutritional niches.Copyright © 2018. Published by Elsevier Inc.


April 21, 2020

Characterization of an NDM-19-producing Klebsiella pneumoniae strain harboring 2 resistance plasmids from China.

Carbapenem-resistant Klebsiella pneumoniae (CRKP) has become a major cause of nosocomial infections and posed challenges on clinical treatments. The main objective of this study was to determinate the genetic characteristics of the NDM-19-producing CRKP strain SCM96. From 2015 to 2017, 18 CRKP strains were recovered from sputum samples of patients in respiratory medicine in 6 hospitals from 5 provinces and cities in China. Polymerase chain reaction results for carbapenem resistance genes detection showed strain SCM96 carried blaNDM-19. Three types of transconjugants harboring different plasmids were selected by conjugation experiment. The Whole Genome Sequencing (WGS) was performed using the PacBio RS platform. The genome size of SCM96 was 5,579,775?bp and composed of chromosomal DNA (5,398,745?bp) and 2 plasmids, IncFII type plasmid pSCM96-1 (134,869?bp) and IncX3 type plasmid pSCM96-2 (46,161?bp). SCM96 belonged to ST15 and K28. In addition to the 4 antibiotic resistance genes located in the chromosome, pSCM96-1 carried a complex resistance region containing 17 resistance genes and several mobile genetic elements (MGEs) like ?Tn6029, In4-like integron, and Tn3, and pSCM96-2 had only 1 blaNDM-19 gene. As far as we know, this was the first description of blaNDM-19 in K. pneumoniae. Up to 22 antibiotic resistance genes, several important MGEs, and transferable plasmids might increase the possibility of co-spreading of blaNDM-19 with other resistance genes.Copyright © 2018 Elsevier Inc. All rights reserved.


April 21, 2020

Characterization and analysis of the transcriptome in Gymnocypris selincuoensis on the Qinghai-Tibetan Plateau using single-molecule long-read sequencing and RNA-seq.

The lakes on the Qinghai-Tibet Plateau (QTP) are the largest and highest lake group in the world. Gymnocypris selincuoensis is the only cyprinid fish living in lake Selincuo, the largest lake on QTP. However, its genetic resource is still blank, limiting studies on molecular and genetic analysis. In this study, the transcriptome of G. selincuoensis was first generated by using PacBio Iso-Seq and Illumina RNA-seq. A full-length (FL) transcriptome with 75,435 transcripts was obtained by Iso-Seq with N50 length of 3,870 bp. Among all transcripts, 75,016 were annotated to public databases, 64,710 contain complete open reading frames and 2,811 were long non-coding RNAs. Based on all- vs.-all BLAST, 2,069 alternative splicing events were detected, and 80% of them were validated by reverse transcription polymerase chain reaction (RT-PCR). Tissue gene expression atlas showed that the number of detected expressed transcripts ranged from 37,397 in brain to 19,914 in muscle, with 10,488 transcripts detected in all seven tissues. Comparative genomic analysis with other cyprinid fishes identified 77 orthologous genes with potential positive selection (Ka/Ks > 0.3). A total of 56,696 perfect simple sequence repeats were identified from FL transcripts. Our results provide valuable genetic resources for further studies on adaptive evolution, gene expression and population genetics in G. selincuoensis and other congeneric fishes. © The Author(s) 2019. Published by Oxford University Press on behalf of Kazusa DNA Research Institute.


April 21, 2020

rMETL: sensitive mobile element insertion detection with long read realignment.

Mobile element insertion (MEI) is a major category of structure variations (SVs). The rapid development of long read sequencing technologies provides the opportunity to detect MEIs sensitively. However, the signals of MEI implied by noisy long reads are highly complex due to the repetitiveness of mobile elements as well as the high sequencing error rates. Herein, we propose the Realignment-based Mobile Element insertion detection Tool for Long read (rMETL). Benchmarking results of simulated and real datasets demonstrate that rMETL enables to handle the complex signals to discover MEIs sensitively. It is suited to produce high-quality MEI callsets in many genomics studies.rMETL is available from https://github.com/hitbc/rMETL.Supplementary data are available at Bioinformatics online. © The Author(s) 2019. Published by Oxford University Press. All rights reserved. For Permissions, please e-mail: journals.permissions@oup.com.


April 21, 2020

Dynamic virulence-related regions of the plant pathogenic fungus Verticillium dahliae display enhanced sequence conservation.

Plant pathogens continuously evolve to evade host immune responses. During host colonization, many fungal pathogens secrete effectors to perturb such responses, but these in turn may become recognized by host immune receptors. To facilitate the evolution of effector repertoires, such as the elimination of recognized effectors, effector genes often reside in genomic regions that display increased plasticity, a phenomenon that is captured in the two-speed genome hypothesis. The genome of the vascular wilt fungus Verticillium dahliae displays regions with extensive presence/absence polymorphisms, so-called lineage-specific regions, that are enriched in in planta-induced putative effector genes. As expected, comparative genomics reveals differential degrees of sequence divergence between lineage-specific regions and the core genome. Unanticipated, lineage-specific regions display markedly higher sequence conservation in coding as well as noncoding regions than the core genome. We provide evidence that disqualifies horizontal transfer to explain the observed sequence conservation and conclude that sequence divergence occurs at a slower pace in lineage-specific regions of the V. dahliae genome. We hypothesize that differences in chromatin organisation may explain lower nucleotide substitution rates in the plastic, lineage-specific regions of V. dahliae. © 2019 The Authors. Molecular Ecology Published by John Wiley & Sons Ltd.


April 21, 2020

Full-length transcriptome analysis of Litopenaeus vannamei reveals transcript variants involved in the innate immune system.

To better understand the immune system of shrimp, this study combined PacBio isoform sequencing (Iso-Seq) and Illumina paired-end short reads sequencing methods to discover full-length immune-related molecules of the Pacific white shrimp, Litopenaeus vannamei. A total of 72,648 nonredundant full-length transcripts (unigenes) were generated with an average length of 2545 bp from five main tissues, including the hepatopancreas, cardiac stomach, heart, muscle, and pyloric stomach. These unigenes exhibited a high annotation rate (62,164, 85.57%) when compared against NR, NT, Swiss-Prot, Pfam, GO, KEGG and COG databases. A total of 7544 putative long noncoding RNAs (lncRNAs) were detected and 1164 nonredundant full-length transcripts (449 UniTransModels) participated in the alternative splicing (AS) events. Importantly, a total of 5279 nonredundant full-length unigenes were successfully identified, which were involved in the innate immune system, including 9 immune-related processes, 19 immune-related pathways and 10 other immune-related systems. We also found wide transcript variants, which increased the number and function complexity of immune molecules; for example, toll-like receptors (TLRs) and interferon regulatory factors (IRFs). The 480 differentially expressed genes (DEGs) were significantly higher or tissue-specific expression patterns in the hepatopancreas compared with that in other four tested tissues (FDR <0.05). Furthermore, the expression levels of six selected immune-related DEGs and putative IRFs were validated using real-time PCR technology, substantiating the reliability of the PacBio Iso-seq results. In conclusion, our results provide new genetic resources of long-read full-length transcripts data and information for identifying immune-related genes, which are an invaluable transcriptomic resource as genomic reference, especially for further exploration of the innate immune and defense mechanisms of shrimp. Copyright © 2019 Elsevier Ltd. All rights reserved.


April 21, 2020

Hybrid sequencing-based personal full-length transcriptomic analysis implicates proteostatic stress in metastatic ovarian cancer.

Comprehensive molecular characterization of myriad somatic alterations and aberrant gene expressions at personal level is key to precision cancer therapy, yet limited by current short-read sequencing technology, individualized catalog of complete genomic and transcriptomic features is thus far elusive. Here, we integrated second- and third-generation sequencing platforms to generate a multidimensional dataset on a patient affected by metastatic epithelial ovarian cancer. Whole-genome and hybrid transcriptome dissection captured global genetic and transcriptional variants at previously unparalleled resolution. Particularly, single-molecule mRNA sequencing identified a vast array of unannotated transcripts, novel long noncoding RNAs and gene chimeras, permitting accurate determination of transcription start, splice, polyadenylation and fusion sites. Phylogenetic and enrichment inference of isoform-level measurements implicated early functional divergence and cytosolic proteostatic stress in shaping ovarian tumorigenesis. A complementary imaging-based high-throughput drug screen was performed and subsequently validated, which consistently pinpointed proteasome inhibitors as an effective therapeutic regime by inducing protein aggregates in ovarian cancer cells. Therefore, our study suggests that clinical application of the emerging long-read full-length analysis for improving molecular diagnostics is feasible and informative. An in-depth understanding of the tumor transcriptome complexity allowed by leveraging the hybrid sequencing approach lays the basis to reveal novel and valid therapeutic vulnerabilities in advanced ovarian malignancies.


April 21, 2020

PacBio full-length cDNA sequencing integrated with RNA-seq reads drastically improves the discovery of splicing transcripts in rice.

In eukaryotes, alternative splicing (AS) greatly expands the diversity of transcripts. However, it is challenging to accurately determine full-length splicing isoforms. Recently, more studies have taken advantage of Pacific Bioscience (PacBio) long-read sequencing to identify full-length transcripts. Nevertheless, the high error rate of PacBio reads seriously offsets the advantages of long reads, especially for accurately identifying splicing junctions. To best capitalize on the features of long reads, we used Illumina RNA-seq reads to improve PacBio circular consensus sequence (CCS) quality and to validate splicing patterns in the rice transcriptome. We evaluated the impact of CCS accuracy on the number and the validation rate of splicing isoforms, and integrated a comprehensive pipeline of splicing transcripts analysis by Iso-Seq and RNA-seq (STAIR) to identify the full-length multi-exon isoforms in rice seedling transcriptome (Oryza sativa L. ssp. japonica). STAIR discovered 11 733 full-length multi-exon isoforms, 6599 more than the SMRT Portal RS_IsoSeq pipeline did. Of these splicing isoforms identified, 4453 (37.9%) were missed in assembled transcripts from RNA-seq reads, and 5204 (44.4%), including 268 multi-exon long non-coding RNAs (lncRNAs), were not reported in the MSU_osa1r7 annotation. Some randomly selected unreported splicing junctions were verified by polymerase chain reaction (PCR) amplification. In addition, we investigated alternative polyadenylation (APA) events in transcripts and identified 829 major polyadenylation [poly(A)] site clusters (PACs). The analysis of splicing isoforms and APA events will facilitate the annotation of the rice genome and studies on the expression and polyadenylation of AS genes in different developmental stages or growth conditions of rice. © 2018 The Authors The Plant Journal © 2018 John Wiley & Sons Ltd.


April 21, 2020

Fudania jinshanensis gen. nov., sp. nov., isolated from faeces of the Tibetan antelope (Pantholops hodgsonii) in China.

Two hitherto unknown bacteria (strains 313T and 352) were recovered from the faeces of Tibetan antelopes on the Tibet-Qinghai Plateau, PR China. Cells were rod-shaped and Gram-stain-positive. The optimal growth conditions were at 37?°C and pH 7. The isolates were closely related to Actinotignum sanguinis (92.6?% 16S rRNA gene sequence similarity), Arcanobacterium haemolyticum (92.5?%), Actinotignum schaalii (92.4?%), Actinobaculum massiliense (92.2?%) and Flaviflexus huanghaiensis (91.6?%). Phylogenetic analyses showed that strains 313T and 352 clustered independently in the vicinity of the genera Actinotignum, Actinobaculum and Flaviflexus, but could not be classified clearly as a member of any of these genera. Phylogenomic analysis also indicated that strains 313T and 352 formed an independent branch in the family Actinomycetaceae. The major cellular fatty acids of the strains were C16?:?0 and C18?:?1?9c. The polar lipids comprised diphosphatidylglycerol, phosphatidylinositol mannoside, phosphatidylglycerol, phosphatidylinositol and five unidentified components. The peptidoglycan contained lysine, alanine and glutamic acid. The respiratory quinone was absent. The whole-cell sugars included glucose and rhamnose. The DNA G+C?content of strain 313T was 60.6?mol%. Based on the low 16S rRNA gene sequence similarities, its taxonomic position in the phylogenetic and phylogenomic trees and its unique lipid pattern, we propose that strains 313T and 352 represent members of a novel species in a new genus, for which the name Fudania jinshanensis gen. nov., sp. nov. is proposed. The type strain is 313T (=CGMCC 4.7453T=DSM 106216T).


April 21, 2020

A prophage and two ICESa2603-family integrative and conjugative elements (ICEs) carrying optrA in Streptococcus suis.

To investigate the presence and transfer of the oxazolidinone/phenicol resistance gene optrA and identify the genetic elements involved in the horizontal transfer of the optrA gene in Streptococcus suis.A total of 237 S. suis isolates were screened for the presence of the optrA gene by PCR. Whole-genome DNA of three optrA-positive strains was completely sequenced using the Illumina MiSeq and Pacbio RSII platforms. MICs were determined by broth microdilution. Transferability of the optrA gene in S. suis was investigated by conjugation. The presence of circular intermediates was examined by inverse PCR.The optrA gene was present in 11.8% (28/237) of the S. suis strains. In three strains, the optrA gene was flanked by two copies of IS1216 elements in the same orientation, located either on a prophage or on ICESa2603-family integrative and conjugative elements (ICEs), including one tandem ICE. In one isolate, the optrA-carrying ICE transferred with a frequency of 2.1?×?10-8. After the transfer, the transconjugant displayed elevated MICs of the respective antimicrobial agents. Inverse PCRs revealed that circular intermediates of different sizes were formed in the three optrA-carrying strains, containing one copy of the IS1216E element and the optrA gene alone or in combination with other resistance genes.A prophage and two ICESa2603-family ICEs (including one tandem ICE) associated with the optrA gene were identified in S. suis. The association of the optrA gene with the IS1216E elements and its location on either a prophage or ICEs will aid its horizontal transfer. © The Author(s) 2019. Published by Oxford University Press on behalf of the British Society for Antimicrobial Chemotherapy. All rights reserved. For permissions, please email: journals.permissions@oup.com.


April 21, 2020

Complete mitochondrial genome of Hemiptelea davidii (Ulmaceae) and phylogenetic analysis

Hemiptelea davidii (Hance) Planch is a potential valuable forest tree in arid sandy environments. Here, the complete mitochondrial genome of H. davidii was assembled using a combination of the PacBio Sequel data and the Illumina Hiseq data. The mitochondrial genome is 460,941bp in length, including 37 protein-coding genes, 19 tRNA genes, and three rRNA genes. The GC content of the whole mito- chondrial genome is 44.84%. Phylogenetic analyses indicated that H. davidii is close with Cannabis and Morus species.


Talk with an expert

If you have a question, need to check the status of an order, or are interested in purchasing an instrument, we're here to help.