Menu
September 22, 2019

Analysis of structural variants in four African cichlids highlights an association with developmental and immune related genes

African Lakes Cichlids are one of the most impressive example of adaptive radiation. Independently in Lake Victoria, Tanganyika, and Malawi, several hundreds of species arose within the last 10 million to 100,000 years. Whereas most analyses in cichlids focused on nucleotide substitutions across species to investigate the genetic bases of this explosive radiation, to date, no study has investigated the contribution of structural variants (SVs) to speciation events (through a reduction of gene flow) and adaptation to different ecological niches. Here, we annotate and characterize the repertoires and evolutionary potential of different SV classes (deletion, duplication, inversion, insertions and translocations) in five cichlid species (Astatotilapia burtoni, Metriaclima zebra, Neolamprologus brichardi, Pundamilia nyererei and Oreochromis niloticus). We investigate the patterns of gain/loss evolution across the phylogeny for each SV type enabling the identification of both lineage specific events and a set of conserved SVs, common to all four species in the radiation. Both deletion and inversion events show a significant overlap with SINE elements, while inversions additionally show a limited, but significant association with DNA transposons. Genes lying inside inverted regions are enriched for genes regulating behaviour, or involved in skeletal and visual system development. Moreover, we find that duplicated genes show enrichment for textquoterightantigen processing and presentationtextquoteright (GO:0019882) and other immune related categories. Altogether, we provide the first, comprehensive overview of rearrangement evolution in East African Cichlids, and some initial insights into their possible contribution to adaptation.


September 22, 2019

Emergence of pathogenic and multiple-antibiotic-resistant Macrococcus caseolyticus in commercial broiler chickens.

Macrococcus caseolyticus is generally considered to be a non-pathogenic bacterium that does not cause human or animal diseases. However, recently, a strain of M. caseolyticus (SDLY strain) that causes high mortality rates was isolated from commercial broiler chickens in China. The main pathological changes caused by SDLY included caseous exudation in cranial cavities, inflammatory infiltration, haemorrhages and multifocal necrosis in various organs. The whole genome of the SDLY strain was sequenced and was compared with that of the non-pathogenic JCSC5402 strain of M. caseolyticus. The results showed that the SDLY strain harboured a large quantity of mutations, antibiotic resistance genes and numerous insertions and deletions of virulence genes. In particular, among the inserted genes, there is a cluster of eight connected genes associated with the synthesis of capsular polysaccharide. This cluster encodes a transferase and capsular polysaccharide synthase, promotes the formation of capsules and causes changes in pathogenicity. Electron microscopy revealed a distinct capsule surrounding the SDLY strain. The pathogenicity test showed that the SDLY strain could cause significant clinical symptoms and pathological changes in both SPF chickens and mice. In addition, these clinical symptoms and pathological changes were the same as those observed in field cases. Furthermore, the anti-microbial susceptibility test demonstrated that the SDLY strain exhibits multiple-antibiotic resistance. The emergence of pathogenic M. caseolyticus indicates that more attention should be paid to the effects of this micro-organism on both poultry and public health.© 2018 Blackwell Verlag GmbH.


September 22, 2019

DNA Methylation by Restriction Modification Systems Affects the Global Transcriptome Profile in Borrelia burgdorferi.

Prokaryote restriction modification (RM) systems serve to protect bacteria from potentially detrimental foreign DNA. Recent evidence suggests that DNA methylation by the methyltransferase (MTase) components of RM systems can also have effects on transcriptome profiles. The type strain of the causative agent of Lyme disease, Borrelia burgdorferi B31, possesses two RM systems with N6-methyladenosine (m6A) MTase activity, which are encoded by the bbe02 gene located on linear plasmid lp25 and bbq67 on lp56. The specific recognition and/or methylation sequences had not been identified for either of these B. burgdorferi MTases, and it was not previously known whether these RM systems influence transcript levels. In the current study, single-molecule real-time sequencing was utilized to map genome-wide m6A sites and to identify consensus modified motifs in wild-type B. burgdorferi as well as MTase mutants lacking either the bbe02 gene alone or both bbe02 and bbq67 genes. Four novel conserved m6A motifs were identified and were fully attributable to the presence of specific MTases. Whole-genome transcriptome changes were observed in conjunction with the loss of MTase enzymes, indicating that DNA methylation by the RM systems has effects on gene expression. Genes with altered transcription in MTase mutants include those involved in vertebrate host colonization (e.g., rpoS regulon) and acquisition by/transmission from the tick vector (e.g., rrp1 and pdeB). The results of this study provide a comprehensive view of the DNA methylation pattern in B. burgdorferi, and the accompanying gene expression profiles add to the emerging body of research on RM systems and gene regulation in bacteria.IMPORTANCE Lyme disease is the most prevalent vector-borne disease in North America and is classified by the Centers for Disease Control and Prevention (CDC) as an emerging infectious disease with an expanding geographical area of occurrence. Previous studies have shown that the causative bacterium, Borrelia burgdorferi, methylates its genome using restriction modification systems that enable the distinction from foreign DNA. Although much research has focused on the regulation of gene expression in B. burgdorferi, the effect of DNA methylation on gene regulation has not been evaluated. The current study characterizes the patterns of DNA methylation by restriction modification systems in B. burgdorferi and evaluates the resulting effects on gene regulation in this important pathogen. Copyright © 2018 American Society for Microbiology.


September 22, 2019

The genome of the tegu lizard Salvator merianae: combining Illumina, PacBio, and optical mapping data to generate a highly contiguous assembly.

Reptiles are a species-rich group with great phenotypic and life history diversity but are highly underrepresented among the vertebrate species with sequenced genomes.Here, we report a high-quality genome assembly of the tegu lizard, Salvator merianae, the first lacertoid with a sequenced genome. We combined 74X Illumina short-read, 29.8X Pacific Biosciences long-read, and optical mapping data to generate a high-quality assembly with a scaffold N50 value of 55.4 Mb. The contig N50 value of this assembly is 521 Kb, making it the most contiguous reptile assembly so far. We show that the tegu assembly has the highest completeness of coding genes and conserved non-exonic elements (CNEs) compared to other reptiles. Furthermore, the tegu assembly has the highest number of evolutionarily conserved CNE pairs, corroborating a high assembly contiguity in intergenic regions. As in other reptiles, long interspersed nuclear elements comprise the most abundant transposon class. We used transcriptomic data, homology- and de novo gene predictions to annotate 22,413 coding genes, of which 16,995 (76%) likely have human orthologs as inferred by CESAR-derived gene mappings. Finally, we generated a multiple genome alignment comprising 10 squamates and 7 other amniote species and identified conserved regions that are under evolutionary constraint. CNEs cover 38 Mb (1.8%) of the tegu genome, with 3.3 Mb in these elements being squamate specific. In contrast to placental mammal-specific CNEs, very few of these squamate-specific CNEs (<20 Kb) overlap transposons, highlighting a difference in how lineage-specific CNEs originated in these two clades.The tegu lizard genome together with the multiple genome alignment and comprehensive conserved element datasets provide a valuable resource for comparative genomic studies of reptiles and other amniotes.


September 22, 2019

Regulation of yeast-to-hyphae transition in Yarrowia lipolytica.

The yeast Yarrowia lipolytica undergoes a morphological transition from yeast-to-hyphal growth in response to environmental conditions. A forward genetic screen was used to identify mutants that reliably remain in the yeast phase, which were then assessed by whole-genome sequencing. All the smooth mutants identified, so named because of their colony morphology, exhibit independent loss of DNA at a repetitive locus made up of interspersed ribosomal DNA and short 10- to 40-mer telomere-like repeats. The loss of repetitive DNA is associated with downregulation of genes with stress response elements (5′-CCCCT-3′) and upregulation of genes with cell cycle box (5′-ACGCG-3′) motifs in their promoter region. The stress response element is bound by the transcription factor Msn2p in Saccharomyces cerevisiae We confirmed that the Y. lipolyticamsn2 (Ylmsn2) ortholog is required for hyphal growth and found that overexpression of Ylmsn2 enables hyphal growth in smooth strains. The cell cycle box is bound by the Mbp1p/Swi6p complex in S. cerevisiae to regulate G1-to-S phase progression. We found that overexpression of either the Ylmbp1 or Ylswi6 homologs decreased hyphal growth and that deletion of either Ylmbp1 or Ylswi6 promotes hyphal growth in smooth strains. A second forward genetic screen for reversion to hyphal growth was performed with the smooth-33 mutant to identify additional genetic factors regulating hyphal growth in Y. lipolytica Thirteen of the mutants sequenced from this screen had coding mutations in five kinases, including the histidine kinases Ylchk1 and Ylnik1 and kinases of the high-osmolarity glycerol response (HOG) mitogen-activated protein (MAP) kinase cascade Ylssk2, Ylpbs2, and Ylhog1 Together, these results demonstrate that Y. lipolytica transitions to hyphal growth in response to stress through multiple signaling pathways.IMPORTANCE Many yeasts undergo a morphological transition from yeast-to-hyphal growth in response to environmental conditions. We used forward and reverse genetic techniques to identify genes regulating this transition in Yarrowia lipolytica We confirmed that the transcription factor Ylmsn2 is required for the transition to hyphal growth and found that signaling by the histidine kinases Ylchk1 and Ylnik1 as well as the MAP kinases of the HOG pathway (Ylssk2, Ylpbs2, and Ylhog1) regulates the transition to hyphal growth. These results suggest that Y. lipolytica transitions to hyphal growth in response to stress through multiple kinase pathways. Intriguingly, we found that a repetitive portion of the genome containing telomere-like and rDNA repeats may be involved in the transition to hyphal growth, suggesting a link between this region and the general stress response. Copyright © 2018 Pomraning et al.


September 22, 2019

Comparative genomics of 84 Pectobacterium genomes reveals the variations related to a pathogenic lifestyle.

Pectobacterium spp. are necrotrophic bacterial plant pathogens of the family Pectobacteriaceae, responsible for a wide spectrum of diseases of important crops and ornamental plants including soft rot, blackleg, and stem wilt. P. carotovorum is a genetically heterogeneous species consisting of three valid subspecies, P. carotovorum subsp. brasiliense (Pcb), P. carotovorum subsp. carotovorum (Pcc), and P. carotovorum subsp. odoriferum (Pco).Thirty-two P. carotovorum strains had their whole genomes sequenced, including the first complete genome of Pco and another circular genome of Pcb, as well as the high-coverage genome sequences for 30 additional strains covering Pcc, Pcb, and Pco. In combination with 52 other publicly available genome sequences, the comparative genomics study of P. carotovorum and other four closely related species P. polaris, P. parmentieri, P. atrosepticum, and Candidatus P. maceratum was conducted focusing on CRISPR-Cas defense systems and pathogenicity determinants. Our analysis identified two CRISPR-Cas types (I-F and I-E) in Pectobacterium, as well as another I-C type in Dickeya that is not found in Pectobacterium. The core pathogenicity factors (e.g., plant cell wall-degrading enzymes) were highly conserved, whereas some factors (e.g., flagellin, siderophores, polysaccharides, protein secretion systems, and regulatory factors) were varied among these species and/or subspecies. Notably, a novel type of T6SS as well as the sorbitol metabolizing srl operon was identified to be specific to Pco in Pectobacterium.This study not only advances the available knowledge about the genetic differentiation of individual subspecies of P. carotovorum, but also delineates the general genetic features of P. carotovorum by comparison with its four closely related species, thereby substantially enriching the extent of information now available for functional genomic investigations about Pectobacterium.


September 22, 2019

Trophoblast organoids as a model for maternal-fetal interactions during human placentation.

The placenta is the extraembryonic organ that supports the fetus during intrauterine life. Although placental dysfunction results in major disorders of pregnancy with immediate and lifelong consequences for the mother and child, our knowledge of the human placenta is limited owing to a lack of functional experimental models1. After implantation, the trophectoderm of the blastocyst rapidly proliferates and generates the trophoblast, the unique cell type of the placenta. In vivo, proliferative villous cytotrophoblast cells differentiate into two main sub-populations: syncytiotrophoblast, the multinucleated epithelium of the villi responsible for nutrient exchange and hormone production, and extravillous trophoblast cells, which anchor the placenta to the maternal decidua and transform the maternal spiral arteries2. Here we describe the generation of long-term, genetically stable organoid cultures of trophoblast that can differentiate into both syncytiotrophoblast and extravillous trophoblast. We used human leukocyte antigen (HLA) typing to confirm that the organoids were derived from the fetus, and verified their identities against four trophoblast-specific criteria3. The cultures organize into villous-like structures, and we detected the secretion of placental-specific peptides and hormones, including human chorionic gonadotropin (hCG), growth differentiation factor 15 (GDF15) and pregnancy-specific glycoprotein (PSG) by mass spectrometry. The organoids also differentiate into HLA-G+ extravillous trophoblast cells, which vigorously invade in three-dimensional cultures. Analysis of the methylome reveals that the organoids closely resemble normal first trimester placentas. This organoid model will be transformative for studying human placental development and for investigating trophoblast interactions with the local and systemic maternal environment.


September 21, 2019

Whole genome sequence of the soybean aphid, Aphis glycines.

Aphids are emerging as model organisms for both basic and applied research. Of the 5,000 estimated species, only three aphids have published whole genome sequences: the pea aphid Acyrthosiphon pisum, the Russian wheat aphid, Diuraphis noxia, and the green peach aphid, Myzus persicae. We present the whole genome sequence of a fourth aphid, the soybean aphid (Aphis glycines), which is an extreme specialist and an important invasive pest of soybean (Glycine max). The availability of genomic resources is important to establish effective and sustainable pest control, as well as to expand our understanding of aphid evolution. We generated a 302.9 Mbp draft genome assembly for Ap. glycines using a hybrid sequencing approach. This assembly shows high completeness with 19,182 predicted genes, 92% of known Ap. glycines transcripts mapping to contigs, and substantial continuity with a scaffold N50 of 174,505 bp. The assembly represents 95.5% of the predicted genome size of 317.1 Mbp based on flow cytometry. Ap. glycines contains the smallest known aphid genome to date, based on updated genome sizes for 19 aphid species. The repetitive DNA content of the Ap. glycines genome assembly (81.6 Mbp or 26.94% of the 302.9 Mbp assembly) shows a reduction in the number of classified transposable elements compared to Ac. pisum, and likely contributes to the small estimated genome size. We include comparative analyses of gene families related to host-specificity (cytochrome P450’s and effectors), which may be important in Ap. glycines evolution. This Ap. glycines draft genome sequence will provide a resource for the study of aphid genome evolution, their interaction with host plants, and candidate genes for novel insect control methods. Copyright © 2017 Elsevier Ltd. All rights reserved.


September 21, 2019

Retrotransposons are the major contributors to the expansion of the Drosophila ananassae Muller F element.

The discordance between genome size and the complexity of eukaryotes can partly be attributed to differences in repeat density. The Muller F element (~5.2 Mb) is the smallest chromosome in Drosophila melanogaster, but it is substantially larger (>18.7 Mb) in D. ananassae To identify the major contributors to the expansion of the F element and to assess their impact, we improved the genome sequence and annotated the genes in a 1.4-Mb region of the D. ananassae F element, and a 1.7-Mb region from the D element for comparison. We find that transposons (particularly LTR and LINE retrotransposons) are major contributors to this expansion (78.6%), while Wolbachia sequences integrated into the D. ananassae genome are minor contributors (0.02%). Both D. melanogaster and D. ananassae F-element genes exhibit distinct characteristics compared to D-element genes (e.g., larger coding spans, larger introns, more coding exons, and lower codon bias), but these differences are exaggerated in D. ananassae Compared to D. melanogaster, the codon bias observed in D. ananassae F-element genes can primarily be attributed to mutational biases instead of selection. The 5′ ends of F-element genes in both species are enriched in dimethylation of lysine 4 on histone 3 (H3K4me2), while the coding spans are enriched in H3K9me2. Despite differences in repeat density and gene characteristics, D. ananassae F-element genes show a similar range of expression levels compared to genes in euchromatic domains. This study improves our understanding of how transposons can affect genome size and how genes can function within highly repetitive domains. Copyright © 2017 Leung et al.


September 21, 2019

PacBio assembly of a Plasmodium knowlesi genome sequence with Hi-C correction and manual annotation of the SICAvar gene family.

Plasmodium knowlesi has risen in importance as a zoonotic parasite that has been causing regular episodes of malaria throughout South East Asia. The P. knowlesi genome sequence generated in 2008 highlighted and confirmed many similarities and differences in Plasmodium species, including a global view of several multigene families, such as the large SICAvar multigene family encoding the variant antigens known as the schizont-infected cell agglutination proteins. However, repetitive DNA sequences are the bane of any genome project, and this and other Plasmodium genome projects have not been immune to the gaps, rearrangements and other pitfalls created by these genomic features. Today, long-read PacBio and chromatin conformation technologies are overcoming such obstacles. Here, based on the use of these technologies, we present a highly refined de novo P. knowlesi genome sequence of the Pk1(A+) clone. This sequence and annotation, referred to as the ‘MaHPIC Pk genome sequence’, includes manual annotation of the SICAvar gene family with 136 full-length members categorized as type I or II. This sequence provides a framework that will permit a better understanding of the SICAvar repertoire, selective pressures acting on this gene family and mechanisms of antigenic variation in this species and other pathogens.


September 21, 2019

Multi-Locus Variable number of tandem repeat Analysis (MLVA) of Yersinia ruckeri confirms the existence of host-specificity, geographic endemism and anthropogenic dissemination of virulent clones.

A Multi-Locus Variable number of tandem repeat Analysis (MLVA) assay was developed for epizootiological study of the internationally significant fish pathogen Yersinia ruckeri, which causes yersiniosis in salmonids. The assay involves amplification of ten Variable Number of Tandem Repeat (VNTR) loci in two five-plex PCR reactions, followed by capillary electrophoresis. A collection of 484 Y. ruckeri isolates, originating from various biological sources and collected from four continents over seven decades, was analysed. Minimum spanning tree cluster analysis of MLVA profiles separated the studied population into nine major clonal complexes, and a number of minor clusters and singletons. The major clonal complexes could be associated with host species, geographic origin and serotype. A single large clonal complex of serotype O1 isolates dominating the yersiniosis situation in international rainbow trout farming suggests anthropogenic spread of this clone, possibly related to transport of fish. Moreover, sub-clustering within this clonal complex indicates putative transmission routes and multiple biotype shift events. In contrast to the situation in rainbow trout, Y. ruckeri strains associated with disease in Atlantic salmon appear as more or less geographically isolated clonal complexes. A single complex of serotype O1 exclusive to Norway was found to be responsible for almost all major yersiniosis outbreaks in modern Norwegian salmon farming, and site-specific sub-clustering further indicates persistent colonisation of freshwater farms in Norway. Identification of genetically diverse Y. ruckeri isolates from clinically healthy fish and environmental sources also suggests the widespread existence of less virulent or avirulent strains.Importance This comprehensive population study substantially improves our understanding of the epizootiological history and nature of an internationally important fish pathogenic bacterium. The MLVA assay developed and presented represents a high-resolution typing tool particularly well suited for Yersinia ruckeri infection tracing, selection of strains for vaccine inclusion, and risk assessment. The ability of the assay to separate isolates into geographically linked and/or possibly host-specific clusters reflects its potential utility for maintenance of national biosecurity. The MLVA is internationally applicable, robust, and provides clear, unambiguous and easily interpreted results. Typing is reasonably inexpensive, with a moderate technological requirement, and may be completed from a harvested colony within a single working day. As the resulting MLVA profiles are readily portable, any Y. ruckeri strain may rapidly be placed in a global epizootiological context. Copyright © 2018 Gulla et al.


September 21, 2019

Assembling large genomes with single-molecule sequencing and locality-sensitive hashing.

Long-read, single-molecule real-time (SMRT) sequencing is routinely used to finish microbial genomes, but available assembly methods have not scaled well to larger genomes. We introduce the MinHash Alignment Process (MHAP) for overlapping noisy, long reads using probabilistic, locality-sensitive hashing. Integrating MHAP with the Celera Assembler enabled reference-grade de novo assemblies of Saccharomyces cerevisiae, Arabidopsis thaliana, Drosophila melanogaster and a human hydatidiform mole cell line (CHM1) from SMRT sequencing. The resulting assemblies are highly continuous, include fully resolved chromosome arms and close persistent gaps in these reference genomes. Our assembly of D. melanogaster revealed previously unknown heterochromatic and telomeric transition sequences, and we assembled low-complexity sequences from CHM1 that fill gaps in the human GRCh38 reference. Using MHAP and the Celera Assembler, single-molecule sequencing can produce de novo near-complete eukaryotic assemblies that are 99.99% accurate when compared with available reference genomes.


September 21, 2019

Discovery and genotyping of structural variation from long-read haploid genome sequence data.

In an effort to more fully understand the full spectrum of human genetic variation, we generated deep single-molecule, real-time (SMRT) sequencing data from two haploid human genomes. By using an assembly-based approach (SMRT-SV), we systematically assessed each genome independently for structural variants (SVs) and indels resolving the sequence structure of 461,553 genetic variants from 2 bp to 28 kbp in length. We find that >89% of these variants have been missed as part of analysis of the 1000 Genomes Project even after adjusting for more common variants (MAF > 1%). We estimate that this theoretical human diploid differs by as much as ~16 Mbp with respect to the human reference, with long-read sequencing data providing a fivefold increase in sensitivity for genetic variants ranging in size from 7 bp to 1 kbp compared with short-read sequence data. Although a large fraction of genetic variants were not detected by short-read approaches, once the alternate allele is sequence-resolved, we show that 61% of SVs can be genotyped in short-read sequence data sets with high accuracy. Uncoupling discovery from genotyping thus allows for the majority of this missed common variation to be genotyped in the human population. Interestingly, when we repeat SV detection on a pseudodiploid genome constructed in silico by merging the two haploids, we find that ~59% of the heterozygous SVs are no longer detected by SMRT-SV. These results indicate that haploid resolution of long-read sequencing data will significantly increase sensitivity of SV detection.© 2017 Huddleston et al.; Published by Cold Spring Harbor Laboratory Press.


September 21, 2019

Identification of a novel RASD1 somatic mutation in a USP8-mutated corticotroph adenoma.

Cushing’s disease (CD) is caused by pituitary corticotroph adenomas that secrete excess adrenocorticotropic hormone (ACTH). In these tumors, somatic mutations in the gene USP8 have been identified as recurrent and pathogenic and are the sole known molecular driver for CD. Although other somatic mutations were reported in these studies, their contribution to the pathogenesis of CD remains unexplored. No molecular drivers have been established for a large proportion of CD cases and tumor heterogeneity has not yet been investigated using genomics methods. Also, even in USP8-mutant tumors, a possibility may exist of additional contributing mutations, following a paradigm from other neoplasm types where multiple somatic alterations contribute to neoplastic transformation. The current study utilizes whole-exome discovery sequencing on the Illumina platform, followed by targeted amplicon-validation sequencing on the Pacific Biosciences platform, to interrogate the somatic mutation landscape in a corticotroph adenoma resected from a CD patient. In this USP8-mutated tumor, we identified an interesting somatic mutation in the gene RASD1, which is a component of the corticotropin-releasing hormone receptor signaling system. This finding may provide insight into a novel mechanism involving loss of feedback control to the corticotropin-releasing hormone receptor and subsequent deregulation of ACTH production in corticotroph tumors.


September 21, 2019

Detecting AGG interruptions in females with a FMR1 premutation by long-read Single-Molecule Sequencing: A 1 year clinical experience.

The fragile X syndrome arises from the FMR1 CGG expansion of a premutation (55-200 repeats) to a full mutation allele (>200 repeats) and is the most frequent cause of inherited X-linked intellectual disability. The risk for a premutation to expand to a full mutation allele depends on the repeat length and AGG triplets interrupting this repeat. In genetic counseling it is important to have information on both these parameters to provide an accurate risk estimate to women carrying a premutation allele and weighing up having children. For example, in case of a small risk a woman might opt for a natural pregnancy followed up by prenatal diagnosis while she might choose for preimplantation genetic diagnosis (PGD) if the risk is high. Unfortunately, the detection of AGG interruptions was previously hampered by technical difficulties complicating their use in diagnostics. Therefore we recently developed, validated and implemented a new methodology which uses long-read single-molecule sequencing to identify AGG interruptions in females with a FMR1 premutation. Here we report on the assets of AGG interruption detection by sequencing and the impact of implementing the assay on genetic counseling.


Talk with an expert

If you have a question, need to check the status of an order, or are interested in purchasing an instrument, we're here to help.