Single-Molecule Real-Time (SMRT) DNA sequencing is unique in that nucleotide incorporation events are monitored in real time, leading to a wealth of kinetic information in addition to the extraction of the primary DNA sequence. The dynamics of the DNA polymerase that is observed adds an additional dimension of sequence-dependent information, and can be used to learn more about the molecule under study. First, the primary sequence itself can be determined more accurately. The kinetic data can be used to corroborate or overturn consensus calls and even enable calling bases in problematic sequence contexts. Second, using the kinetic information, we can…
The assembly of metagenomes is dramatically improved by the long read lengths of SMRT Sequencing. This is demonstrated in an experimental design to sequence a mock community from the Human Microbiome Project, and assemble the data using the hierarchical genome assembly process (HGAP) at Pacific Biosciences. Results of this analysis are promising, and display much improved contiguity in the assembly of the mock community as compared to publicly available short-read data sets and assemblies. Additionally, the use of base modification information to make further associations between contigs provides additional data to improve assemblies, and to distinguish between members within a…
PacBio 2014 User Group Meeting Presentation Slides: Alisha Holloway of the Gladstone Institutes presented on the use of isoform sequencing (Iso-Seq) to improve the annotation of the chicken genome as a model reference for cardiovascular research.
Since the advent of Next-Generation Sequencing (NGS), the cost of de novo genome sequencing and assembly have dropped precipitately, which has spurred interest in genome sequencing overall. Unfortunately the contiguity of the NGS assembled sequences, as well as the accuracy of these assemblies have suffered. Additionally, most NGS de novo assemblies leave large portions of genomes unresolved, and repetitive regions are often collapsed. When compared to the reference quality genome sequences produced before the NGS era, the new sequences are highly fragmented and often prove to be difficult to properly annotate. In some cases the contiguous portions are smaller than…
2015 SMRT Informatics Developers Conference Presentation Slides: Shinichi Morishita of the University of Tokyo presented on how his team has been using SMRT Sequencing to better understand methylomes, metagenomes and structural variation of various eukaryotic genomes.
The human immunoglobulin heavy chain locus (IGH) remains among the most understudied regions of the human genome. Recent efforts have shown that haplotype diversity within IGH is elevated and exhibits population specific patterns; for example, our re-sequencing of the locus from only a single chromosome uncovered >100 Kb of novel sequence, including descriptions of six novel alleles, and four previously unmapped genes. Historically, this complex locus architecture has hindered the characterization of IGH germline single nucleotide, copy number, and structural variants (SNVs; CNVs; SVs), and as a result, there remains little known about the role of IGH polymorphisms in inter-individual…
Over 40% of males and ~16% of female carriers of a FMR1 premutation allele (55-200 CGG repeats) are at risk for developing Fragile X-associated Tremor/Ataxia Syndrome (FXTAS), an adult onset neurodegenerative disorder while, about 20% of female carriers will develop Fragile X-associated Primary Ovarian Insufficiency (FXPOI), in addition to a number of adult-onset clinical problems (FMR1 associated disorders). Marked elevation in FMR1 mRNA levels have been observed with premutation alleles and the resulting RNA toxicity is believed to be the leading molecular mechanism proposed for these disorders. The FMR1 gene, as many housekeeping genes, undergoes alternative splicing. Using long-read isoform…
The PacBio Iso-Seq method produces high-quality, full-length transcripts of up to 10 kb and longer and has been used to annotate many important plant and animal genomes. Here we describe an improved, simplified library workflow and analysis pipeline that reduces library preparation time, RNA input, and cost. The Iso-Seq V2 Express workflow is a one day protocol that requires only ~300 ng of total RNA input while also reducing the number of reverse transcription and amplification steps down to single reactions. Compared with the previous workflow, the Iso-Seq V2 Express workflow increases the percentage of full-length (FL) reads while achieving…
Single cell RNA-seq (scRNA-seq) is an emerging field for characterizing cell heterogeneity in complex tissues. However, most scRNA-seq methodologies are limited to gene count information due to short read lengths. Here, we combine the microfluidics scRNA-seq technique, Drop-Seq, with PacBio Single Molecule, Real-Time (SMRT) Sequencing to generate full-length transcript isoforms that can be confidently assigned to individual cells. We generated single cell Iso-Seq (scIso-Seq) libraries for chimp and human cerebral organoid samples on the Dolomite Nadia platform and sequenced each library with two SMRT Cells 8M on the PacBio Sequel II System. We developed a bioinformatics pipeline to identify, classify,…
With PacBio single-cell RNA sequencing using the Iso-Seq method, you can now distinguish between alternative transcript isoforms at the single-cell level. The highly accurate long reads (HiFi reads) can span the entire 5′ to 3′ end of a transcript, allowing a high-resolution view of isoform diversity and revealing cell-to-cell heterogeneity without the need for assembly.
In this ASHG 2016 virtual poster, Flora Tassone from UC Davis describes her study of the molecular mechanisms linked to fragile X syndrome and associated disorders, such as FXTAS. She is using SMRT Sequencing to resolve the FMR1 gene in premutation carriers because it’s the only technology that can generate full-length transcripts with the causative CGG repeat expansion. Plus: direct confirmation of predicted isoform configurations.
Michael Lutz, from the Duke University Medical Center, discussed a recently published software tool that can now be used in a pipeline with SMRT Sequencing data to find structural variant biomarkers for neurodegenerative diseases with a focus on Alzheimer’s disease, ALS, and Lewy body dementia. His team is particularly interested in short sequence repeats and short tandem repeats, which have already been implicated in neurodegenerative disease.
At AGBT 2017, Margaret Roy from Calico Life Sciences discussed a de novo genome sequencing effort for the naked mole rat. This animal has a remarkably long life span and resistance to cancer, both of which make it interesting for studies of life extension. The team is using SMRT Sequencing for a more complete, contiguous assembly than the two existing short-read-based assemblies. Included: data from the Sequel System.
In this webinar, Emily Hatas of PacBio shares information about the applications and benefits of SMRT Sequencing in plant and animal biology, agriculture, and industrial research fields. This session contains an overview of several applications: whole-genome sequencing for de novo assembly; transcript isoform sequencing (Iso-Seq) method for genome annotation; targeted sequencing solutions; and metagenomics and microbial interactions. High-level workflows and best practices are discussed for key applications.
Long-read sequencing technologies like Iso-Seq analysis present researchers with a powerful tool for probing the transcriptomes of many species. The ability to sequence transcripts from end-to-end has revealed transcription complexity on a scale that was previously impossible. This sequence rich information has also improved our ability to predict transcript functions and biotypes. Researchers can now use Iso-Seq analysis to discover transcript models in almost any species with an accuracy on par with human and mouse annotations. In this webinar, Richard Kuo discusses the core concepts behind Iso-Seq analysis and how to use it to improve or build a new transcriptome…