October 29, 2021  |  

Resolving Complex Pathogenic Alleles using HiFi Long-Range Amplicon Data and a New Clustering Algorithm

Many genetic diseases are mapped to structurally complex loci. These regions contain highly similar paralogous alleles (>99% identity) that span kilobases within the human genome. Comprehensive screening for pathogenic variants is incomplete and labor intensive using short-reads or optical mapping. In contrast, long-range amplification and PacBio HiFi sequencing fully and directly resolve and phase a wide range of pathogenic variants without inference. To capitalize on the accuracy of HiFi data we designed a new amplicon analysis tool, pbAA. pbAA can rapidly deconvolve a mixture of haplotypes, enabling precise diplotyping, and disease allele classification. 


September 7, 2021  |  

Resolving Complex Pathogenic Alleles using HiFi Long-Range Amplicon Data and a New Clustering Algorithm

Many genetic diseases are mapped to structurally complex loci. These regions contain highly similar paralogous alleles (>99% identity) that span kilobases within the human genome. Comprehensive screening for pathogenic variants is incomplete and labor intensive using short-reads or optical mapping. In contrast, long-range amplification and PacBio HiFi sequencing fully and directly resolve and phase a wide range of pathogenic variants without inference. To capitalize on the accuracy of HiFi data we designed a new amplicon analysis tool, pbAA. pbAA can rapidly deconvolve a mixture of haplotypes, enabling precise diplotyping, and disease allele classification. 


August 19, 2021  |  

Case Study: Pioneering a pan-genome reference collection

At DuPont Pioneer, DNA sequencing is paramount for R&D to reveal the genetic basis for traits of interest in commercial crops such as maize, soybean, sorghum, sunflower, alfalfa, canola, wheat, rice, and others. They cannot afford to wait the years it has historically taken for high-quality reference genomes to be produced. Nor can they rely on a single reference to represent the genetic diversity in its germplasm.


June 1, 2021  |  

Genome sequencing of microbial genomes using Single Molecule Real-time sequencing (SMRT) technology.

In the last year, high-throughput sequencing technologies have progressed from proof-of-concept to production quality. Although each technology is able to produce vast quantities of sequence information, in every case the underlying chemistry limits reads to very short lengths. We present a examining de novo assembly comparison with bacterial genome assembly varying genome size (from 3.1Mb to 7.6Mb) and different G+C contents (from 43% to 71%), respectively. We analyzed Solexa reads, 454 reads and Pacbio RS reads from Streptomyces sp. (Genome size, 7.6 Mb; G+C content, 71%), Psychrobacter sp. (Genome size, 3.5 Mb; G+C content, 43%), Salinibacterium sp. (Genome size, 3.1 Mb; G+C content, 61%) and Frigoribacterium sp. (Genome size, 3.3 Mb; G+C content, 63%). We assembly each bacterial genome using Celera assembler 7.0 with and without PacBio RS reads. We found out that the assemble result with Pacbio RS reads have less contigs and scaffolds, and better N50 values.


June 1, 2021  |  

Genome sequencing of endosymbiotic bacterial Streptomyces sp. from Antartic lichen using Single Molecule Real-time Sequencing (SMRT) technology.

Along with the advent of next-generation sequencing (NGS) techniques, it has become possible to sequence a microbial genome very quickly with high coverage. Recently, PacificBioscience developed single molecule real-time sequencing (SMRT) technology, 3rd generation sequencing platform, which provide much longer (average read length: 1.5Kb) reads without PCR amplification. We did de novo sequencing of Streptomyces sp. using Illumina GAIIx, Roche 454 and PacBio RS system and compared the data. The endosymbiotic bacteria Streptomyces sp. PAMC 26508 was isolated from Antarctic lichen Psoroma sp. that grows attached rocks on Barton Peninsula, King George Island, Antarctica (62, 13’S, 58, 47’W). With 4 SMRT cells, we could get more than 15x coverage of corrected sequence data for de novo assembly. Comparing the performance of other sequencing platforms, PacBio platform could generate data on similar manner with general mid-level GC content organism. In conclusion, PacBio RS system, SMRT technology, shows better performance with high GC content organisms and is expected to be the new tool to improve the de novo sequencing and assembly.


June 1, 2021  |  

Long-read, single-molecule applications for protein engineering.

The long read lengths of PacBio’s SMRT Sequencing enable detection of linked mutations across multiple kilobases of sequence. This feature is particularly useful in the context of protein engineering, where large numbers of similar constructs are generated routinely to explore the effects of mutations on function and stability. We have developed a PCR-based barcoded sequencing method to generate high quality, full-length sequence data for batches of constructs generated in a common backbone. Individual barcodes are coupled to primers targeting a common region of the vector of interest. The amplified products are pooled into a single DNA library, and sequencing data are clustered by barcode to generate multi-molecule consensus sequences for each construct present in the pool. As a proof-of-concept dataset, we have generated a library of 384 randomly mutated variants of the Phi29 DNA polymerase, a 575 amino acid protein encoded by a 1.7 kb gene. These variants were amplified with a set of barcoded primers, and the resulting library was sequenced on a single SMRT Cell. The data produced sequences that were completely concordant with independent Sanger sequencing, for a 100% accurate reconstruction of the set of clones.


Talk with an expert

If you have a question, need to check the status of an order, or are interested in purchasing an instrument, we're here to help.