Cold Spring Harbor, NY – The next "next-gen" technology in genome sequencing has gotten a major boost.
A quantitative biologist at Cold Spring Harbor Laboratory (CSHL) and collaborators today published results of experiments that demonstrate the power of so-called single-molecule sequencing, which was recently introduced but whose use has so far been limited by technical issues.
The team, led by CSHL Assistant Professor Michael Schatz and Adam Phillippy and Sergey Koren of the National Biodefense Analysis and Countermeasures Center and the University of Maryland (UMD), has developed a software package that corrects a serious problem inherent in the new sequencing technology: the fact that every fifth or sixth DNA "letter" it generates is incorrect. The high error rate is the flip side of the new method's chief virtue: it generates much longer genome "reads" than other technologies currently used, up to 100 times longer, and thus can provide a much more complete picture of genome structure than can be obtained with current, "2nd-gen" sequencing technology.
Using mathematical algorithms, Schatz and the team have preserved the great advantage of the "3rd-gen" method while all but eliminating its chief flaw. They have reduced the error rate from about 15% or greater to less than one-tenth of one percent. This mathematical "fix" – which has been published in open-source code to the World Wide Web – greatly increases the practical utility of 3rd-gen sequencing for the entire biomedical research community.
The team demonstrates the breadth of potential applications of single-molecule sequencing by applying their fix to sequencing tasks ranging from the tiny bacteriophage virus at one end of the difficulty scale to the large and vastly more complex genome of the parrot, at the other. The parrot genome is more than a third the size of the human genome and is published online today with the team's paper in Nature Biotechnology. The parrot sequence is "far superior to that of any previously sequenced bird genome," Schatz says.
To understand why it is better is to appreciate the advantages of 3rd-gen sequencing. The main advantage has to do with the average length of each "read" (i.e., genome segments read by a sequencer). The individual sequences are assembled into "contigs" -- shorthand for contiguous sequences -- much the way pieces in a jigsaw puzzle are assembled. In currently used 2nd-gen technology, the contigs are very small, and are massively redundant. A "consensus" version of each segment, representing the results of many layered reads, tends to be extremely accurate. But the small size of puzzle pieces prevents accurate assembly of certain genome portions, like those containing long repetitive sequences.
Obtaining superior versions of complete genomes was the objective that motivated Schatz and his collaborators, who also include HHMI Investigator Erich D. Jarvis of Duke University and CSHL Professor W. Richard McCombie, a sequencing pioneer, among others.
Combining the best of both generations
With single-molecule sequencing, the assembled contigs are much longer – affording a much better picture of relatively larger genome segments, including those occupied by lengthy repeats. This is what Schatz and his team wanted to preserve, while at the same time boosting the error-free rate. They did so by effectively taking the best of both 2nd- and 3rd-gen technologies.
"We call our approach 'hybrid error correction,'" Schatz explains.
The team's major insight was to take advantage of the long-read data offered by a 3rd-gen machine like that used in their experiments, a Pacific Biosciences RS sequencer, and mixing in highly accurate short reads obtained from a separate 2nd-gen sequencer. The two data types were run through an open-source genome assembly program called Celera Assembler to generate a clean final assembly that has proven 99.9% error-free and composed of contigs whose median size is at least double that obtainable with 2nd-gen "short-read" sequencers. Contig sizes are expected to increase appreciably in subsequent iterations of the hybrid approach as single molecule long-read sequencing improves.
High-quality genome assemblies are especially important for genome annotation and comparative genome analyses. Many microbial genome analyses depend on finished genomes, but their cost is prohibitive using older technologies. High-quality analysis of the genomes of higher organisms depends upon continuous sequences that capture long stretches of DNA that spell out genes. Discoveries in recent years of spontaneously occurring structural changes in genomes called copy number variations -- such as those made by CSHL Professor Mike Wigler and his team in their research on schizophrenia and autism – make clear the importance of being able to obtain clean and accurate pictures of the entire genomes of affected individuals.
With hybrid error correction, Schatz and his colleagues have "demonstrated that high error rates associated with long reads need not be a barrier to genome assembly," he summarizes. "High-error long reads can be efficiently assembled in combination with complementary short reads to produce assemblies not previously possible."
"Hybrid error correction and de novo assembly of single-molecule sequencing reads" appears online in Nature Biotechnology July 1, 2012. The authors are: Sergey Koren, Michael C. Schatz, Brian P. Walenz, Jeffrey Martin, Jason Howard, Ganeshkumar Ganapathy, Zhong Wang, David A. Rasko, W. Richard McCombie, Erich D. Davis and Adam M. Phillippy. the paper can be obtained online using doi: xxxxxxxxxxxxxxxx.
About Cold Spring Harbor Laboratory
Founded in 1890, Cold Spring Harbor Laboratory (CSHL) has shaped contemporary biomedical research and education with programs in cancer, neuroscience, plant biology and quantitative biology. CSHL is ranked number one in the world by Thomson Reuters for impact of its research in molecular biology and genetics. The Laboratory has been home to eight Nobel Prize winners. Today, CSHL's multidisciplinary scientific community is more than 360 scientists strong and its Meetings & Courses program hosts more than 12,500 scientists from around the world each year to its Long Island campus and its China center. Tens of thousands more benefit from the research, reviews, and ideas published in journals and books distributed internationally by CSHL Press. The Laboratory's education arm also includes a graduate school and programs for undergraduates as well as middle and high school students and teachers. CSHL is a private, not-for-profit institution on the north shore of Long Island. For more information, visit www.cshl.edu.
Peter Tarr | EurekAlert!
Selectively Reactivating Nerve Cells to Retrieve a Memory
01.06.2020 | Universität Heidelberg
CeMM study reveals how a master regulator of gene transcription operates
01.06.2020 | CeMM Forschungszentrum für Molekulare Medizin der Österreichischen Akademie der Wissenschaften
In living cells, enzymes drive biochemical metabolic processes enabling reactions to take place efficiently. It is this very ability which allows them to be used as catalysts in biotechnology, for example to create chemical products such as pharmaceutics. Researchers now identified an enzyme that, when illuminated with blue light, becomes catalytically active and initiates a reaction that was previously unknown in enzymatics. The study was published in "Nature Communications".
Enzymes: they are the central drivers for biochemical metabolic processes in every living cell, enabling reactions to take place efficiently. It is this very...
Early detection of tumors is extremely important in treating cancer. A new technique developed by researchers at the University of California, Davis offers a significant advance in using magnetic resonance imaging to pick out even very small tumors from normal tissue. The work is published May 25 in the journal Nature Nanotechnology.
researchers at the University of California, Davis offers a significant advance in using magnetic resonance imaging to pick out even very small tumors from...
Microelectronics as a key technology enables numerous innovations in the field of intelligent medical technology. The Fraunhofer Institute for Biomedical Engineering IBMT coordinates the BMBF cooperative project "I-call" realizing the first electronic system for ultrasound-based, safe and interference-resistant data transmission between implants in the human body.
When microelectronic systems are used for medical applications, they have to meet high requirements in terms of biocompatibility, reliability, energy...
Thomas Heine, Professor of Theoretical Chemistry at TU Dresden, together with his team, first predicted a topological 2D polymer in 2019. Only one year later, an international team led by Italian researchers was able to synthesize these materials and experimentally prove their topological properties. For the renowned journal Nature Materials, this was the occasion to invite Thomas Heine to a News and Views article, which was published this week. Under the title "Making 2D Topological Polymers a reality" Prof. Heine describes how his theory became a reality.
Ultrathin materials are extremely interesting as building blocks for next generation nano electronic devices, as it is much easier to make circuits and other...
Scientists took a leukocyte as the blueprint and developed a microrobot that has the size, shape and moving capabilities of a white blood cell. Simulating a blood vessel in a laboratory setting, they succeeded in magnetically navigating the ball-shaped microroller through this dynamic and dense environment. The drug-delivery vehicle withstood the simulated blood flow, pushing the developments in targeted drug delivery a step further: inside the body, there is no better access route to all tissues and organs than the circulatory system. A robot that could actually travel through this finely woven web would revolutionize the minimally-invasive treatment of illnesses.
A team of scientists from the Max Planck Institute for Intelligent Systems (MPI-IS) in Stuttgart invented a tiny microrobot that resembles a white blood cell...
19.05.2020 | Event News
07.04.2020 | Event News
06.04.2020 | Event News
29.05.2020 | Materials Sciences
29.05.2020 | Materials Sciences
29.05.2020 | Power and Electrical Engineering