- DNA is a promising medium for archiving data because it will last in the right conditions for 10 000 years or longer
- The data stored in synthetic DNA could be retrieved with 100% accuracy by sequencing the sample and reconstructing the original files
Researchers at the EMBL-European Bioinformatics Institute (EMBL-EBI) have created a way to store data in the form of DNA – a material that lasts for tens of thousands of years. The new method, published today in the journal Nature, makes it possible to store at least 100 million hours of high-definition video in about a cup of DNA.
There is a lot of digital information in the world – about three zettabytes’ worth (that’s 3000 billion billion bytes) – and the constant influx of new digital content poses a real challenge for archivists. Hard disks are expensive and require a constant supply of electricity, while even the best ‘no-power’ archiving materials such as magnetic tape degrade within a decade. This is a growing problem in the life sciences, where massive volumes of data – including DNA sequences – make up the fabric of the scientific record.
"We already know that DNA is a robust way to store information because we can extract it from bones of woolly mammoths, which date back tens of thousands of years, and make sense of it,” explains Nick Goldman of EMBL-EBI. “It’s also incredibly small, dense and does not need any power for storage, so shipping and keeping it is easy.”
Reading DNA is fairly straightforward, but writing it has until now been a major hurdle to making DNA storage a reality. There are two challenges: first, using current methods it is only possible to manufacture DNA in short strings. Secondly, both writing and reading DNA are prone to errors, particularly when the same DNA letter is repeated. Nick Goldman and co-author Ewan Birney, Associate Director of EMBL-EBI, set out to create a code that overcomes both problems.
“We knew we needed to make a code using only short strings of DNA, and to do it in such a way that creating a run of the same letter would be impossible. So we figured, let’s break up the code into lots of overlapping fragments going in both directions, with indexing information showing where each fragment belongs in the overall code, and make a coding scheme that doesn't allow repeats. That way, you would have to have the same error on four different fragments for it to fail – and that would be very rare," says Ewan Birney.
The new method requires synthesising DNA from the encoded information: enter Agilent Technologies, Inc, a California-based company that volunteered its services. Ewan Birney and Nick Goldman sent them encoded versions of: an .mp3 of Martin Luther King’s speech, “I Have a Dream”; a .jpg photo of EMBL-EBI; a .pdf of Watson and Crick’s seminal paper, “Molecular structure of nucleic acids”; a .txt file of all of Shakespeare's sonnets; and a file that describes the encoding.
“We downloaded the files from the Web and used them to synthesise hundreds of thousands of pieces of DNA – the result looks like a tiny piece of dust,” explains Emily Leproust of Agilent. Agilent mailed the sample to EMBL-EBI, where the researchers were able to sequence the DNA and decode the files without errors.
“We’ve created a code that's error tolerant using a molecular form we know will last in the right conditions for 10 000 years, or possibly longer,” says Nick Goldman. “As long as someone knows what the code is, you will be able to read it back if you have a machine that can read DNA.”
Although there are many practical aspects to solve, the inherent density and longevity of DNA makes it an attractive storage medium. The next step for the researchers is to perfect the coding scheme and explore practical aspects, paving the way for a commercially viable DNA storage model.Policy regarding use
Small parts make the difference
12.01.2016 | Fraunhofer-Institut für Organische Elektronik, Elektronenstrahl- und Plasmatechnik FEP
Nanopores could take the salt out of seawater
12.11.2015 | University of Illinois at Urbana-Champaign
Automobiles increase the mobility of their users. However, their maneuverability is pushed to the limit by cramped inner city conditions. Those who need to...
Advance in biomedical imaging: The University of Würzburg's Biocenter has enhanced fluorescence microscopy to label and visualise up to nine different cell structures simultaneously.
Fluorescence microscopy allows researchers to visualise biomolecules in cells. They label the molecules using fluorescent probes, excite them with light and...
NASA's follow-on to the successful ICESat mission will employ a never-before-flown technique for determining the topography of ice sheets and the thickness of sea ice, but that won't be the only first for this mission.
Slated for launch in 2018, NASA's Ice, Cloud and land Elevation Satellite-2 (ICESat-2) also will carry a 3-D printed part made of polyetherketoneketone (PEKK),...
In the last decades, sea level has been rising continuously – about 3.3 mm per year. For reef islands such as the Maldives or the Marshall Islands a sinister picture is being painted evoking the demise of the island states and their cultures. Are the effects of sea-level rise already noticeable on reef islands? Scientists from the ZMT have now answered this question for the Takuu Atoll, a group of Pacific islands, located northeast of Papua New Guinea.
In the last decades, sea level has been rising continuously – about 3.3 mm per year. For reef islands such as the Maldives or the Marshall Islands a sinister...
The ‘Internet of Things’ is growing rapidly. Mobile phones, washing machines and the milk bottle in the fridge: the idea is that minicomputers connected to these will be able to process information, receive and send data. This requires electrical power. Transistors that are capable of switching information with a single electron use far less power than field effect transistors that are commonly used in computers. However, these innovative electronic switches do not yet work at room temperature. Scientists working on the new EU research project ‘Ions4Set’ intend to change this. The program will be launched on February 1. It is coordinated by the Helmholtz-Zentrum Dresden-Rossendorf (HZDR).
“Billions of tiny computers will in future communicate with each other via the Internet or locally. Yet power consumption currently remains a great obstacle”,...
02.02.2016 | Event News
26.01.2016 | Event News
26.01.2016 | Event News
05.02.2016 | Life Sciences
05.02.2016 | Materials Sciences
05.02.2016 | Physics and Astronomy