Showing posts with label Data Storage. Show all posts
Showing posts with label Data Storage. Show all posts

Jul 12, 2023

Capturing the immense potential of microscopic DNA for data storage

In a world first, a 'biological camera' bypasses the constraints of current DNA storage methods, harnessing living cells and their inherent biological mechanisms to encode and store data. This represents a significant breakthrough in encoding and storing images directly within DNA, creating a new model for information storage reminiscent of a digital camera.

Led by Principal Investigator Associate Professor Chueh Loo Poh from the College of Design and Engineering at the National University of Singapore, and the NUS Synthetic Biology for Clinical and Technological Innovation (SynCTI), the team's findings, which could potentially shake up the data-storage industry, were published in Nature Communications on 3 July 2023.

A new paradigm to address global data overload

As the world continues to generate data at an unprecedented rate, data has come to be seen as the 'currency' of the 21st century. Estimated to be 33 ZB in 2018, it has been forecasted that the Global Datasphere will reach 175 ZB by 2025. That has sparked a quest for a storage alternative that can transcend the confines of conventional data storage and address the environmental impact of resource-intensive data centres.

It is only recently that the idea of using DNA to store other types of information, such as images and videos, has garnered attention. This is due to DNA's exceptional storage capacity, stability, and long-standing relevance as a medium for information storage.

"We are facing an impending data overload. DNA, the key biomaterial of every living thing on Earth, stores genetic information that encodes for an array of proteins responsible for various life functions. To put it into perspective, a single gram of DNA can hold over 215,000 terabytes of data -- equivalent to storing 45 million DVDs combined," said Assoc Prof Poh.

"DNA is also easy to manipulate with current molecular biology tools, can be stored in various forms at room temperature, and is so durable it can last centuries," says Cheng Kai Lim, a graduate student working with Assoc Prof Poh.

Despite its immense potential, current research in DNA storage focuses on synthesising DNA strands outside the cells. This process is expensive and relies on complex instruments, which are also prone to errors.

To overcome this bottleneck, Assoc Prof Poh and his team turned to live cells, which contain an abundance of DNA that can act as a 'data bank', circumventing the need to synthesise the genetic material externally.

Through sheer ingenuity and clever engineering, the team developed 'BacCam' -- a novel system that merges various biological and digital techniques to emulate a digital camera's functions using biological components.

"Imagine the DNA within a cell as an undeveloped photographic film," explained Assoc Prof Poh. "Using optogenetics -- a technique that controls the activity of cells with light akin to the shutter mechanism of a camera, we managed to capture 'images' by imprinting light signals onto the DNA 'film'."

Next, using barcoding techniques akin to photo labelling, the researchers marked the captured images for unique identification. Machine-learning algorithms were employed to organise, sort, and reconstruct the stored images. These constitute the 'biological camera', mirroring a digital camera's data capture, storage, and retrieval processes.

The study showcased the camera's ability to capture and store multiple images simultaneously using different light colours. More crucially, compared to earlier methods of DNA data storage, the team's innovative system is easily reproducible and scalable.

"As we push the boundaries of DNA data storage, there is an increasing interest in bridging the interface between biological and digital systems," said Assoc Prof Poh.

Read more at Science Daily

May 4, 2023

The future of data storage lies in DNA microcapsules

Storing data in DNA sounds like science fiction, yet it lies in the near future. Professor Tom de Greef expects the first DNA data center to be up and running within five to ten years. Data won't be stored as zeros and ones in a hard drive but in the base pairs that make up DNA: AT and CG. Such a data center would take the form of a lab, many times smaller than the ones today. De Greef can already picture it all. In one part of the building, new files will be encoded via DNA synthesis. Another part will contain large fields of capsules, each capsule packed with a file. A robotic arm will remove a capsule, read its contents and place it back.

We're talking about synthetic DNA. In the lab, bases are stuck together in a certain order to form synthetically produced strands of DNA. Files and photos that are currently stored in data centers can then be stored in DNA. For now, the technique is suitable only for archival storage. This is because the reading of stored data is very expensive, so you want to consult the DNA files as little as possible.

Large, energy-guzzling data centers made obsolete

Data storage in DNA offers many advantages. A DNA file can be stored much more compactly, for instance, and the lifespan of the data is also many times longer. But perhaps most importantly, this new technology renders large, energy-guzzling data centers obsolete. And this is desperately needed, warns De Greef, "because in three years, we will generate so much data worldwide that we won't be able to store half of it."

Together with PhD student Bas Bögels, Microsoft and a group of university partners, De Greef has developed a new technique to make the innovation of data storage with synthetic DNA scalable. The results have been published today in the journal Nature Nanotechnology. De Greef works at the Department of Biomedical Engineering and the Institute for Complex Molecular Systems (ICMS) at TU Eindhoven and serves as a visiting professor at Radboud University.

Scalable

The idea of using strands of DNA for data storage emerged in the 1980s but was far too difficult and expensive at the time. It became technically possible three decades later, when DNA synthesis started to take off. George Church, a geneticist at Harvard Medical School, elaborated on the idea in 2011. Since then, synthesis and the reading of data have become exponentially cheaper, finally bringing the technology to the market.

In recent years, De Greef and his group have looked mainly into reading the stored data. For the time being, this is the biggest problem facing this new technique. The PCR method currently used for this, called 'random access', is highly error-prone. You can therefore only read one file at a time and, in addition, the data quality deteriorates too much each time you read a file. Not exactly scalable.

Here's how it works: PCR (Polymerase Chain Reaction) creates millions of copies of the piece of DNA that you need by adding a primer with the desired DNA code. Corona tests in the lab, for example, are based on this: even a minuscule amount of coronavirus material from your nose is detectable when copied so many times. But if you want to read multiple files simultaneously, you need multiple primer pairs doing their work at the same time. This creates many errors in the copying process.

Every capsule contains one file

This is where the capsules come into play. De Greef's group developed a microcapsule of proteins and a polymer and then anchored one file per capsule. De Greef: "These capsules have thermal properties that we can use to our advantage." Above 50 degrees Celsius, the capsules seal themselves, allowing the PCR process to take place separately in each capsule. Not much room for error then. De Greef calls this 'thermo-confined PCR'. In the lab, it has so far managed to read 25 files simultaneously without significant error.

If you then lower the temperature again, the copies detach from the capsule and the anchored original remains, meaning that the quality of your original file does not deteriorate. De Greef: "We currently stand at a loss of 0.3 percent after three reads, compared to 35 percent with the existing method."

Searchable with fluorescence

And that's not all. De Greef has also made the data library even easier to search. Each file is given a fluorescent label and each capsule its own color. A device can then recognize the colors and separate them from one another. This brings us back to the imaginary robotic arm at the beginning of this story, which will neatly select the desired file from the pool of capsules in the future.

Read more at Science Daily