Home
Scholarly Works
Reconstruction Codes for DNA Sequences With...
Journal article

Reconstruction Codes for DNA Sequences With Uniform Tandem-Duplication Errors

Abstract

DNA as a data storage medium has several advantages, including far greater data density compared to electronic media. We propose that schemes for data storage in the DNA of living organisms may benefit from studying the reconstruction problem, which is applicable whenever multiple reads of noisy data are available. This strategy is uniquely suited to the medium, which inherently replicates stored data in multiple distinct ways, caused by mutations. We consider noise introduced solely by uniform tandem-duplication, and utilize the relation to constant-weight integer codes in the Manhattan metric. By bounding the intersection of the cross-polytope with hyperplanes, we prove the existence of reconstruction codes with full rate, as well as suggest a construction for a family of reconstruction codes.

Authors

Yehezkeally Y; Schwartz M

Journal

IEEE Transactions on Information Theory, Vol. 66, No. 5, pp. 2658–2668

Publisher

Institute of Electrical and Electronics Engineers (IEEE)

Publication Date

May 1, 2020

DOI

10.1109/tit.2019.2940256

ISSN

0018-9448

Contact the Experts team