Skip to main navigation Skip to search Skip to main content

Making It to First: The Random Access Problem in DNA Storage

  • Avital Boruchovsky
  • , Ohad Elishco
  • , Ryan Gabrys
  • , Anina Gruica
  • , Itzhak Tamo
  • , Eitan Yaakobi

Research output: Contribution to journalArticlepeer-review

1 Scopus citations

Abstract

In this paper, we study the Random Access Problem in DNA storage, which addresses the challenge of retrieving a specific information strand from a DNA-based storage system. In this framework, the data is represented by k information strands which represent the data and are encoded into n strands using a linear code. Then, each sequencing read returns one encoded strand which is chosen uniformly at random. The goal under this paradigm is to design codes that minimize the expected number of reads required to recover an arbitrary information strand. We fully solve the case when k = 2, showing that the best possible code attains a random access expectation of 1 + √22+1 ≈ 0.914 · 2 for q large enough. Moreover, we extend a previous construction, originally developed for k = 3, to arbitrary values of k. Our construction uses Bk−1 sequences over Zq−1, that always exist over large finite fields. We show that for every k ≥ 4, this generalized construction outperforms all previous constructions in terms of reducing the random access expectation.

Original languageEnglish
Pages (from-to)5623-5638
Number of pages16
JournalIEEE Transactions on Information Theory
Volume72
Issue number8
DOIs
StatePublished - 1 Aug 2026

Keywords

  • Coding theory
  • DNA data storage
  • coverage depth
  • error-correcting codes
  • random access

ASJC Scopus subject areas

  • Information Systems
  • Computer Science Applications
  • Library and Information Sciences

Fingerprint

Dive into the research topics of 'Making It to First: The Random Access Problem in DNA Storage'. Together they form a unique fingerprint.

Cite this