Score: 0

Error-Correcting Codes for Labeled DNA Sequences

Published: November 3, 2025 | arXiv ID: 2511.01280v1

By: Dganit Hanania, Eitan Yaakobi

Potential Business Impact:

Fixes mistakes when reading DNA labels.

Business Areas:
Bioinformatics Biotechnology, Data and Analytics, Science and Engineering

Labeling of DNA molecules is a fundamental technique for DNA visualization and analysis. This process was mathematically modeled in [1], where the received sequence indicates the positions of the used labels. In this work, we develop error correcting codes for labeled DNA sequences, establishing bounds and constructing explicit systematic encoders for single substitution, insertion, and deletion errors. We focus on two cases: (1) using the complete set of length-two labels and (2) using the minimal set of length-two labels that ensures the recovery of DNA sequences from their labeling for 'almost' all DNA sequences.

Country of Origin
🇮🇱 Israel

Page Count
6 pages

Category
Computer Science:
Information Theory