activity
20242026
collaborators

6 papers

cs.IT2026

Marker-Delimited Codes for Short-Blocklength, High-Rate Coding over Multi-Read Edit Channels

Sinan Ates Yercan, Marc Antonini, Serge Kas Hanna

The read process of DNA-based data storage systems generates multiple noisy copies of the stored DNA sequences, affected by edit errors consisting of substitutions, deletions, and…

cs.IT2026

The Synthesis-Sequencing Channel for DNA-based Data Storage

Keshav Goyal, Samuel Pearson, João Ribeiro +1

We introduce and study the synthesis-sequencing channel, a two-stage model for DNA-based data storage that jointly captures synthesis and sequencing effects. The synthesis-sequenci…

cs.IT2026

DNA-MGC+: A versatile codec for reliable and resource-efficient data storage on synthetic DNA

Ramy Khabbaz, Jérémy Mateos, Marc Antonini +1

The biochemical processes underlying DNA data storage, including synthesis, amplification, and sequencing, are inherently noisy. Consequently, base-level insertion, deletion, and s…

cs.IT2025

On the Reliability of Information Retrieval From MDS Coded Data in DNA Storage

Serge Kas Hanna

This work presents a theoretical analysis of the probability of successfully retrieving data encoded with MDS codes (e.g., Reed-Solomon codes) in DNA storage systems. We study this…

cs.LG2024

Approximate Gradient Coding for Privacy-Flexible Federated Learning with Non-IID Data

Okko Makkonen, Sampo Niemelä, Camilla Hollanti +1

This work focuses on the challenges of non-IID data and stragglers/dropouts in federated learning. We introduce and explore a privacy-flexible paradigm that models parts of the cli…

cs.IT2024

GC+ Code: A Systematic Short Blocklength Code for Correcting Random Edit Errors in DNA Storage

Serge Kas Hanna

Storing digital data in synthetic DNA faces challenges in ensuring data reliability in the presence of edit errors--deletions, insertions, and substitutions--that occur randomly du…