paper-with-me

홈 › Papers

From Pixels to Nucleotides: End-to-End Token-Based Video Compression for DNA Storage

2026-04-15 · Cihan Ruan, Lebin Zhou, Bingqing Zhao, Rongduo Han, Qiming Yuan, Chenchen Zhu, Linyi Han, Liang Yang, Wei Wang, Wei Jiang, Nam Ling arxiv

DNA-based storage has emerged as a promising approach to the global data crisis, offering molecular-scale density and millennial-scale stability at low maintenance cost. Over the past decade, substantial progress has been made in storing text, images, and files in DNA -- yet video remains an open challenge. The difficulty is not merely technical: effective video DNA storage requires co-designing compression and molecular encoding from the ground up, a challenge that sits at the intersection of two fields that have largely evolved independently. In this work, we present HELIX, the first end-to-end neural network jointly optimizing video compression and DNA encoding -- prior approaches treat the two stages independently, leaving biochemical constraints and compression objectives fundamentally misaligned. Our key insight: token-based representations naturally align with DNA's quaternary alphabet -- discrete semantic units map directly to ATCG bases. We introduce TK-SCONE (Token-Kronecker Structured Constraint-Optimized Neural Encoding), which achieves 1.91 bits per nucleotide through Kronecker-structured mixing that breaks spatial correlations and FSM-based mapping that guarantees biochemical constraints. Unlike two-stage approaches, HELIX learns token distributions simultaneously optimized for visual quality, prediction under masking, and DNA synthesis efficiency. This work demonstrates for the first time that learned compression and molecular storage converge naturally at token representations -- suggesting a new paradigm where neural video codecs are designed for biological substrates from the ground up.

📄 PDF Abstract BibTeX arXiv:2604.13667

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Compression as Adaptation: Implicit Visual Representation with Diffusion Foundation Models

2026-03-08 · Zongyu Guo, Jiajun He, Zhaoyang Jia, Xiaoyi Zhang 외 arxiv

Modern visual generative models acquire rich visual knowledge through large-scale training, yet existing visual representations (such as pixels, latents, or tokens) remain external to the model and cannot directly exploi…

Optimal Video Compression using Pixel Shift Tracking

2024-06-28 · Hitesh Saai Mananchery Panneerselvam, Smit Anand

The Video comprises approximately ~85\% of all internet traffic, but video encoding/compression is being historically done with hard coded rules, which has worked well but only to a certain limit. We have seen a surge in…

Point TrackingVideo Compression

Survey of Information Encoding Techniques for DNA

2019-06-24 · Thomas Heinis, Roman Sokolovskii, Jamie J. Alnasir

The yearly global production of data is growing exponentially, outpacing the capacity of existing storage media, such as tape and disk, and surpassing our ability to store it. DNA storage - the representation of arbitrar…

Information RetrievalRetrievalSurveyTriplet

SeiT: Storage-Efficient Vision Training with Tokens Using 1% of Pixel Storage

2023-03-20 · ICCV 2023 1 · Song Park, Sanghyuk Chun, Byeongho Heo, Wonjae Kim 외

We need billion-scale images to achieve more generalizable and ground-breaking vision models, as well as massive dataset storage to ship the images (e.g., the LAION-4B dataset needs 240TB storage space). However, it has …

Continual Learning

CARP: Compression through Adaptive Recursive Partitioning for Multi-dimensional Images

2019-12-11 · CVPR 2020 6 · Rongjie Liu, Meng Li, Li Ma

Fast and effective image compression for multi-dimensional images has become increasingly important for efficient storage and transfer of massive amounts of high-resolution images and videos. Desirable properties in comp…

Image Compression