paper-with-me

Papers

LightNobel: Improving Sequence Length Limitation in Protein Structure Prediction Model via Adaptive Activation Quantization

2025-05-09 · Seunghee Han, Soongyu Choi, Joo-Young Kim

Recent advances in Protein Structure Prediction Models (PPMs), such as AlphaFold2 and ESMFold, have revolutionized computational biology by achieving unprecedented accuracy in predicting three-dimensional protein folding structures. However, these models face significant scalability challenges, particularly when processing proteins with long amino acid sequences (e.g., sequence length > 1,000). The primary bottleneck that arises from the exponential growth in activation sizes is driven by the unique data structure in PPM, which introduces an additional dimension that leads to substantial memory and computational demands. These limitations have hindered the effective scaling of PPM for real-world applications, such as analyzing large proteins or complex multimers with critical biological and pharmaceutical relevance. In this paper, we present LightNobel, the first hardware-software co-designed accelerator developed to overcome scalability limitations on the sequence length in PPM. At the software level, we propose Token-wise Adaptive Activation Quantization (AAQ), which leverages unique token-wise characteristics, such as distogram patterns in PPM activations, to enable fine-grained quantization techniques without compromising accuracy. At the hardware level, LightNobel integrates the multi-precision reconfigurable matrix processing unit (RMPU) and versatile vector processing unit (VVPU) to enable the efficient execution of AAQ. Through these innovations, LightNobel achieves up to 8.44x, 8.41x speedup and 37.29x, 43.35x higher power efficiency over the latest NVIDIA A100 and H100 GPUs, respectively, while maintaining negligible accuracy loss. It also reduces the peak memory requirement up to 120.05x in PPM, enabling scalable processing for proteins with long sequences.

📄 PDF Abstract BibTeX arXiv:2505.05893

Code (0)

등록된 구현이 없습니다.

Tasks

Protein FoldingProtein Structure PredictionQuantization

Similar Papers 제목 키워드 기반

Variable-Length Generative Protein Design via Generalized Poisson Flow

2026-07-10 · Chaoran Cheng, Zhanghan Ni, Yanru Qu, Yuxin Chen 외 arxiv

The ability to generate variable-length proteins is crucial in protein design, where the optimal length is often unknown and tightly coupled to designability. Current diffusion- and flow-based generative models typically…

Protein Design

ProtTeX-CC: Activating In-Context Learning in Protein LLM via Two-Stage Instruction Compression

2025-08-17 · Chuanliu Fan, Zicheng Ma, Jun Gao, Nan Yu 외 arxiv

Recent advances in protein large language models, such as ProtTeX, represent both side-chain amino acids and backbone structure as discrete token sequences of residue length. While this design enables unified modeling of…

Protein Function Prediction

MAS2HP: A Multi Agent System to Predict Protein Structure in 2D HP model

2022-05-11 · Hossein Parineh, Nasser Mozayani

Protein Structure Prediction (PSP) is an unsolved problem in the field of computational biology. The problem of protein structure prediction is about predicting the native conformation of a protein, while its sequence of…

Protein Structure Prediction

The divergence time of protein structures modelled by Markov matrices and its relation to the divergence of sequences

2023-08-11 · Sandun Rajapaksa, Lloyd Allison, Peter J. Stuckey, Maria Garcia de la Banda 외

A complete time-parameterized statistical model quantifying the divergent evolution of protein structures in terms of the patterns of conservation of their secondary structures is inferred from a large collection of prot…

Evolution and Function of SMC Proteins

2022-10-29 · J. C. Phillips

Structural Maintenance of Chromosomes, SMCs, proteins have long rod like structures immersed in water. Here we use our hydroanalytic methods based on amino acid sequences to discuss their dynamics at multiple length scal…