paper-with-me

Papers

NbBench: Benchmarking Language Models for Comprehensive Nanobody Tasks

2025-05-04 · Yiming Zhang, Koji Tsuda

Nanobodies, single-domain antibody fragments derived from camelid heavy-chain-only antibodies, exhibit unique advantages such as compact size, high stability, and strong binding affinity, making them valuable tools in therapeutics and diagnostics. While recent advances in pretrained protein and antibody language models (PPLMs and PALMs) have greatly enhanced biomolecular understanding, nanobody-specific modeling remains underexplored and lacks a unified benchmark. To address this gap, we introduce NbBench, the first comprehensive benchmark suite for nanobody representation learning. Spanning eight biologically meaningful tasks across nine curated datasets, NbBench encompasses structure annotation, binding prediction, and developability assessment. We systematically evaluate eleven representative models--including general-purpose protein LMs, antibody-specific LMs, and nanobody-specific LMs--in a frozen setting. Our analysis reveals that antibody language models excel in antigen-related tasks, while performance on regression tasks such as thermostability and affinity remains challenging across all models. Notably, no single model consistently outperforms others across all tasks. By standardizing datasets, task definitions, and evaluation protocols, NbBench offers a reproducible foundation for assessing and advancing nanobody modeling.

📄 PDF Abstract BibTeX arXiv:2505.02022

Code (1)

zhymlumine/nbbench 공식 구현 pytorch

Tasks

BenchmarkingRepresentation Learning

Similar Papers 제목 키워드 기반

Sequence-Based Nanobody-Antigen Binding Prediction

2023-07-15 · Usama Sardar, Sarwan Ali, Muhammad Sohaib Ayub, Muhammad Shoaib 외

Nanobodies (Nb) are monomeric heavy-chain fragments derived from heavy-chain only antibodies naturally found in Camelids and Sharks. Their considerably small size (~3-4 nm; 13 kDa) and favorable biophysical properties ma…

Prediction

Nanobody interaction unveils structure, dynamics and proteotoxicity of the Finnish-type amyloidogenic gelsolin variant

2019-03-18

AGel amyloidosis, formerly known as familial amyloidosis of the Finnish-type, is caused by pathological aggregation of proteolytic fragments of plasma gelsolin. So far, four mutations in the gelsolin gene have been repor…

A novel framework to quantify uncertainty in peptide-tandem mass spectrum matches with application to nanobody peptide identification

2021-10-15 · Chris McKennan, Zhe Sang, Yi Shi

Nanobodies are small antibody fragments derived from camelids that selectively bind to antigens. These proteins have marked physicochemical properties that support advanced therapeutics, including treatments for SARS-CoV…

Model Selection

This is the way: designing and compiling LEPISZCZE, a comprehensive NLP benchmark for Polish

2022-11-23 · Łukasz Augustyniak, Kamil Tagowski, Albert Sawczyn, Denis Janiak 외

The availability of compute and data to train larger and larger language models increases the demand for robust methods of benchmarking the true progress of LM training. Recent years witnessed significant progress in sta…

Benchmarking

MAMMAL -- Molecular Aligned Multi-Modal Architecture and Language

2024-10-28 · Yoel Shoshan, Moshiko Raboh, Michal Ozery-Flato, Vadim Ratner 외

Large language models applied to vast biological datasets have the potential to transform biology by uncovering disease mechanisms and accelerating drug development. However, current models are often siloed, trained sepa…

Drug DiscoveryProperty Prediction