paper-with-me

홈 › Papers

Self-Improvement Imitation with Biologically Guided Search for Protein Design Under Oracle Budgets

2026-05-26 · Ashima Khanna, Dominik Grimm arxiv

Protein sequence optimization under tight oracle budgets requires methods that explore vast combinatorial spaces while making each evaluation informative. Existing reinforcement learning and off-policy generative approaches often degrade under surrogate noise, and position-agnostic mutation proposals risk disrupting functionally critical residues. We introduce SILO, a trajectory-level self-improvement imitation framework for oracle-budgeted protein design. SILO uses a hierarchical edit policy that decomposes each mutation into a position choice followed by a residue choice. In each active-learning round, the policy samples candidate trajectories via incremental stochastic beam search without replacement (SBS), and a UCB-based proxy ensemble, combined with an alanine-scan fitness score (AFS), selects candidates with functionally relevant edits for in silico oracle evaluation. The policy is then updated by next-action cross-entropy imitation on the round's best oracle-labeled trajectories, avoiding value-function estimation. Across eight reproduced protein fitness landscapes and five strong baselines from prior work, SILO achieves the highest maximum and top-100 mean fitness on 8 of 8 landscapes within our evaluations, often exhibiting faster early-stage improvement. In low-data and noisy-proxy stress tests on two landscapes per setting, SILO remains competitive or best when several baselines degrade. Ablations show that SBS with AFS account for much of the gains, with iterative imitation providing additional improvement. Code is available at: https://github.com/grimmlab/SILO.git

📄 PDF Abstract BibTeX arXiv:2605.26690

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement LearningProtein Design

Similar Papers 제목 키워드 기반

Graph Transformer-Based Pathway Embedding for Cancer Prognosis

2026-04-17 · Koushik Howlader, Md Tauhidul Islam, Wei Le arxiv

Accurate prediction of cancer progression remains a challenge due to the high heterogeneity of molecular omics data across patients. While biologically informed models have improved the interpretability of these predicti…

Meta-Representational Predictive Coding: Biomimetic Self-Supervised Learning

2025-03-22 · Alexander Ororbia, Karl Friston, Rajesh P. N. Rao

Self-supervised learning has become an increasingly important paradigm in the domain of machine intelligence. Furthermore, evidence for self-supervised adaptation, such as contrastive formulations, has emerged in recent …

FormSelf-Supervised Learning

Adaptive Multi-Scale Goodness Aggregation for Forward-Forward Learning

2026-05-11 · Salar Beigzad, Vansh Verma arxiv

We propose Adaptive Multi-Scale Goodness Aggregation (AMSGA), a novel extension of the Forward-Forward (FF) algorithm designed to improve stability, robustness, and generalization in local-learning neural networks. AMSGA…

GRNFormer: A Biologically-Guided Framework for Integrating Gene Regulatory Networks into RNA Foundation Models

2025-03-03 · Mufan Qiu, Xinyu Hu, Fengwei Zhan, Sukwon Yun 외

Foundation models for single-cell RNA sequencing (scRNA-seq) have shown promising capabilities in capturing gene expression patterns. However, current approaches face critical limitations: they ignore biological prior kn…

Drug Response PredictionGraph Neural Network

Structural Entropy Guided Agent for Detecting and Repairing Knowledge Deficiencies in LLMs

2025-05-12 · Yifan Wei, Xiaoyan Yu, Tengfei Pan, Angsheng Li 외

Large language models (LLMs) have achieved unprecedented performance by leveraging vast pretraining corpora, yet their performance remains suboptimal in knowledge-intensive domains such as medicine and scientific researc…

AI AgentKnowledge DistillationKnowledge GraphsReinforcement Learning+1