paper-with-me

홈 › Papers

Importance Weighted Expectation-Maximization for Protein Sequence Design

2023-04-30 · Zhenqiao Song, Lei LI

Designing protein sequences with desired biological function is crucial in biology and chemistry. Recent machine learning methods use a surrogate sequence-function model to replace the expensive wet-lab validation. How can we efficiently generate diverse and novel protein sequences with high fitness? In this paper, we propose IsEM-Pro, an approach to generate protein sequences towards a given fitness criterion. At its core, IsEM-Pro is a latent generative model, augmented by combinatorial structure features from a separately learned Markov random fields (MRFs). We develop an Monte Carlo Expectation-Maximization method (MCEM) to learn the model. During inference, sampling from its latent space enhances diversity while its MRFs features guide the exploration in high fitness regions. Experiments on eight protein sequence design tasks show that our IsEM-Pro outperforms the previous best methods by at least 55% on average fitness score and generates more diverse and novel protein sequences.

📄 PDF Abstract BibTeX arXiv:2305.00386

Code (1)

JocelynSong/IsEM-Pro 공식 구현 pytorch

Tasks

Diversity

Similar Papers 제목 키워드 기반

Reweighted Expectation Maximization

2019-06-13 · Adji B. Dieng, John Paisley

Training deep generative models with maximum likelihood remains a challenge. The typical workaround is to use variational inference (VI) and maximize a lower bound to the log marginal likelihood of the data. Variational …

Bayesian InferenceDensity EstimationVariational Inference

Expectation-Maximization Attention Networks for Semantic Segmentation

2019-07-31 · ICCV 2019 10 · Xia Li, Zhisheng Zhong, Jianlong Wu, Yibo Yang 외

Self-attention mechanism has been widely used for various tasks. It is designed to compute the representation of each position by a weighted sum of the features at all positions. Thus, it can capture long-range relations…

Semantic Segmentation

Bridge the Inference Gaps of Neural Processes via Expectation Maximization

2025-01-04 · Qi Wang, Marco Federici, Herke van Hoof

The neural process (NP) is a family of computationally efficient models for learning distributions over functions. However, it suffers from under-fitting and shows suboptimal performance in practice. Researchers have pri…

Scalable Importance Sampling in High Dimensions with Low-Rank Mixture Proposals

2025-05-19 · Liam A. Kruse, Marc R. Schlichting, Mykel J. Kochenderfer

Importance sampling is a Monte Carlo technique for efficiently estimating the likelihood of rare events by biasing the sampling distribution towards the rare event of interest. By drawing weighted samples from a learned …

DWFL: Enhancing Federated Learning through Dynamic Weighted Averaging

2024-11-07 · Prakash Chourasia, Tamkanat E Ali, Sarwan Ali, Murray Pattersn

Federated Learning (FL) is a distributed learning technique that maintains data privacy by providing a decentralized training method for machine learning models using distributed big data. This promising Federated Learni…

Federated LearningPrivacy Preserving