paper-with-me

홈 › Papers

Bridge the Inference Gaps of Neural Processes via Expectation Maximization

2025-01-04 · Qi Wang, Marco Federici, Herke van Hoof

The neural process (NP) is a family of computationally efficient models for learning distributions over functions. However, it suffers from under-fitting and shows suboptimal performance in practice. Researchers have primarily focused on incorporating diverse structural inductive biases, \textit{e.g.} attention or convolution, in modeling. The topic of inference suboptimality and an analysis of the NP from the optimization objective perspective has hardly been studied in earlier work. To fix this issue, we propose a surrogate objective of the target log-likelihood of the meta dataset within the expectation maximization framework. The resulting model, referred to as the Self-normalized Importance weighted Neural Process (SI-NP), can learn a more accurate functional prior and has an improvement guarantee concerning the target log-likelihood. Experimental results show the competitive performance of SI-NP over other NPs objectives and illustrate that structural inductive biases, such as attention modules, can also augment our method to achieve SOTA performance. Our code is available at \url{https://github.com/hhq123gogogo/SI_NPs}.

📄 PDF Abstract BibTeX arXiv:2501.03264

Code (1)

hhq123gogogo/si_nps 공식 구현 pytorch

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Attention 설명 없음

Similar Papers 제목 키워드 기반

Prompt Tuning as User Inherent Profile Inference Machine

2024-08-13 · Yusheng Lu, Zhaocheng Du, Xiangyang Li, Xiangyu Zhao 외

Large Language Models (LLMs) have exhibited significant promise in recommender systems by empowering user profiles with their extensive world knowledge and superior reasoning capabilities. However, LLMs face challenges l…

QuantizationRecommendation SystemsWorld Knowledge

Homeostatic plasticity in Bayesian spiking networks as Expectation Maximization with posterior constraints

2012-12-01 · NeurIPS 2012 12 · Stefan Habenschuss, Johannes Bill, Bernhard Nessler

Recent spiking network models of Bayesian inference and unsupervised learning frequently assume either inputs to arrive in a special format or employ complex computations in neuronal activation functions and synaptic pla…

Bayesian InferenceVariational Inference

Entropic Matching for Expectation Propagation of Markov Jump Processes

2023-09-27 · Yannick Eich, Bastian Alt, Heinz Koeppl

We propose a novel, tractable latent state inference scheme for Markov jump processes, for which exact inference is often intractable. Our approach is based on an entropic matching framework that can be embedded into the…

Bayesian Inference

Learning Scripts as Hidden Markov Models

2018-09-11 · J. Walker Orr, Prasad Tadepalli, Janardhan Rao Doppa, Xiaoli Fern 외

Scripts have been proposed to model the stereotypical event sequences found in narratives. They can be applied to make a variety of inferences including filling gaps in the narratives and resolving ambiguous references. …

Clustering

Bayesian Integration of Information Using Top-Down Modulated WTA Networks

2023-08-29 · Otto van der Himst, Leila Bagheriye, Johan Kwisthout

Winner Take All (WTA) circuits a type of Spiking Neural Networks (SNN) have been suggested as facilitating the brain's ability to process information in a Bayesian manner. Research has shown that WTA circuits are capable…