paper-with-me

Papers

A Probabilistic Perspective on Model Collapse

2025-05-20 · SHIRONG XU, Hengzhi He, Guang Cheng

In recent years, model collapse has become a critical issue in language model training, making it essential to understand the underlying mechanisms driving this phenomenon. In this paper, we investigate recursive parametric model training from a probabilistic perspective, aiming to characterize the conditions under which model collapse occurs and, crucially, how it can be mitigated. We conceptualize the recursive training process as a random walk of the model estimate, highlighting how the sample size influences the step size and how the estimation procedure determines the direction and potential bias of the random walk. Under mild conditions, we rigorously show that progressively increasing the sample size at each training step is necessary to prevent model collapse. In particular, when the estimation is unbiased, the required growth rate follows a superlinear pattern. This rate needs to be accelerated even further in the presence of substantial estimation bias. Building on this probabilistic framework, we also investigate the probability that recursive training on synthetic data yields models that outperform those trained solely on real data. Moreover, we extend these results to general parametric model family in an asymptotic regime. Finally, we validate our theoretical results through extensive simulations and a real-world dataset.

📄 PDF Abstract BibTeX arXiv:2505.13947

Code (0)

등록된 구현이 없습니다.

Tasks

model

Similar Papers 제목 키워드 기반

Understanding team collapse via probabilistic graphical models

2024-02-14 · Iasonas Nikolaou, Konstantinos Pelechrinis, Evimaria Terzi

In this work, we develop a graphical model to capture team dynamics. We analyze the model and show how to learn its parameters from data. Using our model we study the phenomenon of team collapse from a computational pers…

Don't Blame the ELBO! A Linear VAE Perspective on Posterior Collapse

2019-11-06 · NeurIPS 2019 12 · James Lucas, George Tucker, Roger Grosse, Mohammad Norouzi

Posterior collapse in Variational Autoencoders (VAEs) arises when the variational posterior distribution closely matches the prior for a subset of latent variables. This paper presents a simple and intuitive explanation …

Variational Inference

Preventing Posterior Collapse with Levenshtein Variational Autoencoder

2020-04-30 · Serhii Havrylov, Ivan Titov

Variational autoencoders (VAEs) are a standard framework for inducing latent variable models that have been shown effective in learning text representations as well as in text generation. The key challenge with using VAE…

SentenceText Generation

Action Gaps and Advantages in Continuous-Time Distributional Reinforcement Learning

2024-10-14 · Harley Wiltzer, Marc G. Bellemare, David Meger, Patrick Shafto 외

When decisions are made at high frequency, traditional reinforcement learning (RL) methods struggle to accurately estimate action values. In turn, their performance is inconsistent and often poor. Whether the performance…

Distributional Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Intelligent Trajectory Planning in UAV-mounted Wireless Networks: A Quantum-Inspired Reinforcement Learning Perspective

2020-07-27 · Yuanjian Li, A. Hamid Aghvami, Daoyi Dong

In this paper, we consider a wireless uplink transmission scenario in which an unmanned aerial vehicle (UAV) serves as an aerial base station collecting data from ground users. To optimize the expected sum uplink transmi…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)Trajectory Planning