paper-with-me

Papers

Escaping the Variance Trap: Jacobian-Free Dynamics for Root-Finding Bilevel Optimization

2026-06-21 · Zhiyu Li, Xi Xuan, Davide Carbone arxiv

Many central machine learning tasks, from entropy tuning in reinforcement learning to equilibrating generative adversarial networks, are fundamentally stochastic root-finding problems rather than loss minimization. Yet, they are frequently forced into a minimization framework via squared residuals, introducing a critical flaw we identify as the Variance Trap. Standard bilevel minimization algorithms require estimating hypergradients involving implicit Jacobians; in stochastic settings, these terms act as noise amplifiers, destabilizing convergence. We formalize Root-Finding Bilevel Optimization (RF-BO) as a distinct problem class that bypasses this pathology. We propose a Jacobian-free solution using Two-Time-Scale Stochastic Approximation (TTSA) that updates directly along the root error, structurally avoiding variance amplification. We provide the first non-asymptotic convergence guarantees for TTSA in this setting under Markovian noise. Extensive experiments demonstrate the decisive advantage of this paradigm: compared to squared-residual and implicit-gradient baselines, our framework achieves a 2.6\% top-1 accuracy gain in SimCLR, 17$\times$ faster convergence in non-linear ODE control where baselines fail, significantly improved entropy stability in reinforcement learning, and an 11.1\% quality improvement in generative modeling.

📄 PDF Abstract BibTeX arXiv:2606.22433

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement LearningBilevel Optimization

Similar Papers 제목 키워드 기반

TAMPC: A Controller for Escaping Traps in Novel Environments

2020-10-23 · Sheng Zhong, Zhenyuan Zhang, Nima Fazeli, Dmitry Berenson

We propose an approach to online model adaptation and control in the challenging case of hybrid and discontinuous dynamics where actions may lead to difficult-to-escape "trap" states, under a given controller. We first l…

Model Predictive Control

Asymptotic Freeness of Layerwise Jacobians Caused by Invariance of Multilayer Perceptron: The Haar Orthogonal Case

2021-03-24 · Benoit Collins, Tomohiro Hayase

Free Probability Theory (FPT) provides rich knowledge for handling mathematical difficulties caused by random matrices that appear in research related to deep neural networks (DNNs), such as the dynamical isometry, Fishe…

How to Stay Curious while Avoiding Noisy TVs using Aleatoric Uncertainty Estimation

2021-02-08 · Augustine N. Mavor-Parker, Kimberly A. Young, Caswell Barry, Lewis D. Griffin

Exploration in environments with sparse rewards is difficult for artificial agents. Curiosity driven learning -- using feed-forward prediction errors as intrinsic rewards -- has achieved some success in these scenarios, …

The Anisotropic Noise in Stochastic Gradient Descent: Its Behavior of Escaping from Sharp Minima and Regularization Effects

2018-03-01 · ICLR 2019 5 · Zhanxing Zhu, Jingfeng Wu, Bing Yu, Lei Wu 외

Understanding the behavior of stochastic gradient descent (SGD) in the context of deep neural networks has raised lots of concerns recently. Along this line, we study a general form of gradient based optimization dynamic…

Fiscal Stimulus of Last Resort

2021-04-06 · Alessandro Piergallini

I examine global dynamics in a monetary model with overlapping generations of finite-horizon agents and a binding lower bound on nominal interest rates. Debt targeting rules exacerbate the possibility of self-fulfilling …