paper-with-me

Papers

Accelerating Optimization via Differentiable Stopping Time

2025-05-28 · Zhonglin Xie, Yiman Fong, Haoran Yuan, Zaiwen Wen

Optimization is an important module of modern machine learning applications. Tremendous efforts have been made to accelerate optimization algorithms. A common formulation is achieving a lower loss at a given time. This enables a differentiable framework with respect to the algorithm hyperparameters. In contrast, its dual, minimizing the time to reach a target loss, is believed to be non-differentiable, as the time is not differentiable. As a result, it usually serves as a conceptual framework or is optimized using zeroth-order methods. To address this limitation, we propose a differentiable stopping time and theoretically justify it based on differential equations. An efficient algorithm is designed to backpropagate through it. As a result, the proposed differentiable stopping time enables a new differentiable formulation for accelerating algorithms. We further discuss its applications, such as online hyperparameter tuning and learning to optimize. Our proposed methods show superior performance in comprehensive experiments across various problems, which confirms their effectiveness.

📄 PDF Abstract BibTeX arXiv:2505.22509

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

DARTS+: Improved Differentiable Architecture Search with Early Stopping

2019-09-13 · Hanwen Liang, Shifeng Zhang, Jiacheng Sun, Xingqiu He 외

Recently, there has been a growing interest in automating the process of neural architecture design, and the Differentiable Architecture Search (DARTS) method makes the process available within a few GPU days. However, t…

GPU

Accelerating Neural Architecture Search using Performance Prediction

2017-05-30 · ICLR 2018 1 · Bowen Baker, Otkrist Gupta, Ramesh Raskar, Nikhil Naik

Methods for neural network hyperparameter optimization and meta-modeling are computationally expensive due to the need to train a large number of model configurations. In this paper, we show that standard frequentist reg…

Hyperparameter OptimizationLanguage ModelingLanguage ModellingNeural Architecture Search+4

Accelerating Visual-Policy Learning through Parallel Differentiable Simulation

2025-05-15 · Haoxiang You, Yilang Liu, Ian Abraham

In this work, we propose a computationally efficient algorithm for visual policy learning that leverages differentiable simulation and first-order analytical policy gradients. Our approach decouple the rendering process …

GPU

Differentiable Power-Flow Optimization

2026-03-30 · Muhammed Öz, Jasmin Hörter, Kaleb Phipps, Charlotte Debus 외 arxiv

With the rise of renewable energy sources and their high variability in generation, the management of power grids becomes increasingly complex and computationally demanding. Conventional AC-power-flow simulations, which …

Accelerating Electronic Stopping Power Predictions by 10 Million Times with a Combination of Time-Dependent Density Functional Theory and Machine Learning

2023-11-01 · Logan Ward, Ben Blaiszik, Cheng-Wei Lee, Troy Martin 외

Knowing the rate at which particle radiation releases energy in a material, the stopping power, is key to designing nuclear reactors, medical treatments, semiconductor and quantum materials, and many other technologies. …