paper-with-me

홈 › Papers

Super Level Sets and Exponential Decay: A Synergistic Approach to Stable Neural Network Training

2024-09-25 · Jatin Chaudhary, Dipak Nidhi, Jukka Heikkonen, Haari Merisaari, Rajiv Kanth

The objective of this paper is to enhance the optimization process for neural networks by developing a dynamic learning rate algorithm that effectively integrates exponential decay and advanced anti-overfitting strategies. Our primary contribution is the establishment of a theoretical framework where we demonstrate that the optimization landscape, under the influence of our algorithm, exhibits unique stability characteristics defined by Lyapunov stability principles. Specifically, we prove that the superlevel sets of the loss function, as influenced by our adaptive learning rate, are always connected, ensuring consistent training dynamics. Furthermore, we establish the "equiconnectedness" property of these superlevel sets, which maintains uniform stability across varying training conditions and epochs. This paper contributes to the theoretical understanding of dynamic learning rate mechanisms in neural networks and also pave the way for the development of more efficient and reliable neural optimization techniques. This study intends to formalize and validate the equiconnectedness of loss function as superlevel sets in the context of neural network training, opening newer avenues for future research in adaptive machine learning algorithms. We leverage previous theoretical discoveries to propose training mechanisms that can effectively handle complex and high-dimensional data landscapes, particularly in applications requiring high precision and reliability.

📄 PDF Abstract BibTeX arXiv:2409.16769

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Exponential Decay Exponential Decay is a learning rate schedule where we decay the learning rate with more iterations using an exponential function: $$ \text{lr} =…

Similar Papers 제목 키워드 기반

Neural network for multi-exponential sound energy decay analysis

2022-05-19 · Georg Götz, Ricardo Falcón Pérez, Sebastian J. Schlecht, Ville Pulkki

An established model for sound energy decay functions (EDFs) is the superposition of multiple exponentials and a noise term. This work proposes a neural-network-based approach for estimating the model parameters from EDF…

Investigation of event-based memory surfaces for high-speed tracking, unsupervised feature extraction and object recognition

2016-03-14 · Saeed Afshar, Gregory Cohen, Tara Julia Hamilton, Jonathan Tapson 외

In this paper we compare event-based decaying and time based-decaying memory surfaces for high-speed eventbased tracking, feature extraction, and object classification using an event-based camera. The high-speed recognit…

Object Recognition

Beyond Exponential Decay: Rethinking Error Accumulation in Large Language Models

2025-05-30 · Mikhail L. Arbuzov, Alexey A. Shvets, Sisong Beir

The prevailing assumption of an exponential decay in large language model (LLM) reliability with sequence length, predicated on independent per-token error probabilities, posits an inherent limitation for long autoregres…

Large Language Model

Black-box model classification under the discriminative factorization

2026-05-08 · Hayden Helm, Merrick Ohata, Carey Priebe arxiv

Access to modern generative systems is often restricted to querying an API (the ``black-box" setting) and many properties of the system are unknown to the user at inference time. While recent work has shown that low-dime…

Exponentially Consistent Kernel Two-Sample Tests

2018-02-23 · Shengyu Zhu, Biao Chen, Zhitang Chen

Given two sets of independent samples from unknown distributions $P$ and $Q$, a two-sample test decides whether to reject the null hypothesis that $P=Q$. Recent attention has focused on kernel two-sample tests as the tes…

Change DetectionVocal Bursts Valence Prediction