paper-with-me

홈 › Papers

Noise to the Rescue: Escaping Local Minima in Neurosymbolic Local Search

2025-03-03 · Alessandro Daniele, Emile van Krieken

Deep learning has achieved remarkable success across various domains, largely thanks to the efficiency of backpropagation (BP). However, BP's reliance on differentiability poses challenges in neurosymbolic learning, where discrete computation is combined with neural models. We show that applying BP to Godel logic, which represents conjunction and disjunction as min and max, is equivalent to a local search algorithm for SAT solving, enabling the optimisation of discrete Boolean formulas without sacrificing differentiability. However, deterministic local search algorithms get stuck in local optima. Therefore, we propose the Godel Trick, which adds noise to the model's logits to escape local optima. We evaluate the Godel Trick on SATLIB, and demonstrate its ability to solve a broad range of SAT problems. Additionally, we apply it to neurosymbolic models and achieve state-of-the-art performance on Visual Sudoku, all while avoiding expensive probabilistic reasoning. These results highlight the Godel Trick's potential as a robust, scalable approach for integrating symbolic reasoning with neural architectures.

📄 PDF Abstract BibTeX arXiv:2503.01817

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Towards Theoretically Understanding Why SGD Generalizes Better Than ADAM in Deep Learning

2020-10-12 · NeurIPS 2020 12 · Pan Zhou, Jiashi Feng, Chao Ma, Caiming Xiong 외

It is not clear yet why ADAM-alike adaptive gradient algorithms suffer from worse generalization performance than SGD despite their faster training speed. This work aims to provide understandings on this generalization g…

The Anisotropic Noise in Stochastic Gradient Descent: Its Behavior of Escaping from Sharp Minima and Regularization Effects

2018-03-01 · ICLR 2019 5 · Zhanxing Zhu, Jingfeng Wu, Bing Yu, Lei Wu 외

Understanding the behavior of stochastic gradient descent (SGD) in the context of deep neural networks has raised lots of concerns recently. Along this line, we study a general form of gradient based optimization dynamic…

The Anisotropic Noise in Stochastic Gradient Descent: Its Behavior of Escaping from Minima and Regularization Effects

2019-05-01 · ICLR 2019 5 · Zhanxing Zhu, Jingfeng Wu, Bing Yu, Lei Wu 외

Understanding the behavior of stochastic gradient descent (SGD) in the context of deep neural networks has raised lots of concerns recently. Along this line, we theoretically study a general form of gradient based optim…

Dynamic of Stochastic Gradient Descent with State-Dependent Noise

2020-06-24 · Qi Meng, Shiqi Gong, Wei Chen, Zhi-Ming Ma 외

Stochastic gradient descent (SGD) and its variants are mainstream methods to train deep neural networks. Since neural networks are non-convex, more and more works study the dynamic behavior of SGD and the impact to its g…

Explore the Loss space with Hill-ADAM

2025-10-04 · Meenakshi Manikandan, Leilani Gilpin arxiv

This paper introduces Hill-ADAM. Hill-ADAM is an optimizer with its focus towards escaping local minima in prescribed loss landscapes to find the global minimum. Hill-ADAM escapes minima by deterministically exploring th…