paper-with-me

Papers

Newton's Lantern: A Reinforcement Learning Framework for Finetuning AC Power Flow Warm Start Models

2026-05-11 · Shourya Bose, Helgi Hilmarsson, Dhruv Suri arxiv

Neural warm starts can sharply reduce the number of Newton-Raphson iterations required to solve the AC power flow problem, but existing supervised approaches generalize poorly on heavily loaded instances near voltage collapse. We prove a lower bound on the Newton-Raphson iteration count that depends on the direction of the warm start error rather than on its magnitude, and show as a corollary that the bound becomes vacuous as the smallest singular value of the power-flow Jacobian shrinks, identifying the failure mode of supervised regression near the saddle-node bifurcation. Motivated by this analysis, we introduce Newton's Lantern, a finetuning pipeline that combines group relative policy optimization with a learned reward model trained on perturbations of the base model's predictions, using the iteration count itself as the supervisory signal. Across IEEE 118-bus, GOC 500-bus, and GOC 2000-bus benchmarks, Newton's Lantern is the only method that converges on every test snapshot while attaining the smallest mean iteration count.

📄 PDF Abstract BibTeX arXiv:2605.11102

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

LANTERN-RD: Enabling Deep Learning for Mitigation of the Invasive Spotted Lanternfly

2022-05-12 · Srivatsa Kundurthy

The Spotted Lanternfly (SLF) is an invasive planthopper that threatens the local biodiversity and agricultural economy of regions such as the Northeastern United States and Japan. As researchers scramble to study the ins…

Pose Estimation

LanteRn: Latent Visual Structured Reasoning

2026-03-26 · André G. Viveiros, Nuno Gonçalves, Matthias Lindemann, André Martins arxiv

While language reasoning models excel in many tasks, visual reasoning remains challenging for current large multimodal models (LMMs). As a result, most LMMs default to verbalizing perceptual content into text, a strong l…

Reinforcement LearningMultimodal ReasoningVisual GroundingVisual Reasoning

Enhancing TCR-Peptide Interaction Prediction with Pretrained Language Models and Molecular Representations

2025-04-22 · Cong Qi, Hanzhang Fang, Siqi Jiang, Tianxing Hu 외

Understanding the binding specificity between T-cell receptors (TCRs) and peptide-major histocompatibility complexes (pMHCs) is central to immunotherapy and vaccine development. However, current predictive models struggl…

BenchmarkingFew-Shot LearningLanguage ModelingLanguage Modelling+2

LANTERN: LLM-Augmented Neurosymbolic Transfer with Experience-Gated Reasoning Networks

2026-05-06 · Mahyar Alinejad, Yue Wang, Amrit Singh Bedi, George Atia arxiv

Transfer learning in reinforcement learning (RL) seeks to accelerate learning in new tasks by leveraging knowledge from related sources. Existing neurosymbolic transfer methods, however, typically rely on manually specif…

Reinforcement LearningTransfer Learning

Push the Limit of Multi-modal Emotion Recognition by Prompting LLMs with Receptive-Field-Aware Attention Weighting

2024-11-26 · Liyun Zhang, Dian Ding, Yu Lu, Yi-Chao Chen 외

Understanding the emotions in a dialogue usually requires external knowledge to accurately understand the contents. As the LLMs become more and more powerful, we do not want to settle on the limited ability of the pre-tr…

Emotion Recognition