paper-with-me

홈 › Papers

Adaptive Constraint Propagation: Scaling Structured Inference for Large Language Models via Meta-Reinforcement Learning

2025-12-31 · Ibne Farabi Shihab, Sanjeda Akter, Anuj Sharma arxiv

Large language models increasingly require structured inference, from JSON schema enforcement to multi-lingual parsing, where outputs must satisfy complex constraints. We introduce MetaJuLS, a meta-reinforcement learning approach that learns universal constraint propagation policies applicable across languages and tasks without task-specific retraining. By formulating structured inference as adaptive constraint propagation and training a Graph Attention Network with meta-learning, MetaJuLS achieves 1.5--2.0$\times$ speedups over GPU-optimized baselines while maintaining within 0.2\% accuracy of state-of-the-art parsers. On Universal Dependencies across 10 languages and LLM-constrained generation (LogicBench, GSM8K-Constrained), MetaJuLS demonstrates rapid cross-domain adaptation: a policy trained on English parsing adapts to new languages and tasks with 5--10 gradient steps (5--15 seconds) rather than requiring hours of task-specific training. Mechanistic analysis reveals the policy discovers human-like parsing strategies (easy-first) and novel non-intuitive heuristics. By reducing propagation steps in LLM deployments, MetaJuLS contributes to Green AI by directly reducing inference carbon footprint.

📄 PDF Abstract BibTeX arXiv:2601.00095

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement LearningDomain Adaptation

Similar Papers 제목 키워드 기반

Embedding Inference for Structured Multilabel Prediction

2015-12-01 · NeurIPS 2015 12 · Farzaneh Mirzazadeh, Siamak Ravanbakhsh, Nan Ding, Dale Schuurmans

A key bottleneck in structured output prediction is the need for inference during training and testing, usually requiring some form of dynamic programming. Rather than using approximate inference or tailoring a speciali…

Prediction

3D Optimization for AI Inference Scaling: Balancing Accuracy, Cost, and Latency

2025-10-21 · Minseok Jung, Abhas Ricky, Muhammad Rameez Chatni arxiv

AI inference scaling is often tuned through 1D heuristics (a fixed reasoning pass) or 2D bivariate trade-offs (e.g., accuracy vs. compute), which fail to consider cost and latency constraints. We introduce a 3D optimizat…

Inference with Aggregate Data: An Optimal Transport Approach

2020-03-31 · Rahul Singh, Isabel Haasler, Qinsheng Zhang, Johan Karlsson 외

We consider inference (filtering) problems over probabilistic graphical models with aggregate data generated by a large population of individuals. We propose a new efficient belief propagation type algorithm over tree-st…

Accelerating Scalable Graph Neural Network Inference with Node-Adaptive Propagation

2023-10-17 · Xinyi Gao, Wentao Zhang, Junliang Yu, Yingxia Shao 외

Graph neural networks (GNNs) have exhibited exceptional efficacy in a diverse array of applications. However, the sheer size of large-scale graphs presents a significant challenge to real-time inference with GNNs. Althou…

Graph Neural Network

The Unified Balance Theory of Second-Moment Exponential Scaling Optimizers in Visual Tasks

2024-05-28 · Gongyue Zhang, Honghai Liu

We have identified a potential method for unifying first-order optimizers through the use of variable Second-Moment Exponential Scaling(SMES). We begin with back propagation, addressing classic phenomena such as gradient…