paper-with-me

홈 › Papers

AbstRaL: Augmenting LLMs' Reasoning by Reinforcing Abstract Thinking

2025-06-09 · Silin Gao, Antoine Bosselut, Samy Bengio, Emmanuel Abbe

Recent studies have shown that large language models (LLMs), especially smaller ones, often lack robustness in their reasoning. I.e., they tend to experience performance drops when faced with distribution shifts, such as changes to numerical or nominal variables, or insertions of distracting clauses. A possible strategy to address this involves generating synthetic data to further "instantiate" reasoning problems on potential variations. In contrast, our approach focuses on "abstracting" reasoning problems. This not only helps counteract distribution shifts but also facilitates the connection to symbolic tools for deriving solutions. We find that this abstraction process is better acquired through reinforcement learning (RL) than just supervised fine-tuning, which often fails to produce faithful abstractions. Our method, AbstRaL -- which promotes abstract reasoning in LLMs using RL on granular abstraction data -- significantly mitigates performance degradation on recent GSM perturbation benchmarks.

📄 PDF Abstract BibTeX arXiv:2506.07751

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning (RL)

Similar Papers 제목 키워드 기반

ABSTRAL: Automatic Design of Multi-Agent Systems Through Iterative Refinement and Topology Optimization

2026-03-24 · Weijia Song, Jiashu Yue, Zhe Pang arxiv

How should multi-agent systems be designed, and can that design knowledge be captured in a form that is inspectable, revisable, and transferable? We introduce ABSTRAL, a framework that treats MAS architecture as an evolv…

Towards Effectively Leveraging Execution Traces for Program Repair with Code LLMs

2025-05-07 · Mirazul Haque, Petr Babkin, Farima Farmahinifarahani, Manuela Veloso

Large Language Models (LLMs) show promising performance on various programming tasks, including Automatic Program Repair (APR). However, most approaches to LLM-based APR are limited to the static analysis of the programs…

Program Repair

Meaningful Learning: Enhancing Abstract Reasoning in Large Language Models via Generic Fact Guidance

2024-03-14 · Kai Xiong, Xiao Ding, Ting Liu, Bing Qin 외

Large language models (LLMs) have developed impressive performance and strong explainability across various reasoning scenarios, marking a significant stride towards mimicking human-like intelligence. Despite this, when …

Memorization

AnomSeer: Reinforcing Multimodal LLMs to Reason for Time-Series Anomaly Detection

2026-02-09 · Junru Zhang, Lang Feng, Haoran Shi, Xu Guo 외 arxiv

Time-series anomaly detection (TSAD) with multimodal large language models (MLLMs) is an emerging area, yet a persistent challenge remains: MLLMs rely on coarse time-series heuristics but struggle with multi-dimensional,…

Anomaly ClassificationReinforcement LearningAnomaly Detection

SpaceR: Reinforcing MLLMs in Video Spatial Reasoning

2025-04-02 · Kun Ouyang, Yuanxin Liu, HaoNing Wu, Yi Liu 외

Video spatial reasoning, which involves inferring the underlying spatial structure from observed video frames, poses a significant challenge for existing Multimodal Large Language Models (MLLMs). This limitation stems pr…

MMESpatial ReasoningVideo MMEVideo Understanding