paper-with-me

홈 › Papers

Beyond Output Critique: Self-Correction via Task Distillation

2026-01-31 · Hossein A. Rahmani, Mengting Wan, Pei Zhou, Longqi Yang, Nick Craswell, Emine Yilmaz, Sujay Kumar Jauhar arxiv

Large language models (LLMs) have shown promising self-correction abilities, where iterative refinement improves the quality of generated responses. However, most existing approaches operate at the level of output critique, patching surface errors while often failing to correct deeper reasoning flaws. We propose SELF-THOUGHT, a framework that introduces an intermediate step of task abstraction before solution refinement. Given an input and an initial response, the model first distills the task into a structured template that captures key variables, constraints, and problem structure. This abstraction then guides solution instantiation, grounding subsequent responses in a clearer understanding of the task and reducing error propagation. Crucially, we show that these abstractions can be transferred across models: templates generated by larger models can serve as structured guides for smaller LLMs, which typically struggle with intrinsic self-correction. By reusing distilled task structures, smaller models achieve more reliable refinements without heavy fine-tuning or reliance on external verifiers. Experiments across diverse reasoning tasks demonstrate that SELF-THOUGHT improves accuracy, robustness, and generalization for both large and small models, offering a scalable path toward more reliable self-correcting language systems.

📄 PDF Abstract BibTeX arXiv:2602.00871

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

CriticBench: Benchmarking LLMs for Critique-Correct Reasoning

2024-02-22 · Zicheng Lin, Zhibin Gou, Tian Liang, Ruilin Luo 외

The ability of Large Language Models (LLMs) to critique and refine their reasoning is crucial for their application in evaluation, feedback provision, and self-improvement. This paper introduces CriticBench, a comprehens…

Benchmarking

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning

2024-12-03 · CVPR 2025 1 · Xueqing Wu, Yuheng Ding, Bingxuan Li, Pan Lu 외

The ability of large vision-language models (LVLMs) to critique and correct their reasoning is an essential building block towards their self-improvement. However, a systematic analysis of such capabilities in LVLMs is s…

BenchmarkingVisual Reasoning

Teaching Large Reasoning Models Effective Reflection

2026-01-19 · Hanbin Wang, Jingwei Song, Jinpeng Li, Qi Zhu 외 arxiv

Large Reasoning Models (LRMs) have recently shown impressive performance on complex reasoning tasks, often by engaging in self-reflective behaviors such as self-critique and backtracking. However, not all reflections are…

Reinforcement Learning

Enabling Scalable Oversight via Self-Evolving Critic

2025-01-10 · Zhengyang Tang, Ziniu Li, Zhenyang Xiao, Tian Ding 외

Despite their remarkable performance, the development of Large Language Models (LLMs) faces a critical challenge in scalable oversight: providing effective feedback for tasks where human evaluation is difficult or where …

Internal Reasoning vs. External Control: A Thermodynamic Analysis of Sycophancy in Large Language Models

2025-12-16 · Edward Y. Chang arxiv

Large Language Models exhibit sycophancy: prioritizing agreeableness over correctness. Current remedies evaluate reasoning outcomes: RLHF rewards correct answers, self-correction critiques outputs. All require ground tru…