paper-with-me

홈 › Papers

From Solving to Verifying: A Unified Objective for Robust Reasoning in LLMs

2025-11-19 · Xiaoxuan Wang, Bo Liu, Song Jiang, Jingzhou Liu, Jingyuan Qi, Xia Chen, Baosheng He arxiv

The reasoning capabilities of large language models (LLMs) have been significantly improved through reinforcement learning (RL). Nevertheless, LLMs still struggle to consistently verify their own reasoning traces. This raises the research question of how to enhance the self-verification ability of LLMs and whether such an ability can further improve reasoning performance. In this work, we propose GRPO-Verif, an algorithm that jointly optimizes solution generation and self-verification within a unified loss function, with an adjustable hyperparameter controlling the weight of the verification signal. Experimental results demonstrate that our method enhances self-verification capability while maintaining comparable performance in reasoning.

📄 PDF Abstract BibTeX arXiv:2511.15137

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

GPT-4 Doesn't Know It's Wrong: An Analysis of Iterative Prompting for Reasoning Problems

2023-10-19 · Kaya Stechly, Matthew Marquez, Subbarao Kambhampati

There has been considerable divergence of opinion on the reasoning abilities of Large Language Models (LLMs). While the initial optimism that reasoning might emerge automatically with scale has been tempered thanks to a …

Scheduling

Neuro-Symbolic Verification on Instruction Following of LLMs

2026-01-25 · Yiming Su, Kunzhao Xu, Yanjie Gao, Fan Yang 외 arxiv

A fundamental problem of applying Large Language Models (LLMs) to important applications is that LLMs do not always follow instructions, and violations are often hard to observe or check. In LLM-based agentic workflows, …

Instruction FollowingLogical Reasoning

Enhancing the Geometric Problem-Solving Ability of Multimodal LLMs via Symbolic-Neural Integration

2025-04-17 · YiCheng Pan, Zhenrong Zhang, Pengfei Hu, Jiefeng Ma 외

Recent advances in Multimodal Large Language Models (MLLMs) have achieved remarkable progress in general domains and demonstrated promise in multimodal mathematical reasoning. However, applying MLLMs to geometry problem …

Geometry Problem SolvingLarge Language ModelLogical ReasoningMathematical Reasoning

The Reasoning-Creativity Trade-off: Toward Creativity-Driven Problem Solving

2026-01-02 · Max Ruiz Luyten, Mihaela van der Schaar arxiv

State-of-the-art large language model (LLM) pipelines rely on bootstrapped reasoning loops: sampling diverse chains of thought and reinforcing the highest-scoring ones, mainly optimizing correctness. We analyze how this …

LeanGeo: Formalizing Competitional Geometry problems in Lean

2025-08-20 · Chendong Song, Zihan Wang, Frederick Pu, Haiming Wang 외 arxiv

Geometry problems are a crucial testbed for AI reasoning capabilities. Most existing geometry solving systems cannot express problems within a unified framework, thus are difficult to integrate with other mathematical fi…