ReflCtrl: Controlling LLM Reflection via Representation Engineering
Large language models (LLMs) with Chain-of-Thought (CoT) reasoning have achieved strong performance across diverse tasks, including mathematics, coding, and general reasoning. A distinctive ability of these reasoning models is self-reflection: the ability to review and revise previous reasoning steps. While self-reflection enhances reasoning performance, it also increases inference cost. In this work, we study self-reflection through the lens of representation engineering. We segment the model's reasoning into steps, identify the steps corresponding to reflection, and extract a reflection direction in the latent space that governs this behavior. Using this direction, we propose a stepwise steering method that can control reflection frequency. We call our framework ReflCtrl. Our experiments show that (1) in many cases reflections are redundant, especially in stronger models (in our experiments, we can save up to 33.6 percent of reasoning tokens while preserving performance), and (2) the model's reflection behavior is highly correlated with an internal uncertainty signal, implying self-reflection may be controlled by the model's uncertainty.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Two-sided Acoustic Metascreen for Broadband and Individual Reflection and Transmission Control
Acoustic wave modulation plays a pivotal role in various applications, including sound-field reconstruction, wireless communication, and particle manipulation, among others. However, current acoustic metamaterial and met…
Generating Equivalent Representations of Code By A Self-Reflection Approach
Equivalent Representations (ERs) of code are textual representations that preserve the same semantics as the code itself, e.g., natural language comments and pseudocode. ERs play a critical role in software development a…
Code GenerationReflection Invariant and Symmetry Detection
Symmetry detection and discrimination are of fundamental meaning in science, technology, and engineering. This paper introduces reflection invariants and defines the directional moment to detect symmetry for shape analys…
Object RecognitionRetrievalSymmetry DetectionControlling Chat Style in Language Models via Single-Direction Editing
Controlling stylistic attributes in large language models (LLMs) remains challenging, with existing approaches relying on either prompt engineering or post-training alignment. This paper investigates this challenge throu…
Prompt EngineeringTarget-to-User Association in ISAC Systems With Vehicle-Lodged RIS
Target-to-user (T2U) association is a prerequisite to fully exploit the potential of the sensing function in communication-centric integrated sensing and communication (ISAC) systems, e.g., for beam and blockage manageme…
Integrated sensing and communicationISACManagement