paper-with-me

Papers

The Illusion of Insight in Reasoning Models

2026-01-02 · Liv G. d'Aliberti, Manoel Horta Ribeiro arxiv

Do reasoning models have "Aha!" moments? Prior work suggests that models like DeepSeek-R1-Zero undergo sudden mid-trace realizations that lead to accurate outputs, implying an intrinsic capacity for self-correction. Yet, it remains unclear whether such intrinsic shifts in reasoning strategy actually improve performance. Here, we study mid-reasoning shifts and instrument training runs to detect them. Our analysis spans 1M+ reasoning traces, hundreds of training checkpoints, three reasoning domains, and multiple decoding temperatures and model architectures. We find that reasoning shifts are rare, do not become more frequent with training, and seldom improve accuracy, indicating that they do not correspond to prior perceptions of model insight. However, their effect varies with model uncertainty. Building on this finding, we show that artificially triggering extrinsic shifts under high entropy reliably improves accuracy. Our results show that mid-reasoning shifts are symptoms of unstable inference behavior rather than an intrinsic mechanism for self-correction.

📄 PDF Abstract BibTeX arXiv:2601.00514

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Can LVLMs Uncover the Truth Behind Visual Illusions? An Analysis of Perceptual and Reasoning Capabilities

2026-07-30 · Liangjie Zhao, Jiaqing Lyu, Kexin Tang, Zecheng Fang 외 arxiv

Large Vision Language Models have integrated reasoning capabilities, elevating cognitive performance to new levels. However, existing evaluations either focus solely on perception or rely on specific domains such as math…

Seeing Is Believing? A Benchmark for Multimodal Large Language Models on Visual Illusions and Anomalies

2026-02-02 · Wenjin Hou, Wei Liu, Han Hu, Xiaoxiao Sun 외 arxiv

Multimodal Large Language Models (MLLMs) have shown remarkable proficiency on general-purpose vision-language benchmarks, reaching or even exceeding human-level performance. However, these evaluations typically rely on s…

Visual Reasoning

A Bioplausible Model for the Expanding Hole Illusion: Insights into Retinal Processing and Illusory Motion

2025-01-15 · Nasim Nematzadeh, David M. W. Powers

The Expanding Hole Illusion is a compelling visual phenomenon in which a static, concentric pattern evokes a strong perception of continuous forward motion. Despite its simplicity, this illusion challenges our understand…

Pupil Dilation

InDL: A New Dataset and Benchmark for In-Diagram Logic Interpretation based on Visual Illusion

2023-05-28 · Haobo Yang, Wenyu Wang, Ze Cao, Zhekai Duan 외

This paper introduces a novel approach to evaluating deep learning models' capacity for in-diagram logic interpretation. Leveraging the intriguing realm of visual illusions, we establish a unique dataset, InDL, designed …

BenchmarkingDecision MakingImage ClassificationLogical Reasoning

From Illusion to Intention: Visual Rationale Learning for Vision-Language Reasoning

2025-11-28 · Changpeng Wang, Haozhe Wang, Xi Chen, Junhan Liu 외 arxiv

Recent advances in vision-language reasoning underscore the importance of thinking with images, where models actively ground their reasoning in visual evidence. Yet, prevailing frameworks treat visual actions as optional…