paper-with-me

홈 › Papers

Draw2Think: Harnessing Geometry Reasoning through Constraint Engine Interaction

2026-05-20 · Juncheng Hu, Jiawei Du, Xin Zhang, Joey Tianyi Zhou arxiv

Vision-language models solve geometry problems with rising accuracy, yet their intermediate states remain latent and unverifiable: a relation expressed in textual reasoning or drawing code carries no guarantee that a constraint-satisfying configuration realizes it. We observe that existing externalization methods based on rendered pixels or one-shot scripts fail to provide exact, per-action geometric guarantees. Enforcing geometric relations by algebraic definition closes this gap: the workspace becomes a constraint-checked evolving canvas. We present Draw2Think, a framework that recasts geometric reasoning from latent spatial inference into agentic interaction with the GeoGebra constraint engine. In a Propose-Draw-Verify loop, Draw2Think externalizes hypotheses onto an executable canvas, measures exact geometric quantities, and feeds structured observations back to the model, so subsequent reasoning proceeds from checked canvas state grounded by the shared workspace. This externalization makes two properties separately auditable: model-level Construction Fidelity (whether the canvas realizes the intended configuration) and engine-level Measurement Faithfulness (exact values and relations from canvas constraints). Across construction, outcome, and rendering evaluations, Draw2Think builds canvases that pass 95.9% predicate-level and 84.0% strict problem-level construction checks on GeoGoal, improves outcome accuracy by up to 4.1%/16.4% on planar/solid benchmarks, and attains 68.2%/90.5% strict/relaxed rendering scores on GenExam-math. Project page is available at https://draw2think.github.io/

📄 PDF Abstract BibTeX arXiv:2605.20743

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Thinking with Geometry: Active Geometry Integration for Spatial Reasoning

2026-02-05 · Haoyuan Li, Qihang Cao, Tao Tang, Kun Xiang 외 arxiv

Recent progress in spatial reasoning with Multimodal Large Language Models (MLLMs) increasingly leverages geometric priors from 3D encoders. However, most existing integration strategies remain passive: geometry is expos…

Autonomous DrivingSpatial Reasoning

VisuoThink: Empowering LVLM Reasoning with Multimodal Tree Search

2025-04-12 · Yikun Wang, Siyin Wang, Qinyuan Cheng, Zhaoye Fei 외

Recent advancements in Large Vision-Language Models have showcased remarkable capabilities. However, they often falter when confronted with complex reasoning tasks that humans typically address through visual aids and de…

Spatial Reasoning

Could Thinking Multilingually Empower LLM Reasoning?

2025-04-16 · Changjiang Gao, Xu Huang, Wenhao Zhu, ShuJian Huang 외

Previous work indicates that large language models exhibit a significant "English bias", i.e. they often perform better when tasks are presented in English. Interestingly, we have observed that using certain other langua…

Answer Selection

Trade-offs in Large Reasoning Models: An Empirical Analysis of Deliberative and Adaptive Reasoning over Foundational Capabilities

2025-03-23 · Weixiang Zhao, Xingyu Sui, Jiahe Guo, Yulin Hu 외

Recent advancements in Large Reasoning Models (LRMs), such as OpenAI's o1/o3 and DeepSeek-R1, have demonstrated remarkable performance in specialized reasoning tasks through human-like deliberative thinking and long chai…

ThinkOmni: Lifting Textual Reasoning to Omni-modal Scenarios via Guidance Decoding

2026-02-26 · Yiran Guan, Sifan Tu, Dingkang Liang, Linghao Zhu 외 arxiv

Omni-modal reasoning is essential for intelligent systems to understand and draw inferences from diverse data sources. While existing omni-modal large language models (OLLM) excel at perceiving diverse modalities, they l…