paper-with-me

홈 › Papers

Reasoning Topology Matters: Network-of-Thought for Complex Reasoning Tasks

2026-03-21 · Fan Huang arxiv

Existing prompting paradigms structure LLM reasoning in limited topologies: Chain-of-Thought (CoT) produces linear traces, while Tree-of-Thought (ToT) performs branching search. Yet complex reasoning often requires merging intermediate results, revisiting hypotheses, and integrating evidence from multiple sources. We propose Network-of-Thought (NoT), a framework that models reasoning as a directed graph with typed nodes and edges, guided by a heuristic-based controller policy. Across four benchmarks (GSM8K, Game of 24, HotpotQA, ProofWriter) and three models (GPT-4o-mini, Llama-3.3-70B-Instruct, Qwen2.5-72B-Instruct), we investigate when network topology outperforms chain or tree structures, whether LLM-generated heuristics can guide graph-based reasoning search, and the computation-accuracy tradeoff across topologies, evaluating each method on accuracy, topology simplicity, and token efficiency. Our results show that CoT remains effective for sequential tasks with GPT-4o-mini (89.5\% on GSM8K), while NoT surpasses ToT on multi-hop reasoning (91.0\% vs.\ 88.0\% on HotpotQA with LLM-as-Judge). With 72B open-source models, NoT achieves the highest accuracy on GSM8K (91.5\%), and Qwen2.5-72B achieves the best multi-hop QA result overall (91.7\% on HotpotQA). Self-generated controller heuristics outperform fixed and random strategies on logical reasoning, with uncertainty-only weighting achieving 57.0\% on ProofWriter. We also find that evaluation methodology significantly impacts method rankings: string-match underestimates all methods on open-ended QA, with the largest gap for NoT, a pattern consistent across all three models (14--18 percentage point gap on HotpotQA).

📄 PDF Abstract BibTeX arXiv:2603.20730

Code (0)

등록된 구현이 없습니다.

Tasks

Logical Reasoning

Results from the Paper

RankTaskDatasetModelMetrics
#19 GSM8K GSM8K Network-of-Thought Accuracy: 91.5

Similar Papers 제목 키워드 기반

Mechanistic Evidence for Faithfulness Decay in Chain-of-Thought Reasoning

2026-02-04 · Donald Ye, Max Loffgren, Om Kotadia, Linus Wong 외 arxiv

Chain-of-Thought (CoT) explanations are widely used to interpret how language models solve complex problems, yet it remains unclear whether these step-by-step explanations reflect how the model actually reaches its answe…

Exchange-of-Thought: Enhancing Large Language Model Capabilities through Cross-Model Communication

2023-12-04 · Zhangyue Yin, Qiushi Sun, Cheng Chang, Qipeng Guo 외

Large Language Models (LLMs) have recently made significant strides in complex reasoning tasks through the Chain-of-Thought technique. Despite this progress, their reasoning is often constrained by their intrinsic unders…

Language ModelingLanguage ModellingLarge Language Modelmodel

DeAR: Decentralized Agentic Reasoning via Capability Grounding and Collaborative Thought Navigation

2026-08-18 · Xing Wei, Changmeng Zheng, XiaoYong Wei, Xiufen Ye 외 arxiv

Existing agentic reasoning systems typically rely on centralized protocols. This design introduces routing bottlenecks and static role allocations that often fail when handling complex multimodal queries. We propose DeAR…

Multimodal Reasoning

Towards Understanding Chain-of-Thought Prompting: An Empirical Study of What Matters

2022-12-20 · Boshi Wang, Sewon Min, Xiang Deng, Jiaming Shen 외

Chain-of-Thought (CoT) prompting can dramatically improve the multi-step reasoning abilities of large language models (LLMs). CoT explicitly encourages the LLM to generate intermediate rationales for solving a problem, b…

Automatic Chain of Thought Prompting in Large Language Models

2022-10-07 · Zhuosheng Zhang, Aston Zhang, Mu Li, Alex Smola

Large language models (LLMs) can perform complex reasoning by generating intermediate reasoning steps. Providing these steps for prompting demonstrations is called chain-of-thought (CoT) prompting. CoT prompting has two …

DiversityQuestion Answering