paper-with-me

홈 › Papers

Consensus is Strategically Insufficient: Reasoning-Trace Disagreement as a Knowledge-Representation Signal

2026-06-02 · Michał Wawer, Jarosław A. Chudziak arxiv

Multi-agent systems are commonly designed to reduce disagreement through voting, consensus protocols, debate, or fault-tolerant aggregation. We argue that this objective is insufficient for value-laden tasks, where disagreement may reflect genuine normative uncertainty rather than agent error. Building on prior work on reasoning-trace disagreement in human-AI collaborative moderation, we propose a knowledge-representation layer in which reasoning traces and agent decisions are abstracted into symbolic disagreement states. Given agents producing explicit reasoning traces and binary decisions, we distinguish four states according to reasoning similarity and conclusion agreement: convergent agreement, divergent agreement, convergent disagreement and divergent disagreement. These states support defeasible strategic routing rules. We instantiate the framework in content moderation and argue that disagreement-aware routing provides a bridge between sub-symbolic LLM deliberation and symbolic knowledge representation for multi-agent strategic reasoning.

📄 PDF Abstract BibTeX arXiv:2606.04223

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Disagreement as Data: Reasoning Trace Analytics in Multi-Agent Systems

2026-01-18 · Elham Tajik, Conrad Borchers, Bahar Shahrokhian, Sebastian Simon 외 arxiv

Learning analytics researchers often analyze qualitative student data such as coded annotations or interview transcripts to understand learning processes. With the rise of generative AI, fully automated and human-AI work…

How Far Are LLMs from Professional Poker Players? Revisiting Game-Theoretic Reasoning with Agentic Tool Use

2026-01-31 · Minhua Lin, Enyan Dai, Hui Liu, Xianfeng Tang 외 arxiv

As Large Language Models (LLMs) are increasingly applied in high-stakes domains, their ability to reason strategically under uncertainty becomes critical. Poker provides a rigorous testbed, requiring not only strong acti…

Reinforcement Learning

Correct Prediction, Wrong Steps? Consensus Reasoning Knowledge Graph for Robust Chain-of-Thought Synthesis

2026-04-15 · Zipeng Ling, Shuliang Liu, Shenghong Fu, Yuehao Tang 외 arxiv

LLM reasoning traces suffer from complex flaws -- *Step Internal Flaws* (logical errors, hallucinations, etc.) and *Step-wise Flaws* (overthinking, underthinking), which vary by sample. A natural approach would be to pro…

Mathematical Reasoning

Reactivating Test-Time Scaling for Plane Geometry Problem Solving

2026-08-31 · Xiaoqiang Kang, Shengen Wu, Maizhen Ning, Xiaobo Jin 외 arxiv

Plane geometry problem (PGP) solving has become a critical benchmark for multimodal reasoning because it requires accurate visual perception and precise multi-step symbolic deduction. Although test-time scaling (TTS) has…

Mathematical ReasoningMultimodal ReasoningVisual Grounding

Minimal-time Deadbeat Consensus and Individual Disagreement Degree Prediction for High-order Linear Multi-agent Systems

2023-04-13 · Fu-Long Hu, Hai-Tao Zhang, Bowen Xu, Zhe Hu 외

In this paper, a Hankel matrix-based fully distributed algorithm is proposed to address a minimal-time deadbeat consensus prediction problem for discrete-time high-order multi-agent systems (MASs). Therein, each agent ca…

PredictionValue prediction