paper-with-me

홈 › Papers

COINBench: Moving Beyond Individual Perspectives to Collective Intent Understanding

2026-03-22 · Xiaozhe Li, Tianyi Lyu, Siyi Yang, Yizhao Yang, Yuxi Gong, Jinxuan Huang, Ligao Zhang, Zhuoyi Huang, Qingwen Liu arxiv

Understanding human intent is a high-level cognitive challenge for Large Language Models (LLMs), requiring sophisticated reasoning over noisy, conflicting, and non-linear discourse. While LLMs excel at following individual instructions, their ability to distill Collective Intent - the process of extracting consensus, resolving contradictions, and inferring latent trends from multi-source public discussions - remains largely unexplored. To bridge this gap, we introduce COIN-BENCH, a dynamic, real-world, live-updating benchmark specifically designed to evaluate LLMs on collective intent understanding within the consumer domain. Unlike traditional benchmarks that focus on transactional outcomes, COIN-BENCH operationalizes intent as a hierarchical cognitive structure, ranging from explicit scenarios to deep causal reasoning. We implement a robust evaluation pipeline that combines a rule-based method with an LLM-as-the-Judge approach. This framework incorporates COIN-TREE for hierarchical cognitive structuring and retrieval-augmented verification (COIN-RAG) to ensure expert-level precision in analyzing raw, collective human discussions. An extensive evaluation of 20 state-of-the-art LLMs across four dimensions - depth, breadth, informativeness, and correctness - reveals that while current models can handle surface-level aggregation, they still struggle with the analytical depth required for complex intent synthesis. COIN-BENCH establishes a new standard for advancing LLMs from passive instruction followers to expert-level analytical agents capable of deciphering the collective voice of the real world. See our project page on COIN-BENCH.

📄 PDF Abstract BibTeX arXiv:2603.21329

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Artificial Collective Intelligence Engineering: a Survey of Concepts and Perspectives

2023-04-11 · Roberto Casadei

Collectiveness is an important property of many systems--both natural and artificial. By exploiting a large number of individuals, it is often possible to produce effects that go far beyond the capabilities of the smarte…

Survey

MirrorMind: Empowering OmniScientist with the Expert Perspectives and Collective Knowledge of Human Scientists

2025-11-21 · Qingbin Zeng, Bingbing Fan, Zhiyu Chen, Sijian Ren 외 arxiv

The emergence of AI Scientists has demonstrated remarkable potential in automating scientific research. However, current approaches largely conceptualize scientific discovery as a solitary optimization or search process,…

Shared Nature, Unique Nurture: PRISM for Pluralistic Reasoning via In-context Structure Modeling

2026-02-24 · Guancheng Tu, Shiyang Zhang, Tianyu Zhang, Yi Zhang 외 arxiv

Large Language Models (LLMs) are converging towards a singular Artificial Hivemind, where shared Nature (pre-training priors) result in a profound collapse of distributional diversity, limiting the distinct perspectives …

Beyond Individuals: Collective Predictive Coding for Memory, Attention, and the Emergence of Language

2025-08-20 · Tadahiro Taniguchi arxiv

This commentary extends the discussion by Parr et al. on memory and attention beyond individual cognitive systems. From the perspective of the Collective Predictive Coding (CPC) hypothesis -- a framework for understandin…

Quantum cognition goes beyond-quantum: modeling the collective participant in psychological measurements

2018-02-24 · Diederik Aerts, Massimiliano Sassoli de Bianchi, Sandro Sozzo, Tomas Veloz

In psychological measurements, two levels should be distinguished: the 'individual level', relative to the different participants in a given cognitive situation, and the 'collective level', relative to the overall statis…

valid