paper-with-me

홈 › Papers

FactGuard: Leveraging Multi-Agent Systems to Generate Answerable and Unanswerable Questions for Enhanced Long-Context LLM Extraction

2025-04-08 · Qian-Wen Zhang, Fang Li, Jie Wang, Lingfeng Qiao, Yifei Yu, Di Yin, Xing Sun

Extractive reading comprehension systems are designed to locate the correct answer to a question within a given text. However, a persistent challenge lies in ensuring these models maintain high accuracy in answering questions while reliably recognizing unanswerable queries. Despite significant advances in large language models (LLMs) for reading comprehension, this issue remains critical, particularly as the length of supported contexts continues to expand. To address this challenge, we propose an innovative data augmentation methodology grounded in a multi-agent collaborative framework. Unlike traditional methods, such as the costly human annotation process required for datasets like SQuAD 2.0, our method autonomously generates evidence-based question-answer pairs and systematically constructs unanswerable questions. Using this methodology, we developed the FactGuard-Bench dataset, which comprises 25,220 examples of both answerable and unanswerable question scenarios, with context lengths ranging from 8K to 128K. Experimental evaluations conducted on seven popular LLMs reveal that even the most advanced models achieve only 61.79% overall accuracy. Furthermore, we emphasize the importance of a model's ability to reason about unanswerable questions to avoid generating plausible but incorrect answers. By implementing efficient data selection and generation within the multi-agent collaborative framework, our method significantly reduces the traditionally high costs associated with manual annotation and provides valuable insights for the training and optimization of LLMs.

📄 PDF Abstract BibTeX arXiv:2504.05607

Code (1)

factguard/factguardbench 공식 구현

Tasks

8kData AugmentationReading Comprehension

Similar Papers 제목 키워드 기반

FactGuard: Agentic Video Misinformation Detection via Reinforcement Learning

2026-02-26 · Zehao Li, Hongwei Yu, Hao Jiang, Qiang Sheng 외 arxiv

Multimodal large language models (MLLMs) have substantially advanced video misinformation detection through unified multimodal reasoning, but they often rely on fixed-depth inference and place excessive trust in internal…

Reinforcement LearningMultimodal ReasoningDecision Making

FactGuard: Event-Centric and Commonsense-Guided Fake News Detection

2025-11-13 · Jing He, Han Zhang, Yuanhui Xiao, Wei Guo 외 arxiv

Fake news detection methods based on writing style have achieved remarkable progress. However, as adversaries increasingly imitate the style of authentic news, the effectiveness of such approaches is gradually diminishin…

Knowledge DistillationFake News Detection

RoboFactory: Exploring Embodied Agent Collaboration with Compositional Constraints

2025-03-20 · Yiran Qin, Li Kang, Xiufeng Song, Zhenfei Yin 외

Designing effective embodied multi-agent systems is critical for solving complex real-world tasks across domains. Due to the complexity of multi-agent embodied systems, existing methods fail to automatically generate saf…

Imitation Learning

HECATE: An ECS-based Framework for Teaching and Developing Multi-Agent Systems

2025-09-08 · Arthur Casals, Anarosa A. F. Brandão arxiv

This paper introduces HECATE, a novel framework based on the Entity-Component-System (ECS) architectural pattern that bridges the gap between distributed systems engineering and MAS development. HECATE is built using the…

MAS$^2$: Self-Generative, Self-Configuring, Self-Rectifying Multi-Agent Systems

2025-09-29 · Kun Wang, Guibin Zhang, ManKit Ye, Xinyu Deng 외 arxiv

The past two years have witnessed the meteoric rise of Large Language Model (LLM)-powered multi-agent systems (MAS), which harness collective intelligence and exhibit a remarkable trajectory toward self-evolution. This p…

Code Generation