paper-with-me

홈 › Papers

DebFlow: Automating Agent Creation via Agent Debate

2025-03-31 · Jinwei Su, Yinghui Xia, Ronghua Shi, Jianhui Wang, Jianuo Huang, Yijin Wang, Tianyu Shi, Yang Jingsong, Lewei He

Large language models (LLMs) have demonstrated strong potential and impressive performance in automating the generation and optimization of workflows. However, existing approaches are marked by limited reasoning capabilities, high computational demands, and significant resource requirements. To address these issues, we propose DebFlow, a framework that employs a debate mechanism to optimize workflows and integrates reflexion to improve based on previous experiences. We evaluated our method across six benchmark datasets, including HotpotQA, MATH, and ALFWorld. Our approach achieved a 3\% average performance improvement over the latest baselines, demonstrating its effectiveness in diverse problem domains. In particular, during training, our framework reduces resource consumption by 37\% compared to the state-of-the-art baselines. Additionally, we performed ablation studies. Removing the Debate component resulted in a 4\% performance drop across two benchmark datasets, significantly greater than the 2\% drop observed when the Reflection component was removed. These findings strongly demonstrate the critical role of Debate in enhancing framework performance, while also highlighting the auxiliary contribution of reflexion to overall optimization.

📄 PDF Abstract BibTeX arXiv:2503.23781

Code (0)

등록된 구현이 없습니다.

Tasks

Math

Similar Papers 제목 키워드 기반

Towards Debate Automation: a Recurrent Model for Predicting Debate Winners

2017-09-01 · EMNLP 2017 9 · Peter Potash, Anna Rumshisky

In this paper we introduce a practical first step towards the creation of an automated debate agent: a state-of-the-art recurrent predictive model for predicting debate winners. By having an accurate predictive model, we…

Text Generation

Multi-Agent Debate for LLM Judges with Adaptive Stability Detection

2025-10-14 · Tianyu Hu, Zhen Tan, Song Wang, Huaizhi Qu 외 arxiv

With advancements in reasoning capabilities, Large Language Models (LLMs) are increasingly employed for automated judgment tasks. While LLMs-as-Judges offer promise in automating evaluations, current approaches often rel…

Computational Efficiency

INDIBATOR: Diverse and Fact-Grounded Individuality for Multi-Agent Debate in Molecular Discovery

2026-02-02 · Yunhui Jang, Seonghyun Park, Jaehyung Kim, Sungsoo Ahn arxiv

Multi-agent systems have emerged as a powerful paradigm for automating scientific discovery. To differentiate agent behavior in the multi-agent system, current frameworks typically assign generic role-based personas such…

An Agentic Approach to Automatic Creation of P&ID Diagrams from Natural Language Descriptions

2024-12-17 · Shreeyash Gowaikar, Srinivasan Iyengar, Sameer Segal, Shivkumar Kalyanaraman

The Piping and Instrumentation Diagrams (P&IDs) are foundational to the design, construction, and operation of workflows in the engineering and process industries. However, their manual creation is often labor-intensive,…

Smart Proofs via Smart Contracts: Succinct and Informative Mathematical Derivations via Decentralized Markets

2021-02-05 · Sylvain Carré, Franck Gabriel, Clément Hongler, Gustavo Lacerda 외

Modern mathematics is built on the idea that proofs should be translatable into formal proofs, whose validity is an objective question, decidable by a computer. Yet, in practice, proofs are informal and may omit many det…

valid