paper-with-me

홈 › Papers

Textualized Agent-Style Reasoning for Complex Tasks by Multiple Round LLM Generation

2024-09-19 · Chen Liang, Zhifan Feng, Zihe Liu, Wenbin Jiang, Jinan Xu, Yufeng Chen, Yong Wang

Chain-of-thought prompting significantly boosts the reasoning ability of large language models but still faces three issues: hallucination problem, restricted interpretability, and uncontrollable generation. To address these challenges, we present AgentCOT, a llm-based autonomous agent framework, which can solve complex problems in an agent-style manner by multiple round LLM generation. At each step, AgentCOT selects an action and executes it to yield an intermediate result with supporting evidence. In addition, we integrate the step's index into the reasoning process to form a graph structure for complex inference logic. We introduce two new strategies to enhance the performance of AgentCOT.We conduct extensive experiments to verify the effectiveness of our method on six common benchmarks. Results exhibit that our method brings in substantial improvements over current competitive approaches.

📄 PDF Abstract BibTeX arXiv:2409.12411

Code (0)

등록된 구현이 없습니다.

Tasks

Hallucination

Similar Papers 제목 키워드 기반

MTBBench: A Multimodal Sequential Clinical Decision-Making Benchmark in Oncology

2025-11-25 · Kiril Vasilev, Alexandre Misrahi, Eeshaan Jain, Phil F Cheng 외 arxiv

Multimodal Large Language Models (LLMs) hold promise for biomedical reasoning, but current benchmarks fail to capture the complexity of real-world clinical workflows. Existing evaluations primarily assess unimodal, decon…

Harnessing Generalist Agents for Contextualized Time Series

2026-06-03 · Zihao Li, Kaifeng Jin, Yuanchen Bei, Jiaru Zou 외 arxiv

Time series are often embedded in rich contexts that are essential for holistic modeling. Moreover, real-world practitioners often require end-to-end workflows for analyzing temporal dynamics, where widely studied tasks …

PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization

2025-06-02 · Zouying Cao, Runze Wang, Yifei Yang, Xinbei Ma 외

Large Language Model (LLM) agents have demonstrated impressive capabilities in handling complex interactive problems. Existing LLM agents mainly generate natural language plans to guide reasoning, which is verbose and in…

Language ModelingLanguage ModellingLarge Language Model

Diagnose, Localize, Align: A Full-Stack Framework for Reliable LLM Multi-Agent Systems under Instruction Conflicts

2025-09-27 · Guancheng Wan, Leixin Sun, Longxu Dou, Zitong Shi 외 arxiv

Large Language Model (LLM)-powered multi-agent systems (MAS) have rapidly advanced collaborative reasoning, tool use, and role-specialized coordination in complex tasks. However, reliability-critical deployment remains h…

OrgAgent: Organize Your Multi-Agent System like a Company

2026-04-01 · Yiru Wang, Xinyue Shen, Yaohui Han, Michael Backes 외 arxiv

While large language model-based multi-agent systems have shown strong potential for complex reasoning, how to effectively organize multiple agents remains an open question. In this paper, we introduce OrgAgent, a compan…