paper-with-me

홈 › Papers

Recursive Chain-of-Feedback Prevents Performance Degradation from Redundant Prompting

2024-02-05 · Jinwoo Ahn, Kyuseung Shin

Large Language Models (LLMs) frequently struggle with complex reasoning tasks, failing to construct logically sound steps towards the solution. In response to this behavior, users often try prompting the LLMs repeatedly in hopes of reaching a better response. This paper studies such repetitive behavior and its effect by defining a novel setting, Chain-of-Feedback (CoF). The setting takes questions that require multi-step reasoning as an input. Upon response, we repetitively prompt meaningless feedback (e.g. 'make another attempt') requesting additional trials. Surprisingly, our preliminary results show that repeated meaningless feedback gradually decreases the quality of the responses, eventually leading to a larger deviation from the intended outcome. To alleviate these troubles, we propose a novel method, Recursive Chain-of-Feedback (R-CoF). Following the logic of recursion in computer science, R-CoF recursively revises the initially incorrect response by breaking down each incorrect reasoning step into smaller individual problems. Our preliminary results show that majority of questions that LLMs fail to respond correctly can be answered using R-CoF without any sample data outlining the logical process.

📄 PDF Abstract BibTeX arXiv:2402.02648

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Silent Collapse in Recursive Learning Systems

2026-05-14 · Zhipeng Zhang arxiv

Recursive learning -- where models are trained on data generated by previous versions of themselves -- is increasingly common in large language models, autonomous agents, and self-supervised systems. However, standard pe…

The Anti-Ouroboros Effect: Emergent Resilience in Large Language Models from Recursive Selective Feedback

2025-09-02 · Sai Teja Reddy Adapala arxiv

The stability of recursively trained large language models (LLMs) is a foundational problem for AI safety. Prevailing theory predicts model collapse, a progressive degradation when models are trained on their own output.…

Learning in Feedback-driven Recurrent Spiking Neural Networks using full-FORCE Training

2022-05-26 · Ankita Paul, Stefan Wagner, Anup Das

Feedback-driven recurrent spiking neural networks (RSNNs) are powerful computational models that can mimic dynamical systems. However, the presence of a feedback loop from the readout to the recurrent layer de-stabilizes…

Can Generative Artificial Intelligence Survive Data Contamination? Theoretical Guarantees under Contaminated Recursive Training

2026-02-17 · Kevin Wang, Hongqian Niu, Didong Li arxiv

As artificial intelligence (AI)-generated content proliferates, models are increasingly trained on their own outputs, risking progressive degradation or collapse. In this article, we provide the first positive, rigorous …

An Agent-based Model of the Cognitive Mechanisms Underlying the Origins of Creative Cultural Evolution

2013-10-14 · Liane Gabora, Maryam Saberi

Human culture is uniquely cumulative and open-ended. Using a computational model of cultural evolution in which neural network based agents evolve ideas for actions through invention and imitation, we tested the hypothes…

Cultural Vocal Bursts Intensity PredictionDiversity