paper-with-me

홈 › Papers

Large Language Models Suffer From Their Own Output: An Analysis of the Self-Consuming Training Loop

2023-11-28 · Martin Briesch, Dominik Sobania, Franz Rothlauf

Large Language Models (LLM) are already widely used to generate content for a variety of online platforms. As we are not able to safely distinguish LLM-generated content from human-produced content, LLM-generated content is used to train the next generation of LLMs, giving rise to a self-consuming training loop. From the image generation domain we know that such a self-consuming training loop reduces both quality and diversity of images finally ending in a model collapse. However, it is unclear whether this alarming effect can also be observed for LLMs. Therefore, we present the first study investigating the self-consuming training loop for LLMs. Further, we propose a novel method based on logic expressions that allows us to unambiguously verify the correctness of LLM-generated content, which is difficult for natural language text. We find that the self-consuming training loop produces correct outputs, however, the output declines in its diversity depending on the proportion of the used generated data. Fresh data can slow down this decline, but not stop it. Given these concerning results, we encourage researchers to study methods to negate this process.

📄 PDF Abstract BibTeX arXiv:2311.16822

Code (0)

등록된 구현이 없습니다.

Tasks

DiversityImage Generation

Similar Papers 제목 키워드 기반

ForgerySleuth: Empowering Multimodal Large Language Models for Image Manipulation Detection

2024-11-29 · Zhihao Sun, Haoran Jiang, Haoran Chen, Yixin Cao 외

Multimodal large language models have unlocked new possibilities for various multimodal tasks. However, their potential in image manipulation detection remains unexplored. When directly applied to the IMD task, M-LLMs of…

Image ManipulationImage Manipulation Detection

Enriching Taxonomies Using Large Language Models

2025-11-21 · Zeinab Ghamlouch, Mehwish Alam arxiv

Taxonomies play a vital role in structuring and categorizing information across domains. However, many existing taxonomies suffer from limited coverage and outdated or ambiguous nodes, reducing their effectiveness in kno…

Dynamic Attention-Guided Context Decoding for Mitigating Context Faithfulness Hallucinations in Large Language Models

2025-01-02 · Yanwen Huang, Yong Zhang, Ning Cheng, Zhitao Li 외

Large language models (LLMs) often suffer from context faithfulness hallucinations, where outputs deviate from retrieved information due to insufficient context utilization and high output uncertainty. Our uncertainty ev…

Computational Efficiency

VALOR-EVAL: Holistic Coverage and Faithfulness Evaluation of Large Vision-Language Models

2024-04-22 · Haoyi Qiu, WenBo Hu, Zi-Yi Dou, Nanyun Peng

Large Vision-Language Models (LVLMs) suffer from hallucination issues, wherein the models generate plausible-sounding but factually incorrect outputs, undermining their reliability. A comprehensive quantitative evaluatio…

HallucinationInformativenessLanguage ModelingLanguage Modelling+1

Mitigating Hallucinations in Large Language Models Via Decoder Layer Skipping

2026-05-30 · Hanze Li, Jinhao You, Yichen Guo, Kai Tang 외 arxiv

Large Language Models (LLMs) have achieved strong performance across diverse natural language tasks, yet their outputs often suffer from hallucinations -- content that is misaligned with factual information. In this work…