Large Language Models Suffer From Their Own Output: An Analysis of the Self-Consuming Training Loop

Large Language Models (LLM) are already widely used to generate content for a variety of online platforms. As we are not able to safely distinguish LLM-generated content from human-produced content, LLM-generated content is used to train the next generation of LLMs, giving rise to a self-consuming training loop. From the image generation domain we know that such a self-consuming training loop reduces both quality and diversity of images finally ending in a model collapse. However, it is unclear whether this alarming effect can also be observed for LLMs. Therefore, we present the first study investigating the self-consuming training loop for LLMs. Further, we propose a novel method based on logic expressions that allows us to unambiguously verify the correctness of LLM-generated content, which is difficult for natural language text. We find that the self-consuming training loop produces correct outputs, however, the output declines in its diversity depending on the proportion of the used generated data. Fresh data can slow down this decline, but not stop it. Given these concerning results, we encourage researchers to study methods to negate this process.

BERTScore: EvaluatingText Generation with…BERTScore: Evaluating Text Generation with BERTEvaluating LargeLanguage Models Trained…Evaluating Large Language Models Trained on CodeWill we run out of data?An analysis of the…Will we run out of data? An analysis of the limits of scaling datasets in Machine LearningThe Curse of Recursion:Training on Generated…The Curse of Recursion: Training on Generated Data Makes Models ForgetSelf-Instruct: AligningLanguage Models with…Self-Instruct: Aligning Language Models with Self-Generated InstructionsCan AI-Generated Text beReliably Detected?Can AI-Generated Text be Reliably Detected?Large Language Model asAttributed Training Dat…Large Language Model as Attributed Training Data Generator: A Tale of Diversity and BiasOn the Stability ofIterative Retraining of…On the Stability of Iterative Retraining of Generative Models on their own DataSelf-ConsumingGenerative Models Go MADSelf-Consuming Generative Models Go MADThe Curious Decline ofLinguistic Diversity…The Curious Decline of Linguistic Diversity: Training Language Models on Synthetic TextTowards Understandingthe Interplay of…Towards Understanding the Interplay of Generative Artificial Intelligence and the InternetMapping the IncreasingUse of LLMs in…Mapping the Increasing Use of LLMs in Scientific PapersIs Model CollapseInevitable? Breaking th…Is Model Collapse Inevitable? Breaking the Curse of Recursion by Accumulating Real and Synthetic DataTowards TheoreticalUnderstandings of…Towards Theoretical Understandings of Self-Consuming Generative ModelsA linguistic analysis ofundesirable outcomes in…A linguistic analysis of undesirable outcomes in the era of generative AIReDiFine: ReusableDiffusion Finetuning fo…ReDiFine: Reusable Diffusion Finetuning for Mitigating Degradation in the Chain of DiffusionA survey on the impactof AI-based recommender…A survey on the impact of AI-based recommenders on human behaviours: methodologies, outcomes and future directionsWhen AI Eats Itself: Onthe Caveats of Data…When AI Eats Itself: On the Caveats of Data Pollution in the Era of Generative AISelf-ConsumingGenerative Models with…Self-Consuming Generative Models with Curated Data Provably Optimize Human PreferencesUnderstandingHallucinations in…Understanding Hallucinations in Diffusion Models through Mode InterpolationHow to Synthesize TextData without Model…How to Synthesize Text Data without Model Collapse?Position: Model CollapseDoes Not Mean What You…Position: Model Collapse Does Not Mean What You ThinkA TheoreticalPerspective: How to…A Theoretical Perspective: How to Prevent Model Collapse in Self-consuming Training LoopsMind the Gap: Examiningthe Self-Improvement…Mind the Gap: Examining the Self-Improvement Capabilities of Large Language ModelsLarge Language ModelsSuffer From Their Own…Large Language Models Suffer From Their Own Output: An Analysis of the Self-Consuming Training Loop過去の参考文献中心の論文この論文を引用する論文古い新しい

ノードをクリックするとフォーカスを固定、空白をクリックすると本論文に戻ります。ホバーで一時的にプレビューできます。各ノードのページはタイトルから開けます。