papersSEP 12 04:00 UTC
arXiv paper proposes fragility spectrum for recursive language-model training
A new arXiv preprint examines what happens when text produced by language models is fed back into their own training data, a practice linked to shrinking output diversity. The authors propose a "fragility spectrum" framework to characterize how different training protocols and data mixtures degrade under this recursive loop. The work aims to give researchers a more systematic way to compare which setups hold up and which break down.