Hacker News new | ask | show | jobs
by satvikpendem 9 days ago
The paradox is analogous to AI poisoning its own training data as more and more AI generated content is released on the Internet. Indeed, I see a hard terminus for both man and machine at some point.
1 comments

Absolutely right. First gen models are trained on 'virgin' non-LLM output. Subsequent models are tainted by ingesting AI replies. Rinse and repeat, and all you end up with is AI copies of AI replies, and the noise will completely overtake the signal.
I'd have been more worried about that had the models not gotten incredibly good over time.

But they did get good and this seems like a non-issue.