Hacker News new | ask | show | jobs
by Razengan 8 days ago
Wait.. If "distillation" produces better results than the "original" models, why don't the US model owners do it to themselves?

Have ChatGPT 5.6 talk to itself to produce 5.7 or whatever?

And since they already know which requests came from China or looked like distillation, can't they just replay those same prompts?

1 comments

They already do and everyone has been using that concept for years now, it's known as RLAIF.