Hacker News new | ask | show | jobs
by SwellJoe 20 hours ago
Sounds like they're explaining magic. Because current models cannot learn and they cannot remember.
1 comments

there is RL for LLMs which actually changes the weights but its more specialization than learning and wont counter the probabalistic nature of the thing
Training is not something you can just bolt on, and it generally requires even larger hardware than inference for a given model, and a huge amount of time. If you're aiming for "free", a OP claims, RL aint it.