|
|
|
|
|
by libraryofbabel
178 days ago
|
|
I don't know - perhaps someone who's more of an expert or who's worked a lot with open source models that haven't been RL-ed can weigh in here! But certainly without the RL step, the LLM would be much worse at coding and would hallucinate more. |
|