|
|
|
|
|
by Cthulhu_
18 days ago
|
|
I have a theory (armchair take here lmao) that AIs are trained on public code, but the biggest codebases are not public. Although I suspect models from Google, Facebook and Microsoft can be trained on their massive internal codebases. Whether they are is another question. |
|
But, you would probably see a difference of scale and architecture. Larger projects that need better organization are probably more likely to be in private codebases (Linux excluded). So you might be right about the lack of private code in LLM being an issue.