|
|
|
|
|
by FromTheFirstIn
32 days ago
|
|
And sitting right next to the data and compute factors in every cross entropy loss equation is the entropy of the language, which is just a fixed constant. There’s such a hard cap on cross entropy loss training and I never hear it come up! |
|