|
|
|
|
|
by blourvim
24 days ago
|
|
These models rely on knowledge that are embedded in their weights, if a new library is released, a new linux version comes out, some new protocol succeeds the previous one, you want your llm to know about it. Sure you can just add that into the context window, but that has its own problems. Unless new research, there are a few which look promising, gives a new method, training is going to be a constant cost sink. On top of this, if you stop training, it is 6 months until someone releases an open weights model and now you are competing to give the lowest price for the same product. Also we can't forget that this is a business that *has to* be in the global labor industry, not just a tech tool, they have to have much better models to justify the trillion dollar evaluation |
|