|
|
|
|
|
by RandyOrion
15 days ago
|
|
Thanks Bonsai team. Now open weight LLMs/VLMs/LMMs are becoming even larger to the extent that consumer-grade hardware are no longer able to run these models. In contrast, quantization and pruning make the model better at the size-performance pareto and provide people with strictly more possibilities. |
|