Hacker News new | ask | show | jobs
by RandyOrion 15 days ago
Thanks Bonsai team.

Now open weight LLMs/VLMs/LMMs are becoming even larger to the extent that consumer-grade hardware are no longer able to run these models. In contrast, quantization and pruning make the model better at the size-performance pareto and provide people with strictly more possibilities.